English
Related papers

Related papers: VolE: A Point-cloud Framework for Food 3D Reconstr…

200 papers

We propose an approach for 3D reconstruction and segmentation of a single object placed on a flat surface from an input video. Our approach is to perform dense depth map estimation for multiple views using a proposed objective function that…

Computer Vision and Pattern Recognition · Computer Science 2016-07-29 Tanmay Gupta , Daeyun Shin , Naren Sivagnanadasan , Derek Hoiem

Volume data is found in many important scientific and engineering applications. Rendering this data for visualization at high quality and interactive rates for demanding applications such as virtual reality is still not easily achievable…

Graphics · Computer Science 2022-09-22 David Bauer , Qi Wu , Kwan-Liu Ma

Recent hand-object interaction datasets show limited real object variability and rely on fitting the MANO parametric model to obtain groundtruth hand shapes. To go beyond these limitations and spur further research, we introduce the SHOWMe…

Computer Vision and Pattern Recognition · Computer Science 2023-09-20 Anilkumar Swamy , Vincent Leroy , Philippe Weinzaepfel , Fabien Baradel , Salma Galaaoui , Romain Bregier , Matthieu Armando , Jean-Sebastien Franco , Gregory Rogez

Three-dimensional ultrasound enables real-time volumetric visualization of anatomical structures. Unlike traditional 2D ultrasound, 3D imaging reduces reliance on precise probe orientation, potentially making ultrasound more accessible to…

Image and Video Processing · Electrical Eng. & Systems 2026-05-05 Tristan S. W. Stevens , Oisín Nolan , Oudom Somphone , Jean-Luc Robert , Ruud J. G. van Sloun

Recent generative models have shown strong performance in generating diverse 3D assets from 2D images, a fundamental research topic in computer vision and graphics. However, these models still struggle to generate voluminous 3D assets when…

Computer Vision and Pattern Recognition · Computer Science 2026-05-01 Hankyeol Lee , Wooyeol Baek , Seongdo Kim , Jongyoo Kim

Cross-modality medical image synthesis is a critical topic and has the potential to facilitate numerous applications in the medical imaging field. Despite recent successes in deep-learning-based generative models, most current medical image…

Image and Video Processing · Electrical Eng. & Systems 2023-07-20 Lingting Zhu , Zeyue Xue , Zhenchao Jin , Xian Liu , Jingzhen He , Ziwei Liu , Lequan Yu

Estimating precise metric depth and scene reconstruction from monocular endoscopy is a fundamental task for surgical navigation in robotic surgery. However, traditional stereo matching adopts binocular images to perceive the depth…

Robotics · Computer Science 2022-11-29 Ruofeng Wei , Bin Li , Hangjie Mo , Fangxun Zhong , Yonghao Long , Qi Dou , Yun-Hui Liu , Dong Sun

Purpose: To develop a framework to reconstruct large-scale volumetric dynamic MRI from rapid continuous and non-gated acquisitions, with applications to pulmonary and dynamic contrast enhanced (DCE) imaging. Theory and Methods: The problem…

Medical imaging is critical for diagnostics, but clinical adoption of advanced AI-driven imaging faces challenges due to patient variability, image artifacts, and limited model generalization. While deep learning has transformed image…

Image and Video Processing · Electrical Eng. & Systems 2025-06-02 Abdul-mojeed Olabisi Ilyas , Adeleke Maradesa , Jamal Banzi , Jianpan Huang , Henry K. F. Mak , Kannie W. Y. Chan

Purpose: To develop a MRI acquisition and reconstruction framework for volumetric cine visualisation of the fetal heart and great vessels in the presence of maternal and fetal motion. Methods: Four-dimensional depiction was achieved using a…

We present an approach for recognizing all objects in a scene and estimating their full pose from an accurate 3D instance-aware semantic reconstruction using an RGB-D camera. Our framework couples convolutional neural networks (CNNs) and a…

Robotics · Computer Science 2019-10-01 Dinh-Cuong Hoang , Todor Stoyanov , Achim J. Lilienthal

Generalizable neural surface reconstruction has become a compelling technique to reconstruct from few images without per-scene optimization, where dense 3D feature volume has proven effective as a global representation of scenes. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-09 Aoxiang Fan , Corentin Dumery , Nicolas Talabot , Hieu Le , Pascal Fua

While a traditional camera only captures one point of view of a scene, a plenoptic or light-field camera, is able to capture spatial and angular information in a single snapshot, enabling depth estimation from a single acquisition. In this…

Image and Video Processing · Electrical Eng. & Systems 2023-08-09 Mathieu Labussière , Céline Teulière , Omar Ait-Aider

The conventional pose estimation of a 3D object usually requires the knowledge of the 3D model of the object. Even with the recent development in convolutional neural networks (CNNs), a 3D model is often necessary in the final estimation.…

Robotics · Computer Science 2019-01-01 Zhongang Cai , Cunjun Yu , Quang-Cuong Pham

Despite the recent development of learning-based gaze estimation methods, most methods require one or more eye or face region crops as inputs and produce a gaze direction vector as output. Cropping results in a higher resolution in the eye…

Computer Vision and Pattern Recognition · Computer Science 2023-05-10 Haldun Balim , Seonwook Park , Xi Wang , Xucong Zhang , Otmar Hilliges

Recent advances in Latent Video Diffusion Models (LVDMs) have revolutionized video generation by leveraging Video Variational Autoencoders (Video VAEs) to compress intricate video data into a compact latent space. However, as LVDM training…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Yu Cheng , Fajie Yuan

This study presents Flower Pose Estimation (FloPE), a real-time flower pose estimation framework for computationally constrained robotic pollination systems. Robotic pollination has been proposed to supplement natural pollination to ensure…

Robotics · Computer Science 2026-02-09 Rashik Shrestha , Madhav Rijal , Trevor Smith , Yu Gu

Automatic dietary assessment based on food images remains a challenge, requiring precise food detection, segmentation, and classification. Vision-Language Models (VLMs) offer new possibilities by integrating visual and textual reasoning. In…

Diffusion models have become a leading approach for high-fidelity medical image synthesis. However, most existing methods for 3D medical image generation rely on convolutional U-Net backbones within latent diffusion frameworks. While…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 Marvin Seyfarth , Salman Ul Hassan Dar , Yannik Frisch , Philipp Wild , Norbert Frey , Florian André , Sandy Engelhardt

Single-image human mesh recovery is a challenging task due to the ill-posed nature of simultaneous body shape, pose, and camera estimation. Existing estimators work well on images taken from afar, but they break down as the person moves…

Computer Vision and Pattern Recognition · Computer Science 2024-12-12 Shengze Wang , Jiefeng Li , Tianye Li , Ye Yuan , Henry Fuchs , Koki Nagano , Shalini De Mello , Michael Stengel