English
Related papers

Related papers: VolTex: Food Volume Estimation using Text-Guided S…

200 papers

Two-dimensional (2D) freehand ultrasonography is one of the most commonly used medical imaging modalities, particularly in obstetrics and gynaecology. However, it only captures 2D cross-sectional views of inherently 3D anatomies, losing…

Image and Video Processing · Electrical Eng. & Systems 2024-04-17 Mark C. Eid , Pak-Hei Yeung , Madeleine K. Wyburd , João F. Henriques , Ana I. L. Namburete

Data-driven methods have shown great potential in solving challenging manipulation tasks; however, their application in the domain of deformable objects has been constrained, in part, by the lack of data. To address this lack, we propose…

Optimal surface segmentation is a state-of-the-art method used for segmentation of multiple globally optimal surfaces in volumetric datasets. The method is widely used in numerous medical image segmentation applications. However, nodes in…

Computer Vision and Pattern Recognition · Computer Science 2019-02-18 Abhay Shah , Michael D. Abramoff , Xiaodong Wu

The advancement of diffusion models has pushed the boundary of text-to-3D object generation. While it is straightforward to composite objects into a scene with reasonable geometry, it is nontrivial to texture such a scene perfectly due to…

Computer Vision and Pattern Recognition · Computer Science 2024-06-05 Qi Wang , Ruijie Lu , Xudong Xu , Jingbo Wang , Michael Yu Wang , Bo Dai , Gang Zeng , Dan Xu

In volume-to-volume translations in medical images, existing models often struggle to capture the inherent volumetric distribution using 3D voxelspace representations, due to high computational dataset demands. We present Score-Fusion, a…

Computer Vision and Pattern Recognition · Computer Science 2025-02-10 Xiyue Zhu , Dou Hoon Kwark , Ruike Zhu , Kaiwen Hong , Yiqi Tao , Shirui Luo , Yudu Li , Zhi-Pei Liang , Volodymyr Kindratenko

The success of the Neural Radiance Fields (NeRF) in novel view synthesis has inspired researchers to propose neural implicit scene reconstruction. However, most existing neural implicit reconstruction methods optimize per-scene parameters…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Yufan Ren , Fangjinhua Wang , Tong Zhang , Marc Pollefeys , Sabine Süsstrunk

In this paper, we tackle the problem of active robotic 3D reconstruction of an object. In particular, we study how a mobile robot with an arm-held camera can select a favorable number of views to recover an object's 3D shape efficiently.…

Computer Vision and Pattern Recognition · Computer Science 2022-09-20 Soomin Lee , Le Chen , Jiahao Wang , Alexander Liniger , Suryansh Kumar , Fisher Yu

Regular nutrient intake monitoring in hospitalised patients plays a critical role in reducing the risk of disease-related malnutrition (DRM). Although several methods to estimate nutrient intake have been developed, there is still a clear…

Computer Vision and Pattern Recognition · Computer Science 2019-06-13 Ya Lu , Thomai Stathopoulou , Maria F. Vasiloglou , Stergios Christodoulidis , Beat Blum , Thomas Walser , Vinzenz Meier , Zeno Stanga , Stavroula G. Mougiakakou

In recent years, sparse voxel-based methods have become the state-of-the-arts for 3D semantic segmentation of indoor scenes, thanks to the powerful 3D CNNs. Nevertheless, being oblivious to the underlying geometry, voxel-based methods…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Zeyu Hu , Xuyang Bai , Jiaxiang Shang , Runze Zhang , Jiayu Dong , Xin Wang , Guangyuan Sun , Hongbo Fu , Chiew-Lan Tai

Large scale text-guided diffusion models have garnered significant attention due to their ability to synthesize diverse images that convey complex visual concepts. This generative power has more recently been leveraged to perform text-to-3D…

Computer Vision and Pattern Recognition · Computer Science 2023-09-20 Etai Sella , Gal Fiebelman , Peter Hedman , Hadar Averbuch-Elor

Reconstructing three-dimensional (3D) scenes with semantic understanding is vital in many robotic applications. Robots need to identify which objects, along with their positions and shapes, to manipulate them precisely with given tasks.…

Robotics · Computer Science 2024-12-17 Khang Nguyen , Tuan Dang , Manfred Huber

Image-based 3D reconstruction is one of the most important tasks in Computer Vision with many solutions proposed over the last few decades. The objective is to extract metric information i.e. the geometry of scene objects directly from…

Computer Vision and Pattern Recognition · Computer Science 2022-09-16 Qiao Chen , Charalambos Poullis

Human shape estimation is an important task for video editing, animation and fashion industry. Predicting 3D human body shape from natural images, however, is highly challenging due to factors such as variation in human bodies, clothing and…

Computer Vision and Pattern Recognition · Computer Science 2018-08-21 Gül Varol , Duygu Ceylan , Bryan Russell , Jimei Yang , Ersin Yumer , Ivan Laptev , Cordelia Schmid

We present a new method for 3D shape reconstruction from a single image, in which a deep neural network directly maps an image to a vector of network weights. The network \textcolor{black}{parametrized by} these weights represents a 3D…

Computer Vision and Pattern Recognition · Computer Science 2019-08-20 Gidi Littwin , Lior Wolf

This paper introduces VolMap, a real-time approach for the semantic segmentation of a 3D LiDAR surrounding view system in autonomous vehicles. We designed an optimized deep convolution neural network that can accurately segment the point…

Computer Vision and Pattern Recognition · Computer Science 2019-07-01 Hager Radi , Waleed Ali

Worldwide, in 2014, more than 1.9 billion adults, 18 years and older, were overweight. Of these, over 600 million were obese. Accurately documenting dietary caloric intake is crucial to manage weight loss, but also presents challenges…

Computer Vision and Pattern Recognition · Computer Science 2016-06-21 Chang Liu , Yu Cao , Yan Luo , Guanling Chen , Vinod Vokkarane , Yunsheng Ma

Learning effective multi-modal 3D representations of objects is essential for numerous applications, such as augmented reality and robotics. Existing methods often rely on task-specific embeddings that are tailored either for semantic…

Computer Vision and Pattern Recognition · Computer Science 2025-11-06 Gaia Di Lorenzo , Federico Tombari , Marc Pollefeys , Daniel Barath

Rapid advancements in text-to-3D generation require robust and scalable evaluation metrics that align closely with human judgment, a need unmet by current metrics such as PSNR and CLIP, which require ground-truth data or focus only on…

Computer Vision and Pattern Recognition · Computer Science 2025-04-14 Shalini Maiti , Lourdes Agapito , Filippos Kokkinos

We present TexMesh, a novel approach to reconstruct detailed human meshes with high-resolution full-body texture from RGB-D video. TexMesh enables high quality free-viewpoint rendering of humans. Given the RGB frames, the captured…

Computer Vision and Pattern Recognition · Computer Science 2020-09-22 Tiancheng Zhi , Christoph Lassner , Tony Tung , Carsten Stoll , Srinivasa G. Narasimhan , Minh Vo

The three-dimensional reconstruction of vocal folds in medicine usually involves endoscopy and an approach to extract depth information like structured light or stereo matching of images. The resulting mesh can accurately represent the…

Fluid Dynamics · Physics 2023-10-06 Daniel Zieger , Christoph Näger , Stefan Becker , Tobias Günther
‹ Prev 1 4 5 6 7 8 10 Next ›