中文
相关论文

相关论文: VolE: A Point-cloud Framework for Food 3D Reconstr…

200 篇论文

Understanding and modeling the 3D scene from a single image is a practical problem. A recent advance proposes a panoptic 3D scene reconstruction task that performs both 3D reconstruction and 3D panoptic segmentation from a single image.…

计算机视觉与模式识别 · 计算机科学 2024-01-17 Tao Chu , Pan Zhang , Qiong Liu , Jiaqi Wang

Volume estimation in large indoor spaces is an important challenge in robotic inspection of industrial warehouses. We propose an approach for volume estimation for autonomous systems using visual features for indoor localization and surface…

系统与控制 · 电气工程与系统科学 2022-11-16 Samuel Balula , Dominic Liao-McPherson , Stefan Stevšić , Alisa Rupenyan , John Lygeros

3D reconstruction is a fundamental task in robotics that gained attention due to its major impact in a wide variety of practical settings, including agriculture, underwater, and urban environments. This task can be carried out via view…

When created faithfully from real-world data, Digital 3D representations of objects can be useful for human or computer-assisted analysis. Such models can also serve for generating training data for machine learning approaches in settings…

计算机视觉与模式识别 · 计算机科学 2024-07-08 David Nakath , Xiangyu Weng , Mengkun She , Kevin Köser

Solving image-to-3D from a single view is an ill-posed problem, and current neural reconstruction methods addressing it through diffusion models still rely on scene-specific optimization, constraining their generalization capability. To…

计算机视觉与模式识别 · 计算机科学 2024-01-09 Christian Simon , Sen He , Juan-Manuel Perez-Rua , Mengmeng Xu , Amine Benhalloum , Tao Xiang

Generating high-fidelity 3D contents remains a fundamental challenge due to the complexity of representing arbitrary topologies-such as open surfaces and intricate internal structures-while preserving geometric details. Prevailing methods…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Xinran Yang , Shuichang Lai , Jiangjing Lyu , Hongjie Li , Bowen Pan , Yuanqi Li , Jie Guo , Zhengkang Zhou , Yanwen Guo

The Gaussian reconstruction kernels have been proposed by Westover (1990) and studied by the computer graphics community back in the 90s, which gives an alternative representation of object 3D geometry from meshes and point clouds. On the…

图形学 · 计算机科学 2024-01-30 Angtian Wang , Peng Wang , Jian Sun , Adam Kortylewski , Alan Yuille

A reasonable and balanced diet is essential for maintaining good health. With the advancements in deep learning, automated nutrition estimation method based on food images offers a promising solution for monitoring daily nutritional intake…

计算机视觉与模式识别 · 计算机科学 2023-10-19 Yuzhe Han , Qimin Cheng , Wenjin Wu , Ziyang Huang

Scale-aware monocular depth estimation poses a significant challenge in computer-aided endoscopic navigation. However, existing depth estimation methods that do not consider the geometric priors struggle to learn the absolute scale from…

计算机视觉与模式识别 · 计算机科学 2024-08-15 Ruofeng Wei , Bin Li , Kai Chen , Yiyao Ma , Yunhui Liu , Qi Dou

High calorie intake in the human body on the one hand, has proved harmful in numerous occasions leading to several diseases and on the other hand, a standard amount of calorie intake has been deemed essential by dieticians to maintain the…

计算机与社会 · 计算机科学 2015-03-24 Pallavi Kuhad , Abdulsalam Yassine , Shervin Shirmohammadi

Quantifying post-consumer food waste in institutional dining settings is essential for supporting data-driven sustainability strategies. This study presents a cost-effective computer vision framework that estimates plate-level food waste by…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Shayan Rokhva , Babak Teimourpour

Point clouds have become an increasingly important representation for 3D medical imaging, offering a compact, surface-preserving alternative to traditional voxel or mesh-based approaches. Recent advances in deep learning have enabled rapid…

图像与视频处理 · 电气工程与系统科学 2026-02-04 Tongxu Zhang , Zhiming Liang , Bei Wang

The interstellar medium (ISM) exhibits complex, multi-scale structures that are challenging to study due to their projection into two-dimensional (2D) column density maps. We present the Volume Density Mapper, a novel algorithm based on…

天体物理仪器与方法 · 物理学 2025-09-23 Guang-Xing Li , Mengke Zhao

Object manipulation is a critical skill required for Embodied AI agents interacting with the world around them. Training agents to manipulate objects, poses many challenges. These include occlusion of the target object by the agent's arm,…

计算机视觉与模式识别 · 计算机科学 2022-03-16 Kiana Ehsani , Ali Farhadi , Aniruddha Kembhavi , Roozbeh Mottaghi

Feed-forward 3D Gaussian Splatting (3DGS) has emerged as a highly effective solution for novel view synthesis. Existing methods predominantly rely on a \emph{pixel-aligned} Gaussian prediction paradigm, where each 2D pixel is mapped to a 3D…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Weijie Wang , Yeqing Chen , Zeyu Zhang , Hengyu Liu , Haoxiao Wang , Zhiyuan Feng , Wenkang Qin , Feng Chen , Zheng Zhu , Donny Y. Chen , Bohan Zhuang

In volume-to-volume translations in medical images, existing models often struggle to capture the inherent volumetric distribution using 3D voxelspace representations, due to high computational dataset demands. We present Score-Fusion, a…

计算机视觉与模式识别 · 计算机科学 2025-02-10 Xiyue Zhu , Dou Hoon Kwark , Ruike Zhu , Kaiwen Hong , Yiqi Tao , Shirui Luo , Yudu Li , Zhi-Pei Liang , Volodymyr Kindratenko

Three-dimensional ultrasound localization microscopy (ULM) enables comprehensive visualization of the vasculature, thereby improving diagnostic reliability. Nevertheless, its clinical translation remains challenging, as the exponential…

A laser scanner can easily acquire the geometric data of physical environments in the form of a point cloud. Recognizing objects from a point cloud is often required for industrial 3D reconstruction, which should include not only geometry…

计算机视觉与模式识别 · 计算机科学 2020-07-01 Hyungki Kim , Moohyun Cha , Duhwan Mun

A monocular 3D object tracking system generally has only up-to-scale pose estimation results without any prior knowledge of the tracked object. In this paper, we propose a novel idea to recover the metric scale of an arbitrary dynamic…

机器人学 · 计算机科学 2018-08-22 Kejie Qiu , Tong Qin , Hongwen Xie , Shaojie Shen

Estimating a scene reconstruction and the camera motion from in-body videos is challenging due to several factors, e.g. the deformation of in-body cavities or the lack of texture. In this paper we present Endo-Depth-and-Motion, a pipeline…

计算机视觉与模式识别 · 计算机科学 2021-07-06 David Recasens , José Lamarca , José M. Fácil , J. M. M. Montiel , Javier Civera