中文
相关论文

相关论文: MFP3D: Monocular Food Portion Estimation Leveragin…

200 篇论文

Monocular depth estimation is a fundamental computer vision task. Recovering 3D depth from a single image is geometrically ill-posed and requires scene understanding, so it is not surprising that the rise of deep learning has led to a…

计算机视觉与模式识别 · 计算机科学 2024-04-04 Bingxin Ke , Anton Obukhov , Shengyu Huang , Nando Metzger , Rodrigo Caye Daudt , Konrad Schindler

Malnutrition is a major public health concern in low-and-middle-income countries (LMICs). Understanding food and nutrient intake across communities, households and individuals is critical to the development of health policies and…

计算机视觉与模式识别 · 计算机科学 2021-05-10 Frank Po Wen Lo , Modou L Jobarteh , Yingnan Sun , Jianing Qiu , Shuo Jiang , Gary Frost , Benny Lo

In recent years, monocular depth estimation is applied to understand the surrounding 3D environment and has made great progress. However, there is an ill-posed problem on how to gain depth information directly from a single image. With the…

计算机视觉与模式识别 · 计算机科学 2021-07-15 Meiqi Pei

3D pose estimation is an invaluable task in computer vision with various practical applications. Especially, 3D pose estimation for multi-person from a monocular video (3DMPPE) is particularly challenging and is still largely uncharted, far…

计算机视觉与模式识别 · 计算机科学 2023-09-19 Sungchan Park , Eunyi You , Inhoe Lee , Joonseok Lee

LiDAR point clouds can effectively depict the motion and posture of objects in three-dimensional space. Many studies accomplish the 3D object detection by voxelizing point clouds. However, in autonomous driving scenarios, the sparsity and…

计算机视觉与模式识别 · 计算机科学 2024-08-13 Yongxin Shao , Aihong Tan , Binrui Wang , Tianhong Yan , Zhetao Sun , Yiyang Zhang , Jiaxin Liu

Recent advances in data-driven geometric multi-view 3D reconstruction foundation models (e.g., DUSt3R) have shown remarkable performance across various 3D vision tasks, facilitated by the release of large-scale, high-quality 3D datasets.…

计算机视觉与模式识别 · 计算机科学 2025-04-21 Wenyu Li , Sidun Liu , Peng Qiao , Yong Dou

Fine-grained grocery object recognition is an important computer vision problem with broad applications in automatic checkout, in-store robotic navigation, and assistive technologies for the visually impaired. Existing datasets on groceries…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Shivanand Venkanna Sheshappanavar , Tejas Anvekar , Shivanand Kundargi , Yufan Wang , Chandra Kambhamettu

This paper introduces KeyDiff3D, a framework for unsupervised monocular 3D keypoints estimation that accurately predicts 3D keypoints from a single image. While previous methods rely on manual annotations or calibrated multi-view images,…

计算机视觉与模式识别 · 计算机科学 2025-07-17 Subin Jeon , In Cho , Junyoung Hong , Seon Joo Kim

The knowledge transfer from 3D printing technology paved the way for unlocking the innovative potential of 3D Food Printing (3DFP) technology. However, this technology-oriented approach neglects userderived issues that could be addressed…

人机交互 · 计算机科学 2025-03-14 Yağmur Kocaman , Taylan U. Bulut , Oğuzhan Özcan

3D vision is of paramount importance for numerous applications ranging from machine intelligence to precision metrology. Despite much recent progress, the majority of 3D imaging hardware remains bulky and complicated and provides much lower…

计算机视觉与模式识别 · 计算机科学 2025-02-12 Zicheng Shen , Feng Zhao , Yibo Ni , Yuanmu Yang

Maintaining health and fitness through a balanced diet is essential for preventing non communicable diseases such as heart disease, diabetes, and cancer. NutriVision combines smart healthcare with computer vision and machine learning to…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Madhumita Veeramreddy , Ashok Kumar Pradhan , Swetha Ghanta , Laavanya Rachakonda , Saraju P Mohanty

Accurate monocular metric depth estimation (MMDE) is crucial to solving downstream tasks in 3D perception and modeling. However, the remarkable accuracy of recent MMDE methods is confined to their training domains. These methods fail to…

计算机视觉与模式识别 · 计算机科学 2024-03-29 Luigi Piccinelli , Yung-Hsu Yang , Christos Sakaridis , Mattia Segu , Siyuan Li , Luc Van Gool , Fisher Yu

Monocular depth estimation plays a crucial role in 3D recognition and understanding. One key limitation of existing approaches lies in their lack of structural information exploitation, which leads to inaccurate spatial layout,…

计算机视觉与模式识别 · 计算机科学 2020-07-23 Tian Chen , Shijie An , Yuan Zhang , Chongyang Ma , Huayan Wang , Xiaoyan Guo , Wen Zheng

Enabling the synthesis of arbitrarily novel viewpoint images within a patient's stomach from pre-captured monocular gastroscopic images is a promising topic in stomach diagnosis. Typical methods to achieve this objective integrate…

计算机视觉与模式识别 · 计算机科学 2024-05-30 Zijie Jiang , Yusuke Monno , Masatoshi Okutomi , Sho Suzuki , Kenji Miki

Food classification is critical to the analysis of nutrients comprising foods reported in dietary assessment. Advances in mobile and wearable sensors, combined with new image based methods, particularly deep learning based approaches, have…

计算机视觉与模式识别 · 计算机科学 2022-06-07 Zeman Shao , Jiangpeng He , Ya-Yuan Yu , Luotao Lin , Alexandra Cowan , Heather Eicher-Miller , Fengqing Zhu

Assessing dietary intake accurately remains an open and challenging research problem. In recent years, image-based approaches have been developed to automatically estimate food intake by capturing eat occasions with mobile devices and…

信息检索 · 计算机科学 2019-10-16 Zeman Shao , Runyu Mao , Fengqing Zhu

Point clouds have become an increasingly important representation for 3D medical imaging, offering a compact, surface-preserving alternative to traditional voxel or mesh-based approaches. Recent advances in deep learning have enabled rapid…

图像与视频处理 · 电气工程与系统科学 2026-02-04 Tongxu Zhang , Zhiming Liang , Bei Wang

Point clouds are a very efficient way to represent volumetric data in medical imaging. First, they do not occupy resources for empty spaces and therefore can avoid trade-offs between resolution and field-of-view for voxel-based 3D…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Mattias Paul Heinrich

In this paper, we propose PASS3D to achieve point-wise semantic segmentation for 3D point cloud. Our framework combines the efficiency of traditional geometric methods with robustness of deep learning methods, consisting of two stages: At…

计算机视觉与模式识别 · 计算机科学 2020-08-27 Xin Kong , Guangyao Zhai , Baoquan Zhong , Yong Liu

Representing 3D shape in deep learning frameworks in an accurate, efficient and compact manner still remains an open challenge. Most existing work addresses this issue by employing voxel-based representations. While these approaches benefit…

计算机视觉与模式识别 · 计算机科学 2019-01-25 Dominic Jack , Jhony K. Pontes , Sridha Sridharan , Clinton Fookes , Sareh Shirazi , Frederic Maire , Anders Eriksson