English
Related papers

Related papers: MozzaVID: Mozzarella Volumetric Image Dataset

200 papers

The rapid progress of Multimodal Large Language Models (MLLMs) has unlocked the potential for enhanced 3D scene understanding and spatial reasoning. A recent line of work explores learning spatial reasoning directly from multi-view images,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Kanghee Lee , Injae Lee , Minseok Kwak , Jungi Hong , Kwonyoung Ryu , Jaesik Park

77% of adults over 50 want to age in place today, presenting a major challenge to ensuring adequate nutritional intake. It has been reported that one in four older adults that are 65 years or older are malnourished and given the direct link…

Computer Vision and Pattern Recognition · Computer Science 2023-04-13 Chi-en Amy Tai , Matthew Keller , Mattie Kerrigan , Yuhao Chen , Saeejith Nair , Pengcheng Xi , Alexander Wong

Dynamic volumetric MRI provides valuable information on in vivo motion and biomechanics, with applications spanning cardiac, musculoskeletal, or pulmonary imaging, amongst others. Developing reconstruction methods for time-resolved…

A volumetric attention(VA) module for 3D medical image segmentation and detection is proposed. VA attention is inspired by recent advances in video processing, enables 2.5D networks to leverage context information along the z direction, and…

Image and Video Processing · Electrical Eng. & Systems 2020-04-07 Xudong Wang , Shizhong Han , Yunqiang Chen , Dashan Gao , Nuno Vasconcelos

With the development of the medical image field, researchers seek to develop a class of datasets to block the need for medical knowledge, such as \text{MedMNIST} (v2). MedMNIST (v2) includes a large number of small-sized (28 $\times$ 28 or…

Computer Vision and Pattern Recognition · Computer Science 2023-04-21 Zhuoran Zheng , Xiuyi Jia

In this work, we introduce RadImageNet-VQA, a large-scale dataset designed to advance radiologic visual question answering (VQA) on CT and MRI exams. Existing medical VQA datasets are limited in scale, dominated by X-ray imaging or…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Léo Butsanets , Charles Corbière , Julien Khlaut , Pierre Manceron , Corentin Dancette

Background: Maintaining a healthy diet is vital to avoid health-related issues, e.g., undernutrition, obesity and many non-communicable diseases. An indispensable part of the health diet is dietary assessment. Traditional manual recording…

Computer Vision and Pattern Recognition · Computer Science 2022-03-08 Wei Wang , Weiqing Min , Tianhao Li , Xiaoxiao Dong , Haisheng Li , Shuqiang Jiang

Food logo detection plays an important role in the multimedia for its wide real-world applications, such as food recommendation of the self-service shop and infringement detection on e-commerce platforms. A large-scale food logo dataset is…

Computer Vision and Pattern Recognition · Computer Science 2021-08-11 Qiang Hou , Weiqing Min , Jing Wang , Sujuan Hou , Yuanjie Zheng , Shuqiang Jiang

Convolutional Neural Networks (CNNs) have been recently employed to solve problems from both the computer vision and medical image analysis fields. Despite their popularity, most approaches are only able to process 2D images while most…

Computer Vision and Pattern Recognition · Computer Science 2016-06-16 Fausto Milletari , Nassir Navab , Seyed-Ahmad Ahmadi

Cluster closure, defined as the progressive filling of gaps between the berries in a grape bunch, is a key trait in vineyard management, impacting disease risk. However, traditional visual scoring methods are labor-intensive, subjective,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Xiangzhi Tong , Chengrui Zhang , Mac Flaherty , Andre Matteo Garcia , Dominic Gorman , Jonathan Jaramillo , Justine E. Vanden Heuvel , Yu Jiang

Deep convolutional neural networks require large amounts of labeled data samples. For many real-world applications, this is a major limitation which is commonly treated by augmentation methods. In this work, we address the problem of…

Computer Vision and Pattern Recognition · Computer Science 2022-08-01 Christoph Reinders , Frederik Schubert , Bodo Rosenhahn

Accurate assessment of dietary intake requires improved tools to overcome limitations of current methods including user burden and measurement error. Emerging technologies such as image-based approaches using advanced machine learning…

Computer Vision and Pattern Recognition · Computer Science 2021-10-06 Zeman Shao , Yue Han , Jiangpeng He , Runyu Mao , Janine Wright , Deborah Kerr , Carol Boushey , Fengqing Zhu

Modern deep learning techniques have enabled advances in image-based dietary assessment such as food recognition and food portion size estimation. Valuable information on the types of foods and the amount consumed are crucial for prevention…

Computer Vision and Pattern Recognition · Computer Science 2021-02-02 Jiangpeng He , Runyu Mao , Zeman Shao , Janine L. Wright , Deborah A. Kerr , Carol J. Boushey , Fengqing Zhu

CT reconstruction provides radiologists with images for diagnosis and treatment, yet current deep learning methods are typically limited to specific anatomies and datasets, hindering generalization ability to unseen anatomies and lesions.…

Image and Video Processing · Electrical Eng. & Systems 2025-10-31 Shaokai Wu , Yapan Guo , Yanbiao Ji , Jing Tong , Yuxiang Lu , Mei Li , Suizhi Huang , Yue Ding , Hongtao Lu

We present Picasso, a CUDA-based library comprising novel modules for deep learning over complex real-world 3D meshes. Hierarchical neural architectures have proved effective in multi-scale feature extraction which signifies the need for…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Huan Lei , Naveed Akhtar , Ajmal Mian

Reconstructing texture-less surfaces poses unique challenges in computer vision, primarily due to the lack of specialized datasets that cater to the nuanced needs of depth and normals estimation in the absence of textural information. We…

Computer Vision and Pattern Recognition · Computer Science 2024-11-06 Muhammad Saif Ullah Khan , Sankalp Sinha , Didier Stricker , Marcus Liwicki , Muhammad Zeshan Afzal

Detection and segmentation of objects in overheard imagery is a challenging task. The variable density, random orientation, small size, and instance-to-instance heterogeneity of objects in overhead imagery calls for approaches distinct from…

Computer Vision and Pattern Recognition · Computer Science 2021-02-25 Nicholas Weir , David Lindenbaum , Alexei Bastidas , Adam Van Etten , Sean McPherson , Jacob Shermeyer , Varun Kumar , Hanlin Tang

The integration of deep learning based systems in clinical practice is often impeded by challenges rooted in limited and heterogeneous medical datasets. In addition, the field has increasingly prioritized marginal performance gains on a…

Image and Video Processing · Electrical Eng. & Systems 2025-03-18 Sebastian Doerrich , Francesco Di Salvo , Julius Brockmann , Christian Ledig

Recent results of deep convolutional networks in visual recognition challenges open the path to a whole new set of disruptive user experiences such as visual search or recommendation. The list of companies offering this type of service is…

Computer Vision and Pattern Recognition · Computer Science 2019-09-20 Arnaud Bellétoile