中文
相关论文

相关论文: MozzaVID: Mozzarella Volumetric Image Dataset

200 篇论文

The rapid progress of Multimodal Large Language Models (MLLMs) has unlocked the potential for enhanced 3D scene understanding and spatial reasoning. A recent line of work explores learning spatial reasoning directly from multi-view images,…

计算机视觉与模式识别 · 计算机科学 2026-04-10 Kanghee Lee , Injae Lee , Minseok Kwak , Jungi Hong , Kwonyoung Ryu , Jaesik Park

77% of adults over 50 want to age in place today, presenting a major challenge to ensuring adequate nutritional intake. It has been reported that one in four older adults that are 65 years or older are malnourished and given the direct link…

计算机视觉与模式识别 · 计算机科学 2023-04-13 Chi-en Amy Tai , Matthew Keller , Mattie Kerrigan , Yuhao Chen , Saeejith Nair , Pengcheng Xi , Alexander Wong

Dynamic volumetric MRI provides valuable information on in vivo motion and biomechanics, with applications spanning cardiac, musculoskeletal, or pulmonary imaging, amongst others. Developing reconstruction methods for time-resolved…

A volumetric attention(VA) module for 3D medical image segmentation and detection is proposed. VA attention is inspired by recent advances in video processing, enables 2.5D networks to leverage context information along the z direction, and…

图像与视频处理 · 电气工程与系统科学 2020-04-07 Xudong Wang , Shizhong Han , Yunqiang Chen , Dashan Gao , Nuno Vasconcelos

With the development of the medical image field, researchers seek to develop a class of datasets to block the need for medical knowledge, such as \text{MedMNIST} (v2). MedMNIST (v2) includes a large number of small-sized (28 $\times$ 28 or…

计算机视觉与模式识别 · 计算机科学 2023-04-21 Zhuoran Zheng , Xiuyi Jia

In this work, we introduce RadImageNet-VQA, a large-scale dataset designed to advance radiologic visual question answering (VQA) on CT and MRI exams. Existing medical VQA datasets are limited in scale, dominated by X-ray imaging or…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Léo Butsanets , Charles Corbière , Julien Khlaut , Pierre Manceron , Corentin Dancette

Background: Maintaining a healthy diet is vital to avoid health-related issues, e.g., undernutrition, obesity and many non-communicable diseases. An indispensable part of the health diet is dietary assessment. Traditional manual recording…

计算机视觉与模式识别 · 计算机科学 2022-03-08 Wei Wang , Weiqing Min , Tianhao Li , Xiaoxiao Dong , Haisheng Li , Shuqiang Jiang

Food logo detection plays an important role in the multimedia for its wide real-world applications, such as food recommendation of the self-service shop and infringement detection on e-commerce platforms. A large-scale food logo dataset is…

计算机视觉与模式识别 · 计算机科学 2021-08-11 Qiang Hou , Weiqing Min , Jing Wang , Sujuan Hou , Yuanjie Zheng , Shuqiang Jiang

Convolutional Neural Networks (CNNs) have been recently employed to solve problems from both the computer vision and medical image analysis fields. Despite their popularity, most approaches are only able to process 2D images while most…

计算机视觉与模式识别 · 计算机科学 2016-06-16 Fausto Milletari , Nassir Navab , Seyed-Ahmad Ahmadi

Cluster closure, defined as the progressive filling of gaps between the berries in a grape bunch, is a key trait in vineyard management, impacting disease risk. However, traditional visual scoring methods are labor-intensive, subjective,…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Xiangzhi Tong , Chengrui Zhang , Mac Flaherty , Andre Matteo Garcia , Dominic Gorman , Jonathan Jaramillo , Justine E. Vanden Heuvel , Yu Jiang

Deep convolutional neural networks require large amounts of labeled data samples. For many real-world applications, this is a major limitation which is commonly treated by augmentation methods. In this work, we address the problem of…

计算机视觉与模式识别 · 计算机科学 2022-08-01 Christoph Reinders , Frederik Schubert , Bodo Rosenhahn

Accurate assessment of dietary intake requires improved tools to overcome limitations of current methods including user burden and measurement error. Emerging technologies such as image-based approaches using advanced machine learning…

计算机视觉与模式识别 · 计算机科学 2021-10-06 Zeman Shao , Yue Han , Jiangpeng He , Runyu Mao , Janine Wright , Deborah Kerr , Carol Boushey , Fengqing Zhu

Modern deep learning techniques have enabled advances in image-based dietary assessment such as food recognition and food portion size estimation. Valuable information on the types of foods and the amount consumed are crucial for prevention…

计算机视觉与模式识别 · 计算机科学 2021-02-02 Jiangpeng He , Runyu Mao , Zeman Shao , Janine L. Wright , Deborah A. Kerr , Carol J. Boushey , Fengqing Zhu

CT reconstruction provides radiologists with images for diagnosis and treatment, yet current deep learning methods are typically limited to specific anatomies and datasets, hindering generalization ability to unseen anatomies and lesions.…

图像与视频处理 · 电气工程与系统科学 2025-10-31 Shaokai Wu , Yapan Guo , Yanbiao Ji , Jing Tong , Yuxiang Lu , Mei Li , Suizhi Huang , Yue Ding , Hongtao Lu

We present Picasso, a CUDA-based library comprising novel modules for deep learning over complex real-world 3D meshes. Hierarchical neural architectures have proved effective in multi-scale feature extraction which signifies the need for…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Huan Lei , Naveed Akhtar , Ajmal Mian

Reconstructing texture-less surfaces poses unique challenges in computer vision, primarily due to the lack of specialized datasets that cater to the nuanced needs of depth and normals estimation in the absence of textural information. We…

计算机视觉与模式识别 · 计算机科学 2024-11-06 Muhammad Saif Ullah Khan , Sankalp Sinha , Didier Stricker , Marcus Liwicki , Muhammad Zeshan Afzal

Detection and segmentation of objects in overheard imagery is a challenging task. The variable density, random orientation, small size, and instance-to-instance heterogeneity of objects in overhead imagery calls for approaches distinct from…

计算机视觉与模式识别 · 计算机科学 2021-02-25 Nicholas Weir , David Lindenbaum , Alexei Bastidas , Adam Van Etten , Sean McPherson , Jacob Shermeyer , Varun Kumar , Hanlin Tang

The integration of deep learning based systems in clinical practice is often impeded by challenges rooted in limited and heterogeneous medical datasets. In addition, the field has increasingly prioritized marginal performance gains on a…

图像与视频处理 · 电气工程与系统科学 2025-03-18 Sebastian Doerrich , Francesco Di Salvo , Julius Brockmann , Christian Ledig

Recent results of deep convolutional networks in visual recognition challenges open the path to a whole new set of disruptive user experiences such as visual search or recommendation. The list of companies offering this type of service is…

计算机视觉与模式识别 · 计算机科学 2019-09-20 Arnaud Bellétoile