中文
相关论文

相关论文: MozzaVID: Mozzarella Volumetric Image Dataset

200 篇论文

Prior work has studied different visual modalities in isolation and developed separate architectures for recognition of images, videos, and 3D data. Instead, in this paper, we propose a single model which excels at classifying images,…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Rohit Girdhar , Mannat Singh , Nikhila Ravi , Laurens van der Maaten , Armand Joulin , Ishan Misra

We introduce TableVista, a comprehensive benchmark for evaluating foundation models in multimodal table reasoning under visual and structural complexity. TableVista consists of 3,000 high-quality table reasoning problems, where each…

计算与语言 · 计算机科学 2026-05-08 Zheyuan Yang , Liqiang Shang , Junjie Chen , Xun Yang , Chenglong Xu , Bo Yuan , Chenyuan Jiao , Yaoru Sun , Yilun Zhao

Food image classification is the fundamental step in image-based dietary assessment, which aims to estimate participants' nutrient intake from eating occasion images. A common challenge of food images is the intra-class diversity and…

计算机视觉与模式识别 · 计算机科学 2024-08-08 Xinyue Pan , Jiangpeng He , Fengqing Zhu

Food image classification is a fundamental step of image-based dietary assessment, enabling automated nutrient analysis from food images. Many current methods employ deep neural networks to train on generic food image datasets that do not…

计算机视觉与模式识别 · 计算机科学 2023-09-19 Xinyue Pan , Jiangpeng He , Fengqing Zhu

Computational pathology tasks have some unique characterises such as multi-gigapixel images, tedious and frequently uncertain annotations, and unavailability of large number of cases [13]. To address some of these issues, we present Deep…

图像与视频处理 · 电气工程与系统科学 2023-01-24 Nima Hatami

Molecular and morphological characters, as important parts of biological taxonomy, are contradictory but need to be integrated. Organism's image recognition and bioinformatics are emerging and hot problems nowadays but with a gap between…

计算机视觉与模式识别 · 计算机科学 2022-06-29 Jiewen Xiao , Wenbin Liao , Ming Zhang , Jing Wang , Jianxin Wang , Yihua Yang

The widespread adoption of camera-equipped mobile devices and wearables has enabled convenient capture of meal images, making food recognition a key component for real time dietary monitoring. However, real-world food images present…

人工智能 · 计算机科学 2026-05-08 Woojin Lee , Pranav Mekkoth , Ye Tian , Onat Gungor , Tajana Rosing

Estimating the number of clusters and cluster structures in unlabeled, complex, and high-dimensional datasets (like images) is challenging for traditional clustering algorithms. In recent years, a matrix reordering-based algorithm called…

Healthcare relies on multiple types of data, such as medical images, genetic information, and clinical records, to improve diagnosis and treatment. However, missing data is a common challenge due to privacy restrictions, cost, and technical…

机器学习 · 计算机科学 2025-03-13 Nazanin Moradinasab , Saurav Sengupta , Jiebei Liu , Sana Syed , Donald E. Brown

Food classification is critical to the analysis of nutrients comprising foods reported in dietary assessment. Advances in mobile and wearable sensors, combined with new image based methods, particularly deep learning based approaches, have…

计算机视觉与模式识别 · 计算机科学 2022-06-07 Zeman Shao , Jiangpeng He , Ya-Yuan Yu , Luotao Lin , Alexandra Cowan , Heather Eicher-Miller , Fengqing Zhu

Mammography, or breast X-ray, is the most widely used imaging modality to detect cancer and other breast diseases. Recent studies have shown that deep learning-based computer-assisted detection and diagnosis (CADe or CADx) tools have been…

图像与视频处理 · 电气工程与系统科学 2023-03-20 Hieu T. Nguyen , Ha Q. Nguyen , Hieu H. Pham , Khanh Lam , Linh T. Le , Minh Dao , Van Vu

Currently, the computational complexity limits the training of high resolution gigapixel images using Convolutional Neural Networks. Therefore, such images are divided into patches or tiles. Since, these high resolution patches are encoded…

计算机视觉与模式识别 · 计算机科学 2022-02-23 Suvidha Tripathi , Satish Kumar Singh , Lee Hwee Kuan

In the realm of food computing, segmenting ingredients from images poses substantial challenges due to the large intra-class variance among the same ingredients, the emergence of new ingredients, and the high annotation costs associated…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Xiongwei Wu , Sicheng Yu , Ee-Peng Lim , Chong-Wah Ngo

Over the past decade, 3D graphics have become highly detailed to mimic the real world, exploding their size and complexity. Certain applications and device constraints necessitate their simplification and/or lossy compression, which can…

Recent deep learning approaches in table detection achieved outstanding performance and proved to be effective in identifying document layouts. Currently, available table detection benchmarks have many limitations, including the lack of…

计算机视觉与模式识别 · 计算机科学 2023-12-01 Mrinal Haloi , Shashank Shekhar , Nikhil Fande , Siddhant Swaroop Dash , Sanjay G

Deep learning underpins a wide range of applications in MRI, including reconstruction, artifact removal, and segmentation. However, progress has been driven largely by public datasets focused on brain and knee imaging, shaping how models…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Paula Arguello , Berk Tinaz , Mohammad Shahab Sepehri , Maryam Soltanolkotabi , Mahdi Soltanolkotabi

In contemporary society, the application of artificial intelligence for automatic food recognition offers substantial potential for nutrition tracking, reducing food waste, and enhancing productivity in food production and consumption…

计算机视觉与模式识别 · 计算机科学 2024-05-21 Shayan Rokhva , Babak Teimourpour , Amir Hossein Soltani

Accurate food volume estimation is crucial for dietary monitoring, medical nutrition management, and food intake analysis. Existing 3D Food Volume estimation methods accurately compute the food volume but lack for food portions selection.…

图形学 · 计算机科学 2025-06-04 Ahmad AlMughrabi , Umair Haroon , Ricardo Marques , Petia Radeva

We introduce a learning-based depth map fusion framework that accepts a set of depth and confidence maps generated by a Multi-View Stereo (MVS) algorithm as input and improves them. This is accomplished by integrating volumetric visibility…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Nathaniel Burgdorfer , Philippos Mordohai

While deep learning has recently achieved great success on multi-view stereo (MVS), limited training data makes the trained model hard to be generalized to unseen scenarios. Compared with other computer vision tasks, it is rather difficult…

计算机视觉与模式识别 · 计算机科学 2020-04-14 Yao Yao , Zixin Luo , Shiwei Li , Jingyang Zhang , Yufan Ren , Lei Zhou , Tian Fang , Long Quan