中文
相关论文

相关论文: MozzaVID: Mozzarella Volumetric Image Dataset

200 篇论文

In response to the increasing demand for efficient and non-invasive methods to estimate food weight, this paper presents a vision-based approach utilizing 2D images. The study employs a dataset of 2380 images comprising fourteen different…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Chathura Wimalasiri , Prasan Kumar Sahoo

The advent of social media platforms has been a catalyst for the development of digital photography that engendered a boom in vision applications. With this motivation, we introduce a large-scale dataset termed 'Photozilla', which includes…

计算机视觉与模式识别 · 计算机科学 2021-06-23 Trisha Singhal , Junhua Liu , Lucienne T. M. Blessing , Kwan Hui Lim

In medical imaging analysis, deep learning has shown promising results. We frequently rely on volumetric data to segment medical images, necessitating the use of 3D architectures, which are commended for their capacity to capture interslice…

图像与视频处理 · 电气工程与系统科学 2023-05-18 Ikboljon Sobirov , Numan Saeed , Mohammad Yaqub

Recent advancements in Large Multimodal Models (LMMs) have shown promising results in mathematical reasoning within visual contexts, with models approaching human-level performance on existing benchmarks such as MathVista. However, we…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Ke Wang , Junting Pan , Weikang Shi , Zimu Lu , Mingjie Zhan , Hongsheng Li

The COVID19 pandemic has had a detrimental impact on the health and welfare of the worlds population. An important strategy in the fight against COVID19 is the effective screening of infected patients, with one of the primary screening…

图像与视频处理 · 电气工程与系统科学 2024-11-05 Nafiz Fahad , Fariha Jahan , Md Kishor Morol , Rasel Ahmed , Md. Abdullah-Al-Jubair

Recognizing food images presents unique challenges due to the variable spatial layout and shape changes of ingredients with different cooking and cutting methods. This study introduces an advanced approach for recognizing ingredients…

计算机视觉与模式识别 · 计算机科学 2024-12-06 Kun Fu , Ying Dai

With the arrival of convolutional neural networks, the complex problem of food recognition has experienced an important improvement in recent years. The best results have been obtained using methods based on very deep convolutional neural…

计算机视觉与模式识别 · 计算机科学 2018-01-23 Eduardo Aguilar , Marc Bolaños , Petia Radeva

Accurate dietary intake estimation is critical for informing policies and programs to support healthy eating, as malnutrition has been directly linked to decreased quality of life. However self-reporting methods such as food diaries suffer…

Surface prediction and completion have been widely studied in various applications. Recently, research in surface completion has evolved from small objects to complex large-scale scenes. As a result, researchers have begun increasing the…

机器人学 · 计算机科学 2024-03-19 Guiyong Zheng , Jinqi Jiang , Chen Feng , Shaojie Shen , Boyu Zhou

Datasets (semi-)automatically collected from the web can easily scale to millions of entries, but a dataset's usefulness is directly related to how clean and high-quality its examples are. In this paper, we describe and publicly release an…

计算机视觉与模式识别 · 计算机科学 2020-08-24 Houda Alberts , Iacer Calixto

Being data-driven is one of the most iconic properties of deep learning algorithms. The birth of ImageNet drives a remarkable trend of "learning from large-scale data" in computer vision. Pretraining on ImageNet to obtain rich universal…

Many state-of-the art visualization techniques must be tailored to the specific type of dataset, its modality (CT, MRI, etc.), the recorded object or anatomical region (head, spine, abdomen, etc.) and other parameters related to the data…

图形学 · 计算机科学 2009-06-15 Dženan Zukić , Christof Rezk-Salama , Andreas Kolb

Food image recognition is a challenging task in computer vision due to the high variability and complexity of food images. In this study, we investigate the potential of Noisy Vision Transformers (NoisyViT) for improving food classification…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Tonmoy Ghosh , Edward Sazonov

Small Video Object Detection (SVOD) is a crucial subfield in modern computer vision, essential for early object discovery and detection. However, existing SVOD datasets are scarce and suffer from issues such as insufficiently small objects,…

计算机视觉与模式识别 · 计算机科学 2024-07-26 Jiahao Guo , Ziyang Xu , Lianjun Wu , Fei Gao , Wenyu Liu , Xinggang Wang

Nowadays, it is common for people to take photographs of every beverage, snack, or meal they eat and then post these photographs on social media platforms. Leveraging these social trends, real-time food recognition and reliable…

计算机视觉与模式识别 · 计算机科学 2023-05-15 Aknur Karabay , Arman Bolatov , Huseyin Atakan Varol , Mei-Yen Chan

We introduce the first comprehensive 3D dataset for the task of unsupervised anomaly detection and localization. It is inspired by real-world visual inspection scenarios in which a model has to detect various types of defects on…

计算机视觉与模式识别 · 计算机科学 2022-02-25 Paul Bergmann , Xin Jin , David Sattlegger , Carsten Steger

Food image segmentation is a critical and indispensible task for developing health-related applications such as estimating food calories and nutrients. Existing food image segmentation models are underperforming due to two reasons: (1)…

计算机视觉与模式识别 · 计算机科学 2021-05-13 Xiongwei Wu , Xin Fu , Ying Liu , Ee-Peng Lim , Steven C. H. Hoi , Qianru Sun

Cancer diseases constitute one of the most significant societal challenges. In this paper, we introduce a novel histopathological dataset for prostate cancer detection. The proposed dataset, consisting of over 2.6 million tissue patches…

We propose Encyclopedic-VQA, a large scale visual question answering (VQA) dataset featuring visual questions about detailed properties of fine-grained categories and instances. It contains 221k unique question+answer pairs each matched…

计算机视觉与模式识别 · 计算机科学 2023-07-25 Thomas Mensink , Jasper Uijlings , Lluis Castrejon , Arushi Goel , Felipe Cadar , Howard Zhou , Fei Sha , André Araujo , Vittorio Ferrari

We present Fashion-MNIST, a new dataset comprising of 28x28 grayscale images of 70,000 fashion products from 10 categories, with 7,000 images per category. The training set has 60,000 images and the test set has 10,000 images. Fashion-MNIST…

机器学习 · 计算机科学 2017-09-19 Han Xiao , Kashif Rasul , Roland Vollgraf