中文
相关论文

相关论文: MozzaVID: Mozzarella Volumetric Image Dataset

200 篇论文

The notion of visual similarity is essential for computer vision, and in applications and studies revolving around vector embeddings of images. However, the scarcity of benchmark datasets poses a significant hurdle in exploring how these…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Tillmann Ohm , Andres Karjus , Mikhail Tamm , Maximilian Schich

The health and function of tissue rely on its vasculature network to provide reliable blood perfusion. Volumetric imaging approaches, such as multiphoton microscopy, are able to generate detailed 3D images of blood vessels that could…

计算机视觉与模式识别 · 计算机科学 2019-06-19 Mohammad Haft-Javaherian , Linjing Fang , Victorine Muse , Chris B. Schaffer , Nozomi Nishimura , Mert R. Sabuncu

Domain shift significantly influences the performance of deep learning algorithms, particularly for object detection within volumetric 3D images. Annotated training data is essential for deep learning-based object detection. However,…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Patrick Møller Jensen , Vedrana Andersen Dahl , Carsten Gundlach , Rebecca Engberg , Hans Martin Kjer , Anders Bjorholm Dahl

The rapid advancement of AI-generated multimodal video-audio content has raised significant concerns regarding information security and content authenticity. Existing synthetic video datasets predominantly focus on the visual modality…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Mengxue Hu , Yunfeng Diao , Changtao Miao , Zhiqing Guo , Jianshu Li , Zhe Li , Joey Tianyi Zhou

The importance of wild video based image set recognition is becoming monotonically increasing. However, the contents of these collected videos are often complicated, and how to efficiently perform set modeling and feature extraction is a…

计算机视觉与模式识别 · 计算机科学 2019-08-07 Rui Wang , XiaoJun Wu , Josef Kittler

Currently, food image recognition tasks are evaluated against fixed datasets. However, in real-world conditions, there are cases in which the number of samples in each class continues to increase and samples from novel classes appear. In…

计算机视觉与模式识别 · 计算机科学 2022-12-23 Shota Horiguchi , Sosuke Amano , Makoto Ogawa , Kiyoharu Aizawa

Convolutional Neural Networks (ConvNets) at present achieve remarkable performance in image classification tasks. However, current ConvNets cannot guarantee the capabilities of the mammalian visual systems such as invariance to contrast and…

计算机视觉与模式识别 · 计算机科学 2021-09-16 E. Ulises Moya-Sánchez , Sebastiá Xambo-Descamps , Abraham Sánchez , Sebastián Salazar-Colores , Ulises Cortés

We present MM-Food-100K, a public 100,000-sample multimodal food intelligence dataset with verifiable provenance. It is a curated approximately 10% open subset of an original 1.2 million, quality-accepted corpus of food images annotated for…

人工智能 · 计算机科学 2025-08-15 Yi Dong , Yusuke Muraoka , Scott Shi , Yi Zhang

High-dimensional structural MRI (sMRI) images are widely used for Alzheimer's Disease (AD) diagnosis. Most existing methods for sMRI representation learning rely on 3D architectures (e.g., 3D CNNs), slice-wise feature extraction with late…

计算机视觉与模式识别 · 计算机科学 2026-01-30 Dexuan Ding , Ciyuan Peng , Endrowednes Kuantama , Jingcai Guo , Jia Wu , Jian Yang , Amin Beheshti , Ming-Hsuan Yang , Yuankai Qi

The irregular geometry and high inter-slice variability in computerized tomography (CT) scans of the human pancreas make an accurate segmentation of this crucial organ a challenging task for existing data-driven deep learning methods. To…

计算机视觉与模式识别 · 计算机科学 2020-10-20 Hao Li , Jun Li , Xiaozhu Lin , Xiaohua Qian

The gastrointestinal (GI) tract of humans can have a wide variety of aberrant mucosal abnormality findings, ranging from mild irritations to extremely fatal illnesses. Prompt identification of gastrointestinal disorders greatly contributes…

Large-scale medical imaging datasets have accelerated deep learning (DL) for medical image analysis. However, the large scale of these datasets poses a challenge for researchers, resulting in increased storage and bandwidth requirements for…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Pranav Kulkarni , Adway Kanhere , Eliot Siegel , Paul H. Yi , Vishwa S. Parekh

Chest X-rays remain the primary diagnostic tool in emergency medicine, yet their limited ability to capture fine anatomical details can result in missed or delayed diagnoses. To address this, we introduce XVertNet, a novel deep-learning…

图像与视频处理 · 电气工程与系统科学 2025-09-03 Ella Eidlin , Assaf Hoogi , Hila Rozen , Mohammad Badarne , Nathan S. Netanyahu

In this article, we present a new unique dataset for dental research - AlphaDent. This dataset is based on the DSLR camera photographs of the teeth of 295 patients and contains over 1200 images. The dataset is labeled for solving the…

Radiology reports contain rich clinical information that can be used to train imaging models without relying on costly manual annotation. However, existing approaches face critical limitations: rule-based methods struggle with linguistic…

We release two artificial datasets, Simulated Flying Shapes and Simulated Planar Manipulator that allow to test the learning ability of video processing systems. In particular, the dataset is meant as a tool which allows to easily assess…

计算机视觉与模式识别 · 计算机科学 2018-07-03 Fabio Ferreira , Jonas Rothfuss , Eren Erdal Aksoy , You Zhou , Tamim Asfour

Breast cancer as a medical condition and mammograms as images exhibit many dimensions of variability across the population. Similarly, the way diagnostic systems are used and maintained by clinicians varies between imaging centres and…

With the rapid development of society and continuous advances in science and technology, the food industry increasingly demands higher production quality and efficiency. Food image classification plays a vital role in enabling automated…

计算机视觉与模式识别 · 计算机科学 2025-09-24 Xinle Gao , Linghui Ye , Zhiyong Xiao

Field archeologists are called upon to identify potsherds, for which purpose they rely on their experience and on reference works. We have developed two complementary machine-learning tools to propose identifications based on images…

计算机视觉与模式识别 · 计算机科学 2019-11-25 Barak Itkin , Lior Wolf , Nachum Dershowitz

Humans make accurate decisions by interpreting complex data from multiple sources. Medical diagnostics, in particular, often hinge on human interpretation of multi-modal information. In order for artificial intelligence to make progress in…

计算机视觉与模式识别 · 计算机科学 2018-11-20 Faisal Mahmood , Ziyun Yang , Thomas Ashley , Nicholas J. Durr