中文
相关论文

相关论文: FungiTastic: A multi-modal dataset and benchmark f…

200 篇论文

In this paper, we introduce a new large-scale face dataset named VGGFace2. The dataset contains 3.31 million images of 9131 subjects, with an average of 362.6 images for each subject. Images are downloaded from Google Image Search and have…

计算机视觉与模式识别 · 计算机科学 2018-05-15 Qiong Cao , Li Shen , Weidi Xie , Omkar M. Parkhi , Andrew Zisserman

Detection and classification of objects in overhead images are two important and challenging problems in computer vision. Among various research areas in this domain, the task of fine-grained classification of objects in overhead images has…

计算机视觉与模式识别 · 计算机科学 2021-05-28 Eran Dahan , Tzvi Diskin , Amit Amram , Amit Moryossef , Omer Koren

High-quality labeled datasets are essential for deep learning. Traditional manual annotation methods are not only costly and inefficient but also pose challenges in specialized domains where expert knowledge is needed. Self-supervised…

计算机视觉与模式识别 · 计算机科学 2023-11-16 Zhaocong liu , Fa Zhang , Lin Cheng , Huanxi Deng , Xiaoyan Yang , Zhenyu Zhang , Chichun Zhou

The automated evaluation of cognitive status utilizing multimedia technologies presents a promising frontier in early dementia diagnosis. However, the development of robust machine learning models for cognitive impairment detection is…

数据库 · 计算机科学 2026-04-03 Liuyu Wu , Rui Feng , Jie Li , Wentao Xiang , Yi Zhang , Yin Cao , Siyang Song , Xiao Gu , Jianqing Li , Wei Wang

Few-shot image classifiers are designed to recognize and classify new data with minimal supervision and limited data but often show reliance on spurious correlations between classes and spurious attributes, known as spurious bias. Spurious…

计算机视觉与模式识别 · 计算机科学 2024-09-05 Guangtao Zheng , Wenqian Ye , Aidong Zhang

The development of successful artificial intelligence models for chest X-ray analysis relies on large, diverse datasets with high-quality annotations. While several databases of chest X-ray images have been released, most include disease…

图像与视频处理 · 电气工程与系统科学 2024-05-21 Nicolás Gaggion , Candelaria Mosquera , Lucas Mansilla , Julia Mariel Saidman , Martina Aineseder , Diego H. Milone , Enzo Ferrante

Traditional video-induced physiological datasets usually rely on whole-trial labels, which introduce temporal label noise in dynamic emotion recognition. We present FIRMED, a peak-centered multimodal dataset based on an immediate-recall…

人机交互 · 计算机科学 2026-04-01 Hao Tang , Songyun Xie , Xinzhou Xie , Can Liao , Bohan Li , Zhongyu Tian , Dalu Zheng

With the widespread application of artificial intelligence (AI), particularly deep learning (DL) and vision large language models (VLLMs), in skin disease diagnosis, the need for interpretability becomes crucial. However, existing…

计算机视觉与模式识别 · 计算机科学 2025-11-11 Yuhao Shen , Liyuan Sun , Yan Xu , Wenbin Liu , Shuping Zhang , Shawn Afvari , Zhongyi Han , Jiaoyan Song , Yongzhi Ji , Tao Lu , Xiaonan He , Xin Gao , Juexiao Zhou

Few-shot learning has recently attracted wide interest in image classification, but almost all the current public benchmarks are focused on natural images. The few-shot paradigm is highly relevant in medical-imaging applications due to the…

计算机视觉与模式识别 · 计算机科学 2022-06-02 Fereshteh Shakeri , Malik Boudiaf , Sina Mohammadi , Ivaxi Sheth , Mohammad Havaei , Ismail Ben Ayed , Samira Ebrahimi Kahou

Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However, existing agricultural multimodal benchmarks mainly evaluate final-answer correctness…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Zi Ye , Yibin Wen , Xiaoya Fan , Xinyu Zhang , Jing Wu , Kun Zeng , Zurong Mai , Jiarui Zhang , Bohan Shi , Juepeng Zheng , Jianxi Huang , Yutong Lu , Haohuan Fu

We introduce MedMNIST v2, a large-scale MNIST-like dataset collection of standardized biomedical images, including 12 datasets for 2D and 6 datasets for 3D. All images are pre-processed into a small size of 28x28 (2D) or 28x28x28 (3D) with…

计算机视觉与模式识别 · 计算机科学 2023-02-20 Jiancheng Yang , Rui Shi , Donglai Wei , Zequan Liu , Lin Zhao , Bilian Ke , Hanspeter Pfister , Bingbing Ni

Face recognition in images is an active area of interest among the computer vision researchers. However, recognizing human face in an unconstrained environment, is a relatively less-explored area of research. Multiple face recognition in…

计算机视觉与模式识别 · 计算机科学 2019-03-29 Shiv Ram Dubey , Snehasis Mukherjee

Visually cataloging and quantifying the natural world requires pushing the boundaries of both detailed visual classification and counting at scale. Despite significant progress, particularly in crowd and traffic analysis, the fine-grained,…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Jinyu Xu , Tianqi Hu , Xiaonan Hu , Letian Zhou , Songliang Cao , Meng Zhang , Hao Lu

Fine-grained classification of microscopic image data with limited samples is an open problem in computer vision and biomedical imaging. Deep learning based vision systems mostly deal with high number of low-resolution images, whereas…

计算机视觉与模式识别 · 计算机科学 2020-10-07 Mengran Fan , Tapabrata Chakrabort , Eric I-Chao Chang , Yan Xu , Jens Rittscher

While artificial intelligence (AI) holds promise for supporting healthcare providers and improving the accuracy of medical diagnoses, a lack of transparency in the composition of datasets exposes AI models to the possibility of…

计算机视觉与模式识别 · 计算机科学 2022-07-08 Matthew Groh , Caleb Harris , Roxana Daneshjou , Omar Badri , Arash Koochek

Forests are vital to ecosystems, supporting biodiversity and essential services, but are rapidly changing due to land use and climate change. Understanding and mitigating negative effects requires parsing data on forests at global scale…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Nikolaos Ioannis Bountos , Arthur Ouaknine , Ioannis Papoutsis , David Rolnick

Phytoplankton are a crucial component of aquatic ecosystems, and effective monitoring of them can provide valuable insights into ocean environments and ecosystem changes. Traditional phytoplankton monitoring methods are often complex and…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Yang Yu , Yuezun Li , Xin Sun , Junyu Dong

Artificial intelligence (AI) is vital in ophthalmology, tackling tasks like diagnosis, classification, and visual question answering (VQA). However, existing AI models in this domain often require extensive annotation and are task-specific,…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Danli Shi , Weiyi Zhang , Xiaolan Chen , Yexin Liu , Jiancheng Yang , Siyu Huang , Yih Chung Tham , Yingfeng Zheng , Mingguang He

Fundus images are essential for the early screening and detection of eye diseases. While deep learning models using fundus images have significantly advanced the diagnosis of multiple eye diseases, variations in images from different…

计算机视觉与模式识别 · 计算机科学 2025-11-10 Qian Zeng , Le Zhang , Yipeng Liu , Ce Zhu , Fan Zhang

We present MaterialFigBench, a benchmark dataset designed to evaluate the ability of multimodal large language models (LLMs) to solve university-level materials science problems that require accurate interpretation of figures. Unlike…

计算与语言 · 计算机科学 2026-03-13 Michiko Yoshitake , Yuta Suzuki , Ryo Igarashi , Yoshitaka Ushiku , Keisuke Nagato