English
Related papers

Related papers: CPPE-5: Medical Personal Protective Equipment Data…

200 papers

Human pose estimation (HPE) is a key building block for developing AI-based context-aware systems inside the operating room (OR). The 24/7 use of images coming from cameras mounted on the OR ceiling can however raise concerns for privacy,…

Computer Vision and Pattern Recognition · Computer Science 2021-08-23 Vinkle Srivastav , Afshin Gangi , Nicolas Padoy

Object recognition is among the fundamental tasks in the computer vision applications, paving the path for all other image understanding operations. In every stage of progress in object recognition research, efforts have been made to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-31 Aria Salari , Abtin Djavadifar , Xiangrui Liu , Homayoun Najjaran

We propose a new model for detecting visual relationships, such as "person riding motorcycle" or "bottle on table". This task is an important step towards comprehensive structured image understanding, going beyond detecting individual…

Computer Vision and Pattern Recognition · Computer Science 2019-05-03 Alexander Kolesnikov , Alina Kuznetsova , Christoph H. Lampert , Vittorio Ferrari

A hallmark of the deep learning era for computer vision is the successful use of large-scale labeled datasets to train feature representations for tasks ranging from object recognition and semantic segmentation to optical flow estimation…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Stefan Stojanov , Anh Thai , Zixuan Huang , James M. Rehg

Many applications of machine learning, such as human health research, involve processing private or sensitive information. Privacy concerns may impose significant hurdles to collaboration in scenarios where there are multiple sites holding…

Machine Learning · Computer Science 2021-02-24 Hafiz Imtiaz , Jafar Mohammadi , Anand D. Sarwate

Deep object recognition models have been very successful over benchmark datasets such as ImageNet. How accurate and robust are they to distribution shifts arising from natural and synthetic variations in datasets? Prior research on this…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Ali Borji

We present MVMO (Multi-View, Multi-Object dataset): a synthetic dataset of 116,000 scenes containing randomly placed objects of 10 distinct classes and captured from 25 camera locations in the upper hemisphere. MVMO comprises…

Computer Vision and Pattern Recognition · Computer Science 2022-06-01 Aitor Alvarez-Gila , Joost van de Weijer , Yaxing Wang , Estibaliz Garrote

Pose-based anomaly detection is a video-analysis technique for detecting anomalous events or behaviors by examining human pose extracted from the video frames. Utilizing pose data alleviates privacy and ethical issues. Also,…

Computer Vision and Pattern Recognition · Computer Science 2023-03-10 Ghazal Alinezhad Noghre , Armin Danesh Pazho , Vinit Katariya , Hamed Tabkhi

Remote photoplethysmography (rPPG) is an attractive method for noninvasive, convenient and concomitant measurement of physiological vital signals. Public benchmark datasets have served a valuable role in the development of this technology…

Computer Vision and Pattern Recognition · Computer Science 2023-05-02 Jiankai Tang , Kequan Chen , Yuntao Wang , Yuanchun Shi , Shwetak Patel , Daniel McDuff , Xin Liu

Automated medical image analysis systems often require large amounts of training data with high quality labels, which are difficult and time consuming to generate. This paper introduces Radiology Object in COntext version 2 (ROCOv2), a…

We present a large scale data set, OpenEDS: Open Eye Dataset, of eye-images captured using a virtual-reality (VR) head mounted display mounted with two synchronized eyefacing cameras at a frame rate of 200 Hz under controlled illumination.…

Computer Vision and Pattern Recognition · Computer Science 2019-05-20 Stephan J. Garbin , Yiru Shen , Immo Schuetz , Robert Cavin , Gregory Hughes , Sachin S. Talathi

Pulmonary embolism (PE) is a life-threatening condition where rapid and accurate diagnosis is imperative yet difficult due to predominantly atypical symptomatology. Computed tomography pulmonary angiography (CTPA) is acknowledged as the…

Image and Video Processing · Electrical Eng. & Systems 2024-07-17 Bizhe Bai , Yan-Jie Zhou , Yujian Hu , Tony C. W. Mok , Yilang Xiang , Le Lu , Hongkun Zhang , Minfeng Xu

When it comes to classifying child sexual abuse images, managing similar inter-class correlations and diverse intra-class correlations poses a significant challenge. Vision transformer models, unlike conventional deep convolutional network…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Hanxian He , Campbell Wilson , Thanh Thi Nguyen , Janis Dalins

The emerging area of computational pathology (CPath) is ripe ground for the application of deep learning (DL) methods to healthcare due to the sheer volume of raw pixel data in whole-slide images (WSIs) of cancerous tissue slides. However,…

Image and Video Processing · Electrical Eng. & Systems 2020-04-23 Jevgenij Gamper , Navid Alemi Koohbanani , Ksenija Benes , Simon Graham , Mostafa Jahanifar , Syed Ali Khurram , Ayesha Azam , Katherine Hewitt , Nasir Rajpoot

In today's Human-Robot Interaction (HRI) scenarios, a prevailing tendency exists to assume that the robot shall cooperate with the closest individual or that the scene involves merely a singular human actor. However, in realistic scenarios,…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Federico Rollo , Andrea Zunino , Nikolaos Tsagarakis , Enrico Mingo Hoffman , Arash Ajoudani

The Remote Sensing Copy-Move Question Answering (RSCMQA) task focuses on interpreting complex tampering scenarios and inferring the relationships between objects. Currently, publicly available datasets often use randomly generated tampered…

Multimedia · Computer Science 2025-06-27 Ze Zhang , Enyuan Zhao , Yi Jiang , Jie Nie , Xinyue Liang

Medical image segmentation remains challenging due to limited annotations for training, ambiguous anatomical features, and domain shifts. While vision-language models such as CLIP offer strong cross-modal representations, their potential…

Computer Vision and Pattern Recognition · Computer Science 2026-02-25 Taha Koleilat , Hojat Asgariandehkordi , Omid Nejati Manzari , Berardino Barile , Yiming Xiao , Hassan Rivaz

Most of existing category-level object pose estimation methods devote to learning the object category information from point cloud modality. However, the scale of 3D datasets is limited due to the high cost of 3D data collection and…

Computer Vision and Pattern Recognition · Computer Science 2024-05-07 Xiao Lin , Minghao Zhu , Ronghao Dang , Guangliang Zhou , Shaolong Shu , Feng Lin , Chengju Liu , Qijun Chen

A fundamental characteristic common to both human vision and natural language is their compositional nature. Yet, despite the performance gains contributed by large vision and language pretraining, we find that: across 7 architectures…

Computation and Language · Computer Science 2023-05-17 Zixian Ma , Jerry Hong , Mustafa Omer Gul , Mona Gandhi , Irena Gao , Ranjay Krishna

The existing face recognition datasets usually lack occlusion samples, which hinders the development of face recognition. Especially during the COVID-19 coronavirus epidemic, wearing a mask has become an effective means of preventing the…

Computer Vision and Pattern Recognition · Computer Science 2021-03-05 Baojin Huang , Zhongyuan Wang , Guangcheng Wang , Kui Jiang , Kangli Zeng , Zhen Han , Xin Tian , Yuhong Yang
‹ Prev 1 4 5 6 7 8 10 Next ›