English
Related papers

Related papers: CT-ScanGaze: A Dataset and Baselines for 3D Volume…

200 papers

3D volume segmentation is a fundamental task in many scientific and medical applications. Producing accurate segmentations efficiently is challenging, in part due to low imaging data quality (e.g., noise and low image resolution) and…

Human-Computer Interaction · Computer Science 2020-04-08 Anahita Sanandaji , Cindy Grimm , Ruth West , Max Parola , Meghan Kajihara , Kathryn Hays , Luke Hillard , Brandon Lane , Molly Beyer

We present a new dataset and benchmark with the goal of advancing research in the intersection of brain activities and eye movements. Our dataset, EEGEyeNet, consists of simultaneous Electroencephalography (EEG) and Eye-tracking (ET)…

Signal Processing · Electrical Eng. & Systems 2021-11-11 Ard Kastrati , Martyna Beata Płomecka , Damián Pascual , Lukas Wolf , Victor Gillioz , Roger Wattenhofer , Nicolas Langer

Gaze estimation methods commonly use facial appearances to predict the direction of a person gaze. However, previous studies show three major challenges with convolutional neural network (CNN)-based, transformer-based, and contrastive…

Computer Vision and Pattern Recognition · Computer Science 2026-05-04 Xinyuan Zhao , Yihang Wu , Ahmad Chaddad , Sarah A. Alkhodair , Reem Kateb

In medical image visualization, path tracing of volumetric medical data like CT scans produces lifelike three-dimensional visualizations. Immersive VR displays can further enhance the understanding of complex anatomies. Going beyond the…

Graphics · Computer Science 2026-01-30 Constantin Kleinbeck , Hannah Schieber , Klaus Engel , Ralf Gutjahr , Daniel Roth

Objective: While recent advances in text-conditioned generative models have enabled the synthesis of realistic medical images, progress has been largely confined to 2D modalities such as chest X-rays. Extending text-to-image generation to…

Computer Vision and Pattern Recognition · Computer Science 2025-10-02 Daniele Molino , Camillo Maria Caruso , Filippo Ruffini , Paolo Soda , Valerio Guarrasi

We present 3DeepCT, a deep neural network for computed tomography, which performs 3D reconstruction of scattering volumes from multi-view images. Our architecture is dictated by the stationary nature of atmospheric cloud fields. The task of…

Image and Video Processing · Electrical Eng. & Systems 2020-12-14 Yael Sde-Chen , Yoav Y. Schechner , Vadim Holodovsky , Eshkol Eytan

The task of predicting 3D eye gaze from eye images can be performed either by (a) end-to-end learning for image-to-gaze mapping or by (b) fitting a 3D eye model onto images. The former case requires 3D gaze labels, while the latter requires…

Computer Vision and Pattern Recognition · Computer Science 2023-11-22 Nikola Popovic , Dimitrios Christodoulou , Danda Pani Paudel , Xi Wang , Luc Van Gool

Radiologic diagnostic errors-under-reading errors, inattentional blindness, and communication failures-remain prevalent in clinical practice. These issues often stem from missed localized abnormalities, limited global context, and…

Computer Vision and Pattern Recognition · Computer Science 2025-09-05 Yuheng Li , Yenho Chen , Yuxiang Lai , Jike Zhong , Vanessa Wildman , Xiaofeng Yang

3D shape is a crucial but heavily underutilized cue in today's computer vision systems, mostly due to the lack of a good generic shape representation. With the recent availability of inexpensive 2.5D depth sensors (e.g. Microsoft Kinect),…

Computer Vision and Pattern Recognition · Computer Science 2015-04-16 Zhirong Wu , Shuran Song , Aditya Khosla , Fisher Yu , Linguang Zhang , Xiaoou Tang , Jianxiong Xiao

We introduce GazeD, a new 3D gaze estimation method that jointly provides 3D gaze and human pose from a single RGB image. Leveraging the ability of diffusion models to deal with uncertainty, it generates multiple plausible 3D gaze and pose…

Computer Vision and Pattern Recognition · Computer Science 2026-01-26 Riccardo Catalini , Davide Di Nucci , Guido Borghi , Davide Davoli , Lorenzo Garattoni , Gianpiero Francesca , Yuki Kawana , Roberto Vezzani

We introduce SPECTRE, a fully transformer-based foundation model for volumetric computed tomography (CT). Our Self-Supervised & Cross-Modal Pretraining for CT Representation Extraction (SPECTRE) approach utilizes scalable 3D Vision…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Cris Claessens , Christiaan Viviers , Giacomo D'Amicantonio , Egor Bondarev , Fons van der Sommen

CT report generation (CTRG) aims to automatically generate diagnostic reports for 3D volumes, relieving clinicians' workload and improving patient care. Despite clinical value, existing works fail to effectively incorporate diagnostic…

Computer Vision and Pattern Recognition · Computer Science 2025-06-27 Xiwei Deng , Xianchun He , Jianfeng Bao , Yudan Zhou , Shuhui Cai , Congbo Cai , Zhong Chen

Eye gaze estimation and simultaneous semantic understanding of a user through eye images is a crucial component in Virtual and Mixed Reality; enabling energy efficient rendering, multi-focal displays and effective interaction with 3D…

Computer Vision and Pattern Recognition · Computer Science 2019-08-27 Zhengyang Wu , Srivignesh Rajendran , Tarrence van As , Joelle Zimmermann , Vijay Badrinarayanan , Andrew Rabinovich

Cancers identified in CT scans are usually accompanied by detailed radiology reports, but publicly available CT datasets often lack these essential reports. This absence limits their usefulness for developing accurate report generation AI.…

Image and Video Processing · Electrical Eng. & Systems 2025-08-20 Pedro R. A. S. Bassi , Mehmet Can Yavuz , Kang Wang , Xiaoxi Chen , Wenxuan Li , Sergio Decherchi , Andrea Cavalli , Yang Yang , Alan Yuille , Zongwei Zhou

Enabling robots to understand human gaze target is a crucial step to allow capabilities in downstream tasks, for example, attention estimation and movement anticipation in real-world human-robot interactions. Prior works have addressed the…

Computer Vision and Pattern Recognition · Computer Science 2025-07-02 Zhuangzhuang Dai , Vincent Gbouna Zakka , Luis J. Manso , Chen Li

High-fidelity gaze redirection is critical for generating augmented data to improve the generalization of gaze estimators. 3D Gaussian Splatting (3DGS) models like GazeGaussian represent the state-of-the-art but can struggle with rendering…

Computer Vision and Pattern Recognition · Computer Science 2025-11-17 Abiram Panchalingam , Indu Bodala , Stuart Middleton

Convolutional Neural Networks (CNNs) have been recently employed to solve problems from both the computer vision and medical image analysis fields. Despite their popularity, most approaches are only able to process 2D images while most…

Computer Vision and Pattern Recognition · Computer Science 2016-06-16 Fausto Milletari , Nassir Navab , Seyed-Ahmad Ahmadi

Efficient and accurate multi-organ segmentation from abdominal CT volumes is a fundamental challenge in medical image analysis. Existing 3D segmentation approaches are computationally and memory intensive, often processing entire volumes…

Image and Video Processing · Electrical Eng. & Systems 2025-05-19 Hania Ghouse , Muzammil Behzad

Accurate delineation of anatomical structures in volumetric CT scans is crucial for diagnosis and treatment planning. While AI has advanced automated segmentation, current approaches typically target individual structures, creating a…

The rapid increase of computed tomography (CT) scans and their time-consuming manual analysis have created an urgent need for robust automated analysis techniques in clinical settings. These aim to assist radiologists and help them managing…

Image and Video Processing · Electrical Eng. & Systems 2026-02-24 Theo Di Piazza , Carole Lazarus , Olivier Nempont , Loic Boussel
‹ Prev 1 3 4 5 6 7 10 Next ›