English
Related papers

Related papers: SLYKLatent: A Learning Framework for Gaze Estimati…

200 papers

Zero-Shot Learning (ZSL) is achieved via aligning the semantic relationships between the global image feature vector and the corresponding class semantic descriptions. However, using the global features to represent fine-grained images may…

Computer Vision and Pattern Recognition · Computer Science 2018-05-22 Yunlong Yu , Zhong Ji , Yanwei Fu , Jichang Guo , Yanwei Pang , Zhongfei Zhang

The purpose of this paper is the detection of salient areas in natural video by using the new deep learning techniques. Salient patches in video frames are predicted first. Then the predicted visual fixation maps are built upon them. We…

Computer Vision and Pattern Recognition · Computer Science 2016-04-28 Souad Chaabouni , Jenny Benois-Pineau , Ofer Hadar , Chokri Ben Amar

In this letter, we propose a new method, Multi-Clue Gaze (MCGaze), to facilitate video gaze estimation via capturing spatial-temporal interaction context among head, face, and eye in an end-to-end learning way, which has not been well…

Computer Vision and Pattern Recognition · Computer Science 2024-01-01 Yiran Guan , Zhuoguang Chen , Wenzheng Zeng , Zhiguo Cao , Yang Xiao

Face images are subject to many different factors of variation, especially in unconstrained in-the-wild scenarios. For most tasks involving such images, e.g. expression recognition from video streams, having enough labeled data is…

Computer Vision and Pattern Recognition · Computer Science 2020-08-19 Marah Halawa , Manuel Wöllhaf , Eduardo Vellasques , Urko Sánchez Sanz , Olaf Hellwich

Event camera, a novel neuromorphic vision sensor, records data with high temporal resolution and wide dynamic range, offering new possibilities for accurate visual representation in challenging scenarios. However, event data is inherently…

Computer Vision and Pattern Recognition · Computer Science 2025-08-08 Lin Zhu , Ruonan Liu , Xiao Wang , Lizhi Wang , Hua Huang

Lip-reading models have been significantly improved recently thanks to powerful deep learning architectures. However, most works focused on frontal or near frontal views of the mouth. As a consequence, lip-reading performance seriously…

Computer Vision and Pattern Recognition · Computer Science 2019-11-15 Shiyang Cheng , Pingchuan Ma , Georgios Tzimiropoulos , Stavros Petridis , Adrian Bulat , Jie Shen , Maja Pantic

This paper explores how deep learning techniques can improve visual-based SLAM performance in challenging environments. By combining deep feature extraction and deep matching methods, we introduce a versatile hybrid visual SLAM system…

Robotics · Computer Science 2024-06-05 Zhang Xiao , Shuaixin Li

Environmental perception systems are crucial for high-precision mapping and autonomous navigation, with LiDAR serving as a core sensor providing accurate 3D point cloud data. Efficiently processing unstructured point clouds while extracting…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Chuang Chen , Yi Lin , Bo Wang , Jing Hu , Xi Wu , Wenyi Ge

Data size is the bottleneck for developing deep saliency models, because collecting eye-movement data is very time consuming and expensive. Most of current studies on human attention and saliency modeling have used high quality stereotype…

Computer Vision and Pattern Recognition · Computer Science 2019-11-20 Zhaohui Che , Ali Borji , Guangtao Zhai , Xiongkuo Min , Guodong Guo , Patrick Le Callet

Numerous ongoing and future large area surveys (e.g. DES, EUCLID, LSST, WFIRST), will increase by several orders of magnitude the volume of data that can be exploited for galaxy morphology studies. The full potential of these surveys can…

Astrophysics of Galaxies · Physics 2018-01-31 D. Tuccillo , M. Huertas-Company , E. Decencière , S. Velasco-Forero , H. Domínguez Sánchez , P. Dimauro

Current multimodal large language models (MLLMs) cannot effectively utilize eye-gaze information for video understanding, even when gaze cues are supplied via visual overlays or text descriptions. We introduce GazeQwen, a parameter…

Computer Vision and Pattern Recognition · Computer Science 2026-03-30 Trong Thang Pham , Hien Nguyen , Ngan Le

Deriving an effective facial expression recognition component is important for a successful human-computer interaction system. Nonetheless, recognizing facial expression remains a challenging task. This paper describes a novel approach…

Computer Vision and Pattern Recognition · Computer Science 2019-04-15 Mundher Al-Shabi , Wooi Ping Cheah , Tee Connie

Deep neural networks have demonstrated superior performance on appearance-based gaze estimation tasks. However, due to variations in person, illuminations, and background, performance degrades dramatically when applying the model to a new…

Computer Vision and Pattern Recognition · Computer Science 2022-10-06 Ruicong Liu , Yiwei Bao , Mingjie Xu , Haofei Wang , Yunfei Liu , Feng Lu

Facial expression recognition is a key task in human-computer interaction and affective computing. However, acquiring a large amount of labeled facial expression data is often costly. Therefore, it is particularly important to design a…

Computer Vision and Pattern Recognition · Computer Science 2026-01-12 Zhongpeng Cai , Jun Yu , Wei Xu , Tianyu Liu , Jianqing Sun , Jiaen Liang

In this paper, we propose a framework for disentangling the appearance and geometry representations in the face recognition task. To provide supervision for this aim, we generate geometrically identical faces by incorporating spatial…

Computer Vision and Pattern Recognition · Computer Science 2020-01-15 Ali Dabouei , Fariborz Taherkhani , Sobhan Soleymani , Jeremy Dawson , Nasser M. Nasrabadi

Many scientific prediction problems have spatiotemporal data- and modeling-related challenges in handling complex variations in space and time using only sparse and unevenly distributed observations. This paper presents a novel deep…

Machine Learning · Computer Science 2021-12-13 Yijun Lin , Yao-Yi Chiang , Meredith Franklin , Sandrah P. Eckel , José Luis Ambite

Self-supervised learning for depth estimation uses geometry in image sequences for supervision and shows promising results. Like many computer vision tasks, depth network performance is determined by the capability to learn accurate spatial…

Computer Vision and Pattern Recognition · Computer Science 2021-11-22 Hang Zhou , David Greenwood , Sarah Taylor

Reliable facial expression learning (FEL) involves the effective learning of distinctive facial expression characteristics for more reliable, unbiased and accurate predictions in real-life settings. However, current systems struggle with…

Computer Vision and Pattern Recognition · Computer Science 2024-12-10 Azmine Toushik Wasi , Taki Hasan Rafi , Raima Islam , Karlo Serbetar , Dong Kyu Chae

We propose a novel neural pipeline, MSGazeNet, that learns gaze representations by taking advantage of the eye anatomy information through a multistream framework. Our proposed solution comprises two components, first a network for…

Computer Vision and Pattern Recognition · Computer Science 2024-03-04 Zunayed Mahmud , Paul Hungler , Ali Etemad

Eye gaze, encompassing fixations and saccades, provides critical insights into human intentions and future actions. This study introduces a gaze-regularized framework that enhances Vision Language Models (VLMs) for egocentric behavior…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Anupam Pani , Yanchao Yang
‹ Prev 1 4 5 6 7 8 10 Next ›