English
Related papers

Related papers: GRAPE: Generalizable and Robust Multi-view Facial …

200 papers

Recently, appearance-based gaze estimation has been attracting attention in computer vision, and remarkable improvements have been achieved using various deep learning techniques. Despite such progress, most methods aim to infer gaze…

Computer Vision and Pattern Recognition · Computer Science 2024-01-26 Suneung Kim , Woo-Jeoung Nam , Seong-Whan Lee

In this paper, we present a sparsity-aware deep network for automatic 4D facial expression recognition (FER). Given 4D data, we first propose a novel augmentation method to combat the data limitation problem for deep learning. This is…

Computer Vision and Pattern Recognition · Computer Science 2020-08-20 Muzammil Behzad , Nhat Vo , Xiaobai Li , Guoying Zhao

Robotic vision plays a key role for perceiving the environment in grasping applications. However, the conventional framed-based robotic vision, suffering from motion blur and low sampling rate, may not meet the automation needs of evolving…

DeepFakes are synthetic videos generated by swapping a face of an original image with the face of somebody else. In this paper, we describe our work to develop general, deep learning-based models to classify DeepFake content. We propose a…

Computer Vision and Pattern Recognition · Computer Science 2022-03-02 Pratikkumar Prajapati , Chris Pollett

Learning image representations to capture fine-grained semantics has been a challenging and important task enabling many applications such as image search and clustering. In this paper, we present Graph-Regularized Image Semantic Embedding…

Computer Vision and Pattern Recognition · Computer Science 2019-03-01 Da-Cheng Juan , Chun-Ta Lu , Zhen Li , Futang Peng , Aleksei Timofeev , Yi-Ting Chen , Yaxi Gao , Tom Duerig , Andrew Tomkins , Sujith Ravi

This paper aims to learn a compact representation of a video for video face recognition task. We make the following contributions: first, we propose a meta attention-based aggregation scheme which adaptively and fine-grained weighs the…

Computer Vision and Pattern Recognition · Computer Science 2019-09-13 Zhaoxiang Liu , Huan Hu , Jinqiang Bai , Shaohua Li , Shiguo Lian

Efficient and robust grasp pose detection is vital for robotic manipulation. For general 6 DoF grasping, conventional methods treat all points in a scene equally and usually adopt uniform sampling to select grasp candidates. However, we…

Robotics · Computer Science 2024-06-18 Chenxi Wang , Hao-Shu Fang , Minghao Gou , Hongjie Fang , Jin Gao , Cewu Lu

Vision-based models for robotic grasping automate critical, repetitive, and draining industrial tasks. Existing approaches are typically limited in two ways: they either target a single gripper and are potentially applied on costly dual-arm…

Class-incremental learning is a challenging problem, where the goal is to train a model that can classify data from an increasing number of classes over time. With the advancement of vision-language pre-trained models such as CLIP, they…

Computer Vision and Pattern Recognition · Computer Science 2024-07-22 Linlan Huang , Xusheng Cao , Haori Lu , Xialei Liu

Despite recent advances in facial recognition, there remains a fundamental issue concerning degradations in performance due to substantial perspective (pose) differences between enrollment and query (probe) imagery. Therefore, we propose a…

Computer Vision and Pattern Recognition · Computer Science 2025-05-15 J. Brennan Peace , Shuowen Hu , Benjamin S. Riggan

Deep learning methods have been achieved brilliant results in face recognition. One of the important tasks to improve the performance is to collect and label images as many as possible. However, labeling identities and checking qualities of…

Computer Vision and Pattern Recognition · Computer Science 2023-05-15 Myung-cheol Roh , Pyoung-gang Lim , Jongju Shin

A common architectural choice for deep metric learning is a convolutional neural network followed by global average pooling (GAP). Albeit simple, GAP is a highly effective way to aggregate information. One possible explanation for the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-23 Yeti Z. Gurbuz , Ozan Sener , A. Aydın Alatan

A series of region-based methods succeed in extracting regional features and enhancing grasp detection quality. However, faced with a cluttered scene with potential collision, the definition of the grasp-relevant region stays inconsistent,…

Robotics · Computer Science 2024-11-15 Siang Chen , Pengwei Xie , Wei Tang , Dingchang Hu , Yixiang Dai , Guijin Wang

Learning robotic grasps from visual observations is a promising yet challenging task. Recent research shows its great potential by preparing and learning from large-scale synthetic datasets. For the popular, 6 degree-of-freedom (6-DOF)…

Computer Vision and Pattern Recognition · Computer Science 2020-09-29 Chaozheng Wu , Jian Chen , Qiaoyu Cao , Jianchi Zhang , Yunxin Tai , Lin Sun , Kui Jia

Recognising remote sensing scene images remains challenging due to large visual-semantic discrepancies. These mainly arise due to the lack of detailed annotations that can be employed to align pixel-level representations with high-level…

Computer Vision and Pattern Recognition · Computer Science 2020-04-10 S. Wang , Y. Guan , L. Shao

The image matching field has been witnessing a continuous emergence of novel learnable feature matching techniques, with ever-improving performance on conventional benchmarks. However, our investigation shows that despite these gains, their…

Computer Vision and Pattern Recognition · Computer Science 2024-05-22 Hanwen Jiang , Arjun Karpur , Bingyi Cao , Qixing Huang , Andre Araujo

The task of 2D animal pose estimation plays a crucial role in advancing deep learning applications in animal behavior analysis and ecological research. Despite notable progress in some existing approaches, our study reveals that the…

Computer Vision and Pattern Recognition · Computer Science 2025-04-02 Lei Wang , Yujie Zhong , Xiaopeng Sun , Jingchun Cheng , Chengjian Feng , Qiong Cao , Lin Ma , Zhaoxin Fan

The semantic representation of deep features is essential for image context understanding, and effective fusion of features with different semantic representations can significantly improve the model's performance on salient object…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Han Sun , Jun Cen , Ningzhong Liu , Dong Liang , Huiyu Zhou

Cross modal face matching between the thermal and visible spectrum is a much desired capability for night-time surveillance and security applications. Due to a very large modality gap, thermal-to-visible face recognition is one of the most…

Computer Vision and Pattern Recognition · Computer Science 2016-08-01 M. Saquib Sarfraz , Rainer Stiefelhagen

Cross modal face matching between the thermal and visible spectrum is a much de- sired capability for night-time surveillance and security applications. Due to a very large modality gap, thermal-to-visible face recognition is one of the…

Computer Vision and Pattern Recognition · Computer Science 2015-07-13 M. Saquib Sarfraz , Rainer Stiefelhagen