中文
相关论文

相关论文: Robust LSTM-Autoencoders for Face De-Occlusion in …

200 篇论文

Plenty of effective methods have been proposed for face recognition during the past decade. Although these methods differ essentially in many aspects, a common practice of them is to specifically align the facial area based on the prior…

计算机视觉与模式识别 · 计算机科学 2017-08-02 Yuanyi Zhong , Jiansheng Chen , Bo Huang

Accurate lane detection is essential for effective path planning and lane following in autonomous driving, especially in scenarios with significant occlusion from vehicles and pedestrians. Existing models often struggle under such…

计算机视觉与模式识别 · 计算机科学 2024-08-20 Aayush Agrawal , Ashmitha Jaysi Sivakumar , Ibrahim Kaif , Chayan Banerjee

Very low-resolution face recognition (VLRFR) poses unique challenges, such as tiny regions of interest and poor resolution due to extreme standoff distance or wide viewing angle of the acquisition devices. In this paper, we study principled…

计算机视觉与模式识别 · 计算机科学 2023-04-21 Jacky Chen Long Chai , Tiong-Sik Ng , Cheng-Yaw Low , Jaewoo Park , Andrew Beng Jin Teoh

Lensless cameras, innovatively replacing traditional lenses for ultra-thin, flat optics, encode light directly onto sensors, producing images that are not immediately recognizable. This compact, lightweight, and cost-effective imaging…

计算机视觉与模式识别 · 计算机科学 2024-06-07 Xin Cai , Hailong Zhang , Chenchen Wang , Wentao Liu , Jinwei Gu , Tianfan Xue

Person re-identification is vital for monitoring and tracking crowd movement to enhance public security. However, re-identification in the presence of occlusion substantially reduces the performance of existing systems and is a challenging…

计算机视觉与模式识别 · 计算机科学 2023-04-18 Prathistith Raj Medi , Ghanta Sai Krishna , Praneeth Nemani , Satyanarayana Vollala , Santosh Kumar

Vision language models (VLMs) demonstrate impressive capabilities in visual question answering and image captioning, acting as a crucial link between visual and language models. However, existing open-source VLMs heavily rely on pretrained…

计算机视觉与模式识别 · 计算机科学 2024-07-24 Aristeidis Panos , Rahaf Aljundi , Daniel Olmeda Reino , Richard E Turner

Face anti-spoofing (FAS) is crucial for protecting facial recognition systems from presentation attacks. Previous methods approached this task as a classification problem, lacking interpretability and reasoning behind the predicted results.…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Hongyang Wang , Yichen Shi , Zhuofu Tao , Yuhao Gao , Liepiao Zhang , Xun Lin , Jun Feng , Xiaochen Yuan , Zitong Yu , Xiaochun Cao

Occlusion poses a major challenge for person re-identification (ReID). Existing approaches typically rely on outside tools to infer visible body parts, which may be suboptimal in terms of both computational efficiency and ReID accuracy. In…

计算机视觉与模式识别 · 计算机科学 2022-01-04 Pengfei Wang , Changxing Ding , Zhiyin Shao , Zhibin Hong , Shengli Zhang , Dacheng Tao

Facial Expression Recognition is an active area of research in computer vision with a wide range of applications. Several approaches have been developed to solve this problem for different benchmark datasets. However, Facial Expression…

计算机视觉与模式识别 · 计算机科学 2016-01-12 Abubakrelsedik Karali , Ahmad Bassiouny , Motaz El-Saban

In this paper, we tackle the challenge of face recognition in the wild, where images often suffer from low quality and real-world distortions. Traditional heuristic approaches-either training models directly on these degraded images or…

计算机视觉与模式识别 · 计算机科学 2024-04-05 Yunhao Liu , Yu-Ju Tsai , Kelvin C. K. Chan , Xiangtai Li , Lu Qi , Ming-Hsuan Yang

This paper presents an easy and efficient face detection and face recognition approach using free software components from the internet. Face detection and face recognition problems have wide applications in home and office security.…

计算机视觉与模式识别 · 计算机科学 2019-01-23 Hira Ahmad

Place recognition is a critical and challenging task for mobile robots, aiming to retrieve an image captured at the same place as a query image from a database. Existing methods tend to fail while robots move autonomously under occlusion…

计算机视觉与模式识别 · 计算机科学 2023-02-21 Yue Chen , Xingyu Chen , Yicen Li

High-quality facial appearance capture has traditionally required costly studio recording. Recent works consider an in-the-wild smartphone-based setup; however, their model-based inverse rendering paradigm struggles with the complex…

计算机视觉与模式识别 · 计算机科学 2026-05-08 Yuxuan Han , Xin Ming , Tianxiao Li , Zhuofan Shen , Qixuan Zhang , Lan Xu , Feng Xu

Super-resolution (SR) and landmark localization of tiny faces are highly correlated tasks. On the one hand, landmark localization could obtain higher accuracy with faces of high-resolution (HR). On the other hand, face SR would benefit from…

计算机视觉与模式识别 · 计算机科学 2019-11-21 Yu Yin , Joseph P. Robinson , Yulun Zhang , Yun Fu

Traditionally, video conferencing is a widely adopted solution for telecommunication, but a lack of immersiveness comes inherently due to the 2D nature of facial representation. The integration of Virtual Reality (VR) in a…

计算机视觉与模式识别 · 计算机科学 2021-12-03 Surabhi Gupta , Ashwath Shetty , Avinash Sharma

Ophthalmic images may contain identical-looking pathologies that can cause failure in automated techniques to distinguish different retinal degenerative diseases. Additionally, reliance on large annotated datasets and lack of knowledge…

图像与视频处理 · 电气工程与系统科学 2022-08-02 Sharif Amit Kamran , Khondker Fariha Hossain , Alireza Tavakkoli , Stewart Lee Zuckerbrod , Salah A. Baker

Multimodal large language models (MLLMs) have recently become a focal point of research due to their formidable multimodal understanding capabilities. For example, in the audio and speech domains, an LLM can be equipped with (automatic)…

计算机视觉与模式识别 · 计算机科学 2025-03-10 Umberto Cappellazzo , Minsu Kim , Honglie Chen , Pingchuan Ma , Stavros Petridis , Daniele Falavigna , Alessio Brutti , Maja Pantic

Robust face reconstruction from monocular image in general lighting conditions is challenging. Methods combining deep neural network encoders with differentiable rendering have opened up the path for very fast monocular reconstruction of…

计算机视觉与模式识别 · 计算机科学 2021-11-23 Abdallah Dib , Cedric Thebault , Junghyun Ahn , Philippe-Henri Gosselin , Christian Theobalt , Louis Chevallier

In unsupervised anomaly detection (UAD) research, while state-of-the-art models have reached a saturation point with extensive studies on public benchmark datasets, they adopt large-scale tailor-made neural networks (NN) for detection…

计算机视觉与模式识别 · 计算机科学 2024-07-08 YeongHyeon Park , Sungho Kang , Myung Jin Kim , Hyeong Seok Kim , Juneho Yi

This paper reveals that large language models (LLMs), despite being trained solely on textual data, are surprisingly strong encoders for purely visual tasks in the absence of language. Even more intriguingly, this can be achieved by a…

计算机视觉与模式识别 · 计算机科学 2024-05-07 Ziqi Pang , Ziyang Xie , Yunze Man , Yu-Xiong Wang