中文
相关论文

相关论文: MiVOLO: Multi-input Transformer for Age and Gender…

200 篇论文

The ability to accurately recognize an individual's face with respect to human aging factor holds significant importance for various private as well as government sectors such as customs and public security bureaus, passport office, and…

计算机视觉与模式识别 · 计算机科学 2024-10-23 Wang Yao , Muhammad Ali Farooq , Joseph Lemley , Peter Corcoran

To minimize the impact of age variation on face recognition, age-invariant face recognition (AIFR) extracts identity-related discriminative features by minimizing the correlation between identity- and age-related features while face age…

计算机视觉与模式识别 · 计算机科学 2022-11-08 Zhizhong Huang , Junping Zhang , Hongming Shan

To minimize the effects of age variation in face recognition, previous work either extracts identity-related discriminative features by minimizing the correlation between identity- and age-related features, called age-invariant face…

计算机视觉与模式识别 · 计算机科学 2021-03-04 Zhizhong Huang , Junping Zhang , Hongming Shan

Early detection of developmental disorders can be aided by analyzing infant craniofacial morphology, but modeling infant faces is challenging due to limited data and frequent spontaneous expressions. We introduce BabyFlow, a generative AI…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Antonia Alomar , Mireia Masias , Marius George Linguraru , Federico M. Sukno , Gemma Piella

Gender is one of the most common attributes used to describe an individual. It is used in multiple domains such as human computer interaction, marketing, security, and demographic reports. Research has been performed to automate the task of…

计算机视觉与模式识别 · 计算机科学 2018-05-22 Maneet Singh , Shruti Nagpal , Richa Singh , Mayank Vatsa

Face recognition in unconstrained environments such as surveillance, video, and web imagery must contend with extreme variation in pose, blur, illumination, and occlusion, where conventional visual quality metrics fail to predict whether…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Allen Tu , Kartik Narayan , Joshua Gleason , Jennifer Xu , Matthew Meyn , Tom Goldstein , Vishal M. Patel

Identification using biometrics is an important yet challenging task. Abundant research has been conducted on identifying personal identity or gender using given signals. Various types of biometrics such as electrocardiogram (ECG),…

计算机视觉与模式识别 · 计算机科学 2019-09-13 Hyoung-Kyu Song , Ebrahim AlAlkeem , Jaewoong Yun , Tae-Ho Kim , Tae-Ho Kim , Hyerin Yoo , Dasom Heo , Chan Yeob Yeun , Myungsu Chae

General-purpose Large Vision-Language Models (LVLMs), despite their massive scale, often falter in dermatology due to "diffuse attention" - the inability to disentangle subtle pathological lesions from background noise. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2026-01-15 Lijun Liu , Linwei Chen , Zhishou Zhang , Meng Tian , Hengfu Cui , Ruiyang Li , Zhaocheng Liu , Qiang Ju , Qianxi Li , Hong-Yu Zhou

While leveraging abundant human videos and simulated robot data poses a scalable solution to the scarcity of real-world robot data, the generalization capability of existing vision-language-action models (VLAs) remains limited by mismatches…

Recent advances in image-to-video (I2V) generation have achieved remarkable progress in synthesizing high-quality, temporally coherent videos from static images. Among all the applications of I2V, human-centric video generation includes a…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Liao Shen , Wentao Jiang , Yiran Zhu , Jiahe Li , Tiezheng Ge , Zhiguo Cao , Bo Zheng

Recognizing speaking in humans is a central task towards understanding social interactions. Ideally, speaking would be detected from individual voice recordings, as done previously for meeting scenarios. However, individual voice recordings…

计算机视觉与模式识别 · 计算机科学 2024-03-05 Jose Vargas Quiros , Chirag Raman , Stephanie Tan , Ekin Gedik , Laura Cabrera-Quiros , Hayley Hung

Facial Emotion Recognition is an inherently difficult problem, due to vast differences in facial structures of individuals and ambiguity in the emotion displayed by a person. Recently, a lot of work is being done in the field of Facial…

计算机视觉与模式识别 · 计算机科学 2021-10-29 Aakash Saroop , Pathik Ghugare , Sashank Mathamsetty , Vaibhav Vasani

This paper presents a detailed study of improving visual representations for vision language (VL) tasks and develops an improved object detection model to provide object-centric representations of images. Compared to the most widely used…

计算机视觉与模式识别 · 计算机科学 2021-03-11 Pengchuan Zhang , Xiujun Li , Xiaowei Hu , Jianwei Yang , Lei Zhang , Lijuan Wang , Yejin Choi , Jianfeng Gao

Predicting if a person is an adult or a minor has several applications such as inspecting underage driving, preventing purchase of alcohol and tobacco by minors, and granting restricted access. The challenging nature of this problem arises…

计算机视觉与模式识别 · 计算机科学 2018-03-21 Maneet Singh , Shruti Nagpal , Mayank Vatsa , Richa Singh

Video understanding requires reasoning at multiple spatiotemporal resolutions -- from short fine-grained motions to events taking place over longer durations. Although transformer architectures have recently advanced the state-of-the-art,…

计算机视觉与模式识别 · 计算机科学 2022-06-01 Shen Yan , Xuehan Xiong , Anurag Arnab , Zhichao Lu , Mi Zhang , Chen Sun , Cordelia Schmid

The last decade or two has witnessed a boom of images. With the increasing ubiquity of cameras and with the advent of selfies, the number of facial images available in the world has skyrocketed. Consequently, there has been a growing…

计算机视觉与模式识别 · 计算机科学 2021-10-26 Vikas Sheoran , Shreyansh Joshi , Tanisha R. Bhayani

We present VINO, a unified visual generator that performs image and video generation and editing within a single framework. Instead of relying on task-specific models or independent modules for each modality, VINO uses a shared diffusion…

计算机视觉与模式识别 · 计算机科学 2026-01-19 Junyi Chen , Tong He , Zhoujie Fu , Pengfei Wan , Kun Gai , Weicai Ye

Estimating motion from images is a well-studied problem in computer vision and robotics. Previous work has developed techniques to estimate the motion of a moving camera in a largely static environment (e.g., visual odometry) and to segment…

机器人学 · 计算机科学 2019-03-01 Kevin M. Judd , Jonathan D. Gammell , Paul Newman

In this paper, we propose a novel age estimation method based on GLOH feature descriptor and multi-task learning (MTL). The GLOH feature descriptor, one of the state-of-the-art feature descriptor, is used to capture the age-related local…

计算机视觉与模式识别 · 计算机科学 2011-05-09 Yixiong Liang , Lingbo Liu , Ying Xu , Yao Xiang , Beiji Zou

Recent progress in face restoration has shifted from visual fidelity to identity fidelity, driving a transition from reference-free to reference-based paradigms that condition restoration on reference images of the same person. However,…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Teer Song , Yue Zhang , Yu Tian , Ziyang Wang , Xianlin Zhang , Guixuan Zhang , Xuan Liu , Xueming Li , Yasen Zhang