English
Related papers

Related papers: FERA: A Pose-Based Framework for Rule-Grounded Mul…

200 papers

The question-answering (QA) capabilities of foundation models are highly sensitive to prompt variations, rendering their performance susceptible to superficial, non-meaning-altering changes. This vulnerability often stems from the model's…

Machine Learning · Computer Science 2024-06-07 Dyah Adila , Shuai Zhang , Boran Han , Yuyang Wang

Safety filters in commercial text-to-image (T2I) models systematically block legitimate artistic content involving the human figure, treating classical nude photography with the same restrictiveness as explicit material. While prior…

Multimedia · Computer Science 2026-03-24 Luca Cazzaniga

Federated learning (FL) remains highly vulnerable to adaptive backdoor attacks that preserve stealth by closely imitating benign update statistics. Existing defenses predominantly rely on anomaly detection in parameter or gradient space,…

Machine Learning · Computer Science 2026-02-13 Chibueze Peace Obioma , Youcheng Sun , Mustafa A. Mustafa

Face Emotion Recognition (FER) is essential for social interactions and understanding others' mental states. Utilizing eye tracking to investigate FER has yielded insights into cognitive processes. In this study, we utilized an…

Human-Computer Interaction · Computer Science 2025-03-21 Meisam J. Seikavandi , Maria J. Barrett , Paolo Burelli

Diffusion models have achieved remarkable success in generative modeling, yet how to effectively adapt large pretrained models to new tasks remains challenging. We revisit the reconstruction behavior of diffusion models during denoising to…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Bo Yin , Xiaobin Hu , Xingyu Zhou , Peng-Tao Jiang , Yue Liao , Junwei Zhu , Jiangning Zhang , Ying Tai , Chengjie Wang , Shuicheng Yan

Facial expression recognition (FER) must remain robust under both cultural variation and perceptually degraded visual conditions, yet most existing evaluations assume homogeneous data and high-quality imagery. We introduce an agent-based,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-16 David Freire-Obregón , José Salas-Cáceres , Javier Lorenzo-Navarro , Oliverio J. Santana , Daniel Hernández-Sosa , Modesto Castrillón-Santana

Rule-based models are essential for high-stakes decision-making due to their transparency and interpretability, but their discrete nature creates challenges for optimization and scalability. In this work, we present the Fuzzy Rule-based…

Machine Learning · Computer Science 2025-09-25 Javier Fumanal-Idocin , Raquel Fernandez-Peralta , Javier Andreu-Perez

Fair representation learning (FRL) is a popular class of methods aiming to produce fair classifiers via data preprocessing. Recent regulatory directives stress the need for FRL methods that provide practical certificates, i.e., provable…

Machine Learning · Computer Science 2023-06-09 Nikola Jovanović , Mislav Balunović , Dimitar I. Dimitrov , Martin Vechev

We present ASTRA (A} Scene-aware TRAnsformer-based model for trajectory prediction), a light-weight pedestrian trajectory forecasting model that integrates the scene context, spatial dynamics, social inter-agent interactions and temporal…

Computer Vision and Pattern Recognition · Computer Science 2025-01-20 Izzeddin Teeti , Aniket Thomas , Munish Monga , Sachin Kumar , Uddeshya Singh , Andrew Bradley , Biplab Banerjee , Fabio Cuzzolin

In smart manufacturing environments, accurate and real-time recognition of worker actions is essential for productivity, safety, and human-machine collaboration. While skeleton-based human activity recognition (HAR) offers robustness to…

Computer Vision and Pattern Recognition · Computer Science 2025-08-21 Vinit Hegiste , Vidit Goyal , Tatjana Legler , Martin Ruskowski

Over the past decade, the technology used by referees in football has improved substantially, enhancing the fairness and accuracy of decisions. This progress has culminated in the implementation of the Video Assistant Referee (VAR), an…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Jan Held , Anthony Cioppa , Silvio Giancola , Abdullah Hamdi , Christel Devue , Bernard Ghanem , Marc Van Droogenbroeck

Multi-view facial expression recognition (FER) is a challenging task because the appearance of an expression varies in poses. To alleviate the influences of poses, recent methods either perform pose normalization or learn separate FER…

Computer Vision and Pattern Recognition · Computer Science 2019-05-27 Yuanyuan Liu , Jiyao Peng , Jiabei Zeng , Shiguang Shan

Accurate localization serves as an important component in autonomous driving systems. Traditional rule-based localization involves many standalone modules, which is theoretically fragile and requires costly hyperparameter tuning, therefore…

Autonomous racing has rapidly gained research attention. Traditionally, racing cars rely on 2D LiDAR as their primary visual system. In this work, we explore the integration of an event camera with the existing system to provide enhanced…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Zhuyun Zhou , Zongwei Wu , Florian Bolli , Rémi Boutteau , Fan Yang , Radu Timofte , Dominique Ginhac , Tobi Delbruck

Radars and cameras are mature, cost-effective, and robust sensors and have been widely used in the perception stack of mass-produced autonomous driving systems. Due to their complementary properties, outputs from radar detection (radar…

Computer Vision and Pattern Recognition · Computer Science 2021-06-21 Xu Dong , Binnan Zhuang , Yunxiang Mao , Langechuan Liu

Parameter-efficient fine-tuning(PEFT) has largely focused on LoRA and its accuracy-oriented variants, leaving the original goal of reducing trainable parameters has receivedcomparatively little attention. We introduce FoRA, which revisits…

Computation and Language · Computer Science 2026-05-29 Juneyoung Park , Seongbae Lee , Han-Sang Lee , Kyuho Lee , Minjae Kim , Seungheon Hyeon , Kiduk Kwon , Seongwan Kim , Jaeho Lee

Sign language plays a crucial role in bridging communication gaps between the deaf and hard-of-hearing communities. However, existing sign language video generation models often rely on complex intermediate representations, which limits…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Liuzhou Zhang , Zeyu Zhang , Biao Wu , Luyao Tang , Zirui Song , Hongyang He , Renda Han , Guangzhen Yao , Huacan Wang , Ronghao Chen , Xiuying Chen , Guan Huang , Zheng Zhu

Humans (and many vertebrates) face the problem of fusing together multiple fixations of a scene in order to obtain a representation of the whole, where each fixation uses a high-resolution fovea and decreasing resolution in the periphery.…

Computer Vision and Pattern Recognition · Computer Science 2025-11-20 Christopher K. I. Williams

This paper presents DeRA, a novel 1D video tokenizer that decouples the spatial-temporal representation learning in video tokenization to achieve better training efficiency and performance. Specifically, DeRA maintains a compact 1D latent…

Computer Vision and Pattern Recognition · Computer Science 2025-12-05 Pengbo Guo , Junke Wang , Zhen Xing , Chengxu Liu , Daoguo Dong , Xueming Qian , Zuxuan Wu

Despite decades of work, surveillance still struggles to find specific targets across long, multi-camera video. Prior methods -- tracking pipelines, CLIP based models, and VideoRAG -- require heavy manual filtering, capture only shallow…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Hyojin Park , Yi Li , Janghoon Cho , Sungha Choi , Jungsoo Lee , Taotao Jing , Shuai Zhang , Munawar Hayat , Dashan Gao , Ning Bi , Fatih Porikli