中文
相关论文

相关论文: VPD-100K: Towards Generalizable and Fine-grained V…

200 篇论文

Remote user verification in Internet-based applications is becoming increasingly important nowadays. A popular scenario for it consists of submitting a picture of the user's Identity Document (ID) to a service platform, authenticating its…

密码学与安全 · 计算机科学 2025-08-29 Javier Muñoz-Haro , Ruben Tolosana , Julian Fierrez , Ruben Vera-Rodriguez , Aythami Morales

Detecting 3D objects keypoints is of great interest to the areas of both graphics and computer vision. There have been several 2D and 3D keypoint datasets aiming to address this problem in a data-driven way. These datasets, however, either…

计算机视觉与模式识别 · 计算机科学 2020-08-10 Yang You , Yujing Lou , Chengkun Li , Zhoujun Cheng , Liangwei Li , Lizhuang Ma , Weiming Wang , Cewu Lu

Data-driven robotic manipulation learning depends on large-scale, high-quality expert demonstration datasets. However, existing datasets, which primarily rely on human teleoperated robot collection, are limited in terms of scalability,…

The deployment of facial recognition systems has created an ethical dilemma: achieving high accuracy requires massive datasets of real faces collected without consent, leading to dataset retractions and potential legal liabilities under…

Open data sets that contain personal information are susceptible to adversarial attacks even when anonymized. By performing low-cost joins on multiple datasets with shared attributes, malicious users of open data portals might get access to…

密码学与安全 · 计算机科学 2022-11-30 Kaustav Bhattacharjee , Akm Islam , Jaideep Vaidya , Aritra Dasgupta

Video anomaly detection (VAD) without human monitoring is a complex computer vision task that can have a positive impact on society if implemented successfully. While recent advances have made significant progress in solving this task, most…

计算机视觉与模式识别 · 计算机科学 2023-08-23 Joseph Fioresi , Ishan Rajendrakumar Dave , Mubarak Shah

While fine-grained object recognition is an important problem in computer vision, current models are unlikely to accurately classify objects in the wild. These fully supervised models need additional annotated images to classify objects in…

计算机视觉与模式识别 · 计算机科学 2017-09-11 Timnit Gebru , Judy Hoffman , Li Fei-Fei

Visual localization is the task of estimating the camera pose of an image relative to a scene representation. In practice, visual localization systems are often cloud-based. Naturally, this raises privacy concerns in terms of revealing…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Vojtech Panek , Patrik Beliansky , Zuzana Kukelova , Torsten Sattler

In this paper, we introduce ScenePilot-4K, a large-scale first-person dataset for safety-aware vision-language learning and evaluation in autonomous driving. Built from public online driving videos, ScenePilot-4K contains 3,847 hours of…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Yujin Wang , Yutong Zheng , Wenxian Fan , Tianyi Wang , Hongqing Chu , Li Zhang , Bingzhao Gao , Daxin Tian , Jianqiang Wang , Hong Chen

Video Anomaly Detection (VAD) aims to automatically analyze spatiotemporal patterns in surveillance videos collected from open spaces to detect anomalous events that may cause harm, such as fighting, stealing, and car accidents. However,…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Yang Liu , Siao Liu , Xiaoguang Zhu , Jielin Li , Hao Yang , Liangyu Teng , Juncen Guo , Yan Wang , Dingkang Yang , Jing Liu

We present the first systematic study on concealed object detection (COD), which aims to identify objects that are "perfectly" embedded in their background. The high intrinsic similarities between the concealed objects and their background…

计算机视觉与模式识别 · 计算机科学 2024-02-21 Deng-Ping Fan , Ge-Peng Ji , Ming-Ming Cheng , Ling Shao

Lane detection plays a key role in autonomous driving. While car cameras always take streaming videos on the way, current lane detection works mainly focus on individual images (frames) by ignoring dynamics along the video. In this work, we…

计算机视觉与模式识别 · 计算机科学 2021-08-20 Yujun Zhang , Lei Zhu , Wei Feng , Huazhu Fu , Mingqian Wang , Qingxia Li , Cheng Li , Song Wang

In recent years, differential privacy has seen significant advancements in image classification; however, its application to video activity recognition remains under-explored. This paper addresses the challenges of applying differential…

计算机视觉与模式识别 · 计算机科学 2023-06-29 Zelun Luo , Yuliang Zou , Yijin Yang , Zane Durante , De-An Huang , Zhiding Yu , Chaowei Xiao , Li Fei-Fei , Animashree Anandkumar

We present High-Density Visual Particle Dynamics (HD-VPD), a learned world model that can emulate the physical dynamics of real scenes by processing massive latent point clouds containing 100K+ particles. To enable efficiency at this scale,…

机器学习 · 计算机科学 2024-07-01 William F. Whitney , Jacob Varley , Deepali Jain , Krzysztof Choromanski , Sumeet Singh , Vikas Sindhwani

Techniques to deliver privacy-preserving synthetic datasets take a sensitive dataset as input and produce a similar dataset as output while maintaining differential privacy. These approaches have the potential to improve data sharing and…

数据库 · 计算机科学 2018-08-24 Luke Rodriguez , Bill Howe

Visual Prompt Tuning (VPT) has emerged as a parameter-efficient fine-tuning paradigm for vision transformers, with conventional approaches utilizing dataset-level prompts that remain the same across all input instances. We observe that this…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Xi Xiao , Yunbei Zhang , Xingjian Li , Tianyang Wang , Xiao Wang , Yuxiang Wei , Jihun Hamm , Min Xu

Advances in radiance fields have enabled photorealistic novel view synthesis. In several domains, large-scale real-world datasets have been developed to support comprehensive benchmarking and to facilitate progress beyond scene-specific…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Cheng-You Lu , Yi-Shan Hung , Wei-Ling Chi , Hao-Ping Wang , Charlie Li-Ting Tsai , Yu-Cheng Chang , Yu-Lun Liu , Thomas Do , Chin-Teng Lin

In this paper, we introduce VCSL (Video Copy Segment Localization), a new comprehensive segment-level annotated video copy dataset. Compared with existing copy detection datasets restricted by either video-level annotation or small-scale,…

计算机视觉与模式识别 · 计算机科学 2022-06-17 Sifeng He , Xudong Yang , Chen Jiang , Gang Liang , Wei Zhang , Tan Pan , Qing Wang , Furong Xu , Chunguang Li , Jingxiong Liu , Hui Xu , Kaiming Huang , Yuan Cheng , Feng Qian , Xiaobo Zhang , Lei Yang

Securing personal identity against deepfake attacks is increasingly critical in the digital age, especially for celebrities and political figures whose faces are easily accessible and frequently targeted. Most existing deepfake detection…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Kaiqing Lin , Zhiyuan Yan , Ke-Yue Zhang , Li Hao , Yue Zhou , Yuzhen Lin , Weixiang Li , Taiping Yao , Shouhong Ding , Bin Li

Visual commonsense plays a vital role in understanding and reasoning about the visual world. While commonsense knowledge bases like ConceptNet provide structured collections of general facts, they lack visually grounded representations.…

计算机视觉与模式识别 · 计算机科学 2025-06-06 Xiangqing Shen , Fanfan Wang , Siwei Wu , Rui Xia