中文
相关论文

相关论文: TacEva: A Performance Evaluation Framework For Vis…

200 篇论文

Evaluating employee performance in organizations with varying workloads and tasks is challenging. Specifically, it is important to understand how quantitative measurements of employee achievements relate to supervisor expectations, what the…

Sensor visibility is crucial for safety-critical applications in automotive, robotics, smart infrastructure and others: In addition to object detection and occupancy mapping, visibility describes where a sensor can potentially measure or is…

计算机视觉与模式识别 · 计算机科学 2022-11-14 Joachim Börger , Marc Patrick Zapf , Marat Kopytjuk , Xinrun Li 2 , Claudius Gläser

Vision Transformers (ViTs) excel in computer vision tasks but lack flexibility for edge devices' diverse needs. A vital issue is that ViTs pre-trained to cover a broad range of tasks are \textit{over-qualified} for edge devices that usually…

计算机视觉与模式识别 · 计算机科学 2025-04-07 Ziteng Wei , Qiang He , Bing Li , Feifei Chen , Yun Yang

Traffic Surveillance Systems (TSS) have become increasingly crucial in modern intelligent transportation systems, with vision technologies playing a central role for scene perception and understanding. While existing surveys typically focus…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Wei Zhou , Li Yang , Lei Zhao , Runyu Zhang , Yifan Cui , Hongpu Huang , Kun Qie , Chen Wang

Robotic in-hand manipulation requires reliable object-motion tracking under frequent visual occlusion, yet low-texture visuotactile images provide few stable correspondences for conventional image- or geometry-matching methods. This paper…

机器人学 · 计算机科学 2026-05-28 Zhongyuan Liao , Junzhe Wang , Qingyang Liu , Zhenmin Huang , Jun Ma , Yi Cai , Fei Meng , Haobo Liang , Michael Yu Wang

Being intensively studied, visual tracking has seen great recent advances in either speed (e.g., with correlation filters) or accuracy (e.g., with deep features). Real-time and high accuracy tracking algorithms, however, remain scarce. In…

计算机视觉与模式识别 · 计算机科学 2017-08-02 Heng Fan , Haibin Ling

Tactile sensing offers rich and complementary information to vision and language, enabling robots to perceive fine-grained object properties. However, existing tactile sensors lack standardization, leading to redundant features that hinder…

机器人学 · 计算机科学 2026-02-03 Yiyun Zhou , Mingjing Xu , Jingwei Shi , Quanjiang Li , Jingyuan Chen

While vision-language-action models (VLAs) have shown promising robotic behaviors across a diverse set of manipulation tasks, they achieve limited success rates when deployed on novel tasks out of the box. To allow these policies to safely…

机器人学 · 计算机科学 2025-10-31 Qiao Gu , Yuanliang Ju , Shengxiang Sun , Igor Gilitschenski , Haruki Nishimura , Masha Itkina , Florian Shkurti

Tactile sensing is critical in advanced interactive systems by emulating the human sense of touch to detect stimuli. Vision-based tactile sensors are promising for providing multimodal capabilities and high robustness, yet existing…

机器人学 · 计算机科学 2025-04-07 Mayue Shi , Yongqi Zhang , Xiaotong Guo , Eric M. Yeatman

Deformable object manipulation is a classical and challenging research area in robotics. Compared with rigid object manipulation, this problem is more complex due to the deformation properties including elastic, plastic, and elastoplastic…

机器人学 · 计算机科学 2024-05-14 Jianhua Shan , Yuhao Sun , Shixin Zhang , Fuchun Sun , Zixi Chen , Zirong Shen , Cesare Stefanini , Yiyong Yang , Shan Luo , Bin Fang

Recent advances in text-to-video (T2V) technology, as demonstrated by models such as Runway Gen-3, Pika, Sora, and Kling, have significantly broadened the applicability and popularity of the technology. This progress has created a growing…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Zelu Qi , Ping Shi , Shuqi Wang , Chaoyang Zhang , Fei Zhao , Zefeng Ying , Da Pan , Xi Yang , Zheqi He , Teng Dai

Time-to-Collision (TTC) estimation lies in the core of the forward collision warning (FCW) functionality, which is key to all Automatic Emergency Braking (AEB) systems. Although the success of solutions using frame-based cameras (e.g.,…

机器人学 · 计算机科学 2025-04-24 Kaizhen Sun , Jinghang Li , Kuan Dai , Bangyan Liao , Wei Xiong , Yi Zhou

Existing text-to-video (T2V) evaluation benchmarks, such as VBench and EvalCrafter, suffer from two limitations. (i) While the emphasis is on subject-centric prompts or static camera scenes, camera motion essential for producing cinematic…

计算机视觉与模式识别 · 计算机科学 2025-10-10 Nithin C. Babu , Aniruddha Mahapatra , Harsh Rangwani , Rajiv Soundararajan , Kuldeep Kulkarni

Elastomers are central to vision-based tactile sensors (VBTSs), where they transduce external contact into observable deformation. Different VBTS architectures, however, require distinct optical and mechanical properties, particularly…

系统与控制 · 电气工程与系统科学 2026-04-14 Wen Fan , Dandan Zhang

The problem of visual tracking evaluation is sporting a large variety of performance measures, and largely suffers from lack of consensus about which measures should be used in experiments. This makes the cross-paper tracker comparison…

计算机视觉与模式识别 · 计算机科学 2016-03-08 Luka Čehovin , Aleš Leonardis , Matej Kristan

Autonomous vehicles (AVs) rely on multi-modal fusion for safety, but current visual and optical sensors fail to detect road-induced excitations which are critical for vehicles' dynamic control. Inspired by human synesthesia, we propose the…

人工智能 · 计算机科学 2026-02-03 Rui Wang , Yaoguang Cao , Yuyi Chen , Jianyi Xu , Zhuoyang Li , Jiachen Shang , Shichun Yang

Due to the complexity of modeling the elastic properties of materials, the use of machine learning algorithms is continuously increasing for tactile sensing applications. Recent advances in deep neural networks applied to computer vision…

机器人学 · 计算机科学 2020-06-05 Carmelo Sferrazza , Raffaello D'Andrea

Autoregressive (AR) models have recently shown strong performance in image generation, where a critical component is the visual tokenizer (VT) that maps continuous pixel inputs to discrete token sequences. The quality of the VT largely…

计算机视觉与模式识别 · 计算机科学 2025-05-20 Huawei Lin , Tong Geng , Zhaozhuo Xu , Weijie Zhao

We propose TacFiLM, a lightweight modality-fusion approach that integrates visual-tactile signals into vision-language-action (VLA) models. While recent advances in VLA models have introduced robot policies that are both generalizable and…

Large vision language models (VLMs) combine large language models with vision encoders, demonstrating promise across various tasks. However, they often underperform in task-specific applications due to domain gaps between pre-training and…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Yang Bai , Yang Zhou , Jun Zhou , Rick Siow Mong Goh , Daniel Shu Wei Ting , Yong Liu