中文
相关论文

相关论文: Learning Oculomotor Behaviors from Scanpath

200 篇论文

Predicting pedestrian crossing intentions is crucial for the navigation of mobile robots and intelligent vehicles. Although recent deep learning-based models have shown significant success in forecasting intentions, few consider incomplete…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Yu Liu , Zhijie Liu , Zedong Yang , You-Fu Li , He Kong

Semantic segmentation and activity classification are key components to creating intelligent surgical systems able to understand and assist clinical workflow. In the Operating Room, semantic segmentation is at the core of creating robots…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Idris Hamoud , Alexandros Karargyris , Aidean Sharghi , Omid Mohareri , Nicolas Padoy

Understanding people's actions and interactions typically depends on seeing them. Automating the process of action recognition from visual data has been the topic of much research in the computer vision community. But what if it is too…

计算机视觉与模式识别 · 计算机科学 2019-09-23 Tianhong Li , Lijie Fan , Mingmin Zhao , Yingcheng Liu , Dina Katabi

Computational neuroscience studies that have examined human visual system through functional magnetic resonance imaging (fMRI) have identified a model where the mammalian brain pursues two distinct pathways (for recognition of biological…

计算机视觉与模式识别 · 计算机科学 2015-09-15 Bardia Yousefi , C. K. Loo

Trajectory Prediction of dynamic objects is a widely studied topic in the field of artificial intelligence. Thanks to a large number of applications like predicting abnormal events, navigation system for the blind, etc. there have been many…

机器学习 · 计算机科学 2017-05-29 Daksh Varshneya , G. Srinivasaraghavan

Wearable collaborative robots stand to assist human wearers who need fall prevention assistance or wear exoskeletons. Such a robot needs to be able to constantly adapt to the surrounding scene based on egocentric vision, and predict the ego…

计算机视觉与模式识别 · 计算机科学 2024-08-08 Weizhuo Wang , C. Karen Liu , Monroe Kennedy

Visual imitation learning enables robotic agents to acquire skills by observing expert demonstration videos. In the one-shot setting, the agent generates a policy after observing a single expert demonstration without additional fine-tuning.…

机器人学 · 计算机科学 2026-01-01 Raktim Gautam Goswami , Prashanth Krishnamurthy , Yann LeCun , Farshad Khorrami

Autonomous vehicles (AVs) must navigate dynamic urban environments where occlusions and perception limitations introduce significant uncertainties. This research builds upon and extends existing approaches in risk-aware motion planning and…

机器人学 · 计算机科学 2025-08-19 Korbinian Moller , Luis Schwarzmeier , Johannes Betz

3D understanding and rendering of moving humans from monocular videos is a challenging task. Despite recent progress, the task remains difficult in real-world scenarios, where obstacles may block the camera view and cause partial occlusions…

计算机视觉与模式识别 · 计算机科学 2023-08-10 Tiange Xiang , Adam Sun , Jiajun Wu , Ehsan Adeli , Li Fei-Fei

While an exciting diversity of new imaging devices is emerging that could dramatically improve robotic perception, the challenges of calibrating and interpreting these cameras have limited their uptake in the robotics community. In this…

机器人学 · 计算机科学 2021-03-23 S. Tejaswi Digumarti , Joseph Daniel , Ahalya Ravendran , Donald G. Dansereau

Zero-shot skeleton-based action recognition aims to classify unseen skeleton-based human actions without prior exposure to such categories during training. This task is extremely challenging due to the difficulty in generalizing from known…

计算机视觉与模式识别 · 计算机科学 2025-07-25 Kai Zhou , Shuhai Zhang , Zeng You , Jinwu Hu , Mingkui Tan , Fei Liu

Object recognition has made great advances in the last decade, but predominately still relies on many high-quality training examples per object category. In contrast, learning new objects from only a few examples could enable many impactful…

3D animation aims to generate a 3D animated video from an input image and a target 3D motion sequence. Recent advances in image-to-3D models enable the creation of animations directly from user-hand drawings. Distinguished from conventional…

图形学 · 计算机科学 2025-08-04 Sunjae Yoon , Gwanhyeong Koo , Younghwan Lee , Ji Woo Hong , Chang D. Yoo

To reach human performance on complex tasks, a key ability for artificial systems is to understand physical interactions between objects, and predict future outcomes of a situation. This ability, often referred to as intuitive physics, has…

计算机视觉与模式识别 · 计算机科学 2020-05-04 Ronan Riochet , Josef Sivic , Ivan Laptev , Emmanuel Dupoux

Given sufficient pairs of resting-state and task-evoked fMRI scans from subjects, it is possible to train ML models to predict subject-specific task-evoked activity using resting-state functional MRI (rsfMRI) scans. However, while rsfMRI…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Minh Nguyen , Gia H. Ngo , Mert R. Sabuncu

We propose an image-classification method to predict the perceived-relevance of text documents from eye-movements. An eye-tracking study was conducted where participants read short news articles, and rated them as relevant or irrelevant for…

人机交互 · 计算机科学 2020-01-16 Nilavra Bhattacharya , Somnath Rakshit , Jacek Gwizdka , Paul Kogut

Autonomous driving has achieved rapid development over the last few decades, including the machine perception as an important issue of it. Although object detection based on conventional cameras has achieved remarkable results in 2D/3D,…

机器人学 · 计算机科学 2021-07-20 Rui Yang , Zhi Yan , Tao Yang , Yassine Ruichek

In this study, we demonstrate a novel self-navigated motion correction method that suppresses eye motion and blinking artifacts on wide-field optical coherence tomographic angiography (OCTA) without requiring any hardware modification.…

医学物理 · 物理学 2020-05-25 Xiang Wei , Tristan T. Hormel , Yukun Guo , Thomas S. Hwang , Yali Jia

We naturally step sideways or lean to see around the obstacle when our view is blocked, and recover a more informative observation. Enabling robots to make the same kind of viewpoint choice is critical for human-centered operations,…

机器人学 · 计算机科学 2026-03-13 Boxun Hu , Chang Chang , Jiawei Ge , Man Namgung , Xiaomin Lin , Axel Krieger , Tinoosh Mohsenin

Open vocabulary object detection (OVD) aims at seeking an optimal object detector capable of recognizing objects from both base and novel categories. Recent advances leverage knowledge distillation to transfer insightful knowledge from…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Jiaming Li , Jiacheng Zhang , Jichang Li , Ge Li , Si Liu , Liang Lin , Guanbin Li