中文
相关论文

相关论文: Visuo-Locomotive Complexity as a Component of Para…

200 篇论文

We aim to build complex humanoid agents that integrate perception, motor control, and memory. In this work, we partly factor this problem into low-level motor control from proprioception and high-level coordination of the low-level skills…

人工智能 · 计算机科学 2019-01-16 Josh Merel , Arun Ahuja , Vu Pham , Saran Tunyasuvunakool , Siqi Liu , Dhruva Tirumala , Nicolas Heess , Greg Wayne

Composite visualization represents a widely embraced design that combines multiple visual representations to create an integrated view. However, the traditional approach of creating composite visualizations in immersive environments…

人机交互 · 计算机科学 2024-08-08 Qian Zhu , Tao Lu , Shunan Guo , Xiaojuan Ma , Yalong Yang

This paper presents a framework to navigate visually impaired people through unfamiliar environments by means of a mobile manipulator. The Human-Robot system consists of three key components: a mobile base, a robotic arm, and the human…

The visuomotor system of any animal is critical for its survival, and the development of a complex one within humans is large factor in our success as a species on Earth. This system is an essential part of our ability to adapt to our…

机器学习 · 计算机科学 2021-03-29 Tamir Blum , Gabin Paillet , Watcharawut Masawat , Mickael Laine , Kazuya Yoshida

Designing and building visual analytics (VA) systems is a complex, iterative process that requires the seamless integration of data processing, analytics capabilities, and visualization techniques. While prior research has extensively…

人机交互 · 计算机科学 2025-08-12 Leonardo Ferreira , Gustavo Moreira , Fabio Miranda

In this paper, we introduce VisioPath, a novel framework combining vision-language models (VLMs) with model predictive control (MPC) to enable safe autonomous driving in dynamic traffic environments. The proposed approach leverages a…

系统与控制 · 电气工程与系统科学 2025-07-10 Shanting Wang , Panagiotis Typaldos , Chenjun Li , Andreas A. Malikopoulos

Vision Language Models (VLMs) play a crucial role in robotic manipulation by enabling robots to understand and interpret the visual properties of objects and their surroundings, allowing them to perform manipulation based on this multimodal…

机器人学 · 计算机科学 2025-05-21 Nurhan Bulus Guran , Hanchi Ren , Jingjing Deng , Xianghua Xie

The navigation of complex labyrinths with tens of rooms under visual partially observable state is typically addressed using recurrent deep reinforcement learning architectures. In this work, we show that navigation can be achieved through…

神经与进化计算 · 计算机科学 2024-04-11 Caleidgh Bayer , Robert J. Smith , Malcolm I. Heywood

Recent advances in autonomous vehicle (AV) behavior planning have shown impressive social interaction capabilities when interacting with other road users. However, achieving human-like prediction and decision-making in interactions with…

机器人学 · 计算机科学 2025-07-29 Meiting Dang , Yanping Wu , Yafei Wang , Dezong Zhao , David Flynn , Chongfeng Wei

Walking and cycling, commonly referred to as active travel, have become integral components of modern transport planning. Recently, there has been growing recognition of the substantial role that active travel can play in making cities more…

物理与社会 · 物理学 2025-05-06 Ivann Schlosser , Valentina Marín Maureira , Richard Milton , Elsa Arcaute , Michael Batty

A crucial ability of mobile intelligent agents is to integrate the evidence from multiple sensory inputs in an environment and to make a sequence of actions to reach their goals. In this paper, we attempt to approach the problem of…

计算机视觉与模式识别 · 计算机科学 2020-03-10 Chuang Gan , Yiwei Zhang , Jiajun Wu , Boqing Gong , Joshua B. Tenenbaum

The advent of cyber-physical systems, such as robots and autonomous vehicles (AVs), brings new opportunities and challenges for the domain of interaction design. Though there is consensus about the value of human-centred development, there…

We propose VLM-Social-Nav, a novel Vision-Language Model (VLM) based navigation approach to compute a robot's motion in human-centered environments. Our goal is to make real-time decisions on robot actions that are socially compliant with…

机器人学 · 计算机科学 2024-11-27 Daeun Song , Jing Liang , Amirreza Payandeh , Amir Hossain Raj , Xuesu Xiao , Dinesh Manocha

This paper explores the principles for transforming a quadrupedal robot into a guide robot for individuals with visual impairments. A guide robot has great potential to resolve the limited availability of guide animals that are accessible…

机器人学 · 计算机科学 2024-04-08 J. Taery Kim , Wenhao Yu , Yash Kothari , Jie Tan , Greg Turk , Sehoon Ha

How can a robot navigate successfully in rich and diverse environments, indoors or outdoors, along office corridors or trails on the grassland, on the flat ground or the staircase? To this end, this work aims to address three challenges:…

机器人学 · 计算机科学 2022-06-01 Bo Ai , Wei Gao , Vinay , David Hsu

We present a novel high-level planning framework that leverages vision-language models (VLMs) to improve autonomous navigation in unknown indoor environments with many dead ends. Traditional exploration methods often take inefficient routes…

机器人学 · 计算机科学 2025-10-14 D. Schwartz , K. Kondo , J. P. How

Multiple-view visualization (MV) has been used for visual analytics in various fields (e.g., bioinformatics, cybersecurity, and intelligence analysis). Because each view encodes data from a particular perspective, analysts often use a set…

人机交互 · 计算机科学 2022-07-18 Abdul Rahman Shaikh , David Koop , Hamed Alhoori , Maoyuan Sun

Contact-based decision and planning methods are becoming increasingly important to endow higher levels of autonomy for legged robots. Formal synthesis methods derived from symbolic systems have great potential for reasoning about high-level…

机器人学 · 计算机科学 2022-01-04 Ye Zhao , Yinan Li , Luis Sentis , Ufuk Topcu , Jun Liu

This study focuses on Embodied Complex-Question Answering task, which means the embodied robot need to understand human questions with intricate structures and abstract semantics. The core of this task lies in making appropriate plans based…

机器人学 · 计算机科学 2025-04-02 Ning Lan , Baoshan Ou , Xuemei Xie , Guangming Shi

Interpretable communication is essential for safe and trustworthy autonomous driving, yet current vision-language models (VLMs) often operate under idealized assumptions and struggle to capture user intent in real-world scenarios. Existing…

计算机视觉与模式识别 · 计算机科学 2025-07-02 Djamahl Etchegaray , Yuxia Fu , Zi Huang , Yadan Luo