中文
相关论文

相关论文: HMR-1: Hierarchical Massage Robot with Vision-Lang…

200 篇论文

Are current Vision Language Models (VLMs) ready to comprehend and reason about complex embodied interactions in 3D environments? We introduce Embodied3DBench, a robot-centric benchmark targeting low-level spatial intelligence in embodied 3D…

计算机视觉与模式识别 · 计算机科学 2026-05-29 Jiyao Zhang , Mingxu Zhang , Yitong Peng , Haoxuan Liu , Chenshuo Wang , Yuxing Long , Haoyang Huang , Dongjiang Li , Nan Duan , Hui Shen , Hao Dong

Leveraging Multi-modal Large Language Models (MLLMs) to create embodied agents offers a promising avenue for tackling real-world tasks. While language-centric embodied agents have garnered substantial attention, MLLM-based embodied agents…

This paper introduces and overviews a multidisciplinary project aimed at developing responsible and adaptive multi-human multi-robot (MHMR) systems for complex, dynamic settings. The project integrates co-design, ethical frameworks, and…

Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precision. However, autonomous medical robotics has been limited by a fundamental data problem:…

机器人学 · 计算机科学 2026-04-30 Open-H-Embodiment Consortium , : , Nigel Nelson , Juo-Tung Chen , Jesse Haworth , Xinhao Chen , Lukas Zbinden , Dianye Huang , Alaa Eldin Abdelaal , Alberto Arezzo , Ayberk Acar , Farshid Alambeigi , Carlo Alberto Ammirati , Yunke Ao , Pablo David Aranda Rodriguez , Soofiyan Atar , Mattia Ballo , Noah Barnes , Federica Barontini , Filip Binkiewicz , Peter Black , Sebastian Bodenstedt , Leonardo Borgioli , Nikola Budjak , Benjamin Calmé , Fabio Carrillo , Nicola Cavalcanti , Changwei Chen , Haoxin Chen , Sihang Chen , Qihan Chen , Zhongyu Chen , Ziyang Chen , Shing Shin Cheng , Meiqing Cheng , Min Cheng , Zih-Yun Sarah Chiu , Xiangyu Chu , Camilo Correa-Gallego , Giulio Dagnino , Anton Deguet , Jacob Delgado , Jonathan C. DeLong , Kaizhong Deng , Alexander Dimitrakakis , Qingpeng Ding , Hao Ding , Giovanni Distefano , Daniel Donoho , Anqing Duan , Marco Esposito , Shane Farritor , Jad Fayad , Zahi Fayad , Mario Ferradosa , Filippo Filicori , Chelsea Finn , Philipp Fürnstahl , Jiawei Ge , Stamatia Giannarou , Xavier Giralt Ludevid , Frederic Giraud , Aditya Amit Godbole , Ken Goldberg , Antony Goldenberg , Diego Granero Marana , Xiaoqing Guo , Tamás Haidegger , Evan Hailey , Pascal Hansen , Ziyi Hao , Kush Hari , Kengo Hayashi , Jonathon Hawkins , Shelby Haworth , Ortrun Hellig , S. Duke Herrell , Zhouyang Hong , Andrew Howe , Junlei Hu , Zhaoyang Jacopo Hu , Ria Jain , Mohammad Rafiee Javazm , Howard Ji , Rui Ji , Jianmin Ji , Zhongliang Jiang , Dominic Jones , Jeffrey Jopling , Britton Jordan , Ran Ju , Michael Kam , Luoyao Kang , Fausto Kang , Siddhartha Kapuria , Peter Kazanzides , Sonika Kiehler , Ethan Kilmer , Ji Woong Kim , Przemysław Korzeniowski , Chandra Kuchi , Nithesh Kumar , Alan Kuntz , Federico Lavagno , Yu Chung Lee , Hao-Chih Lee , Hang Li , Zhen Li , Xiao Liang , Xinxin Lin , Jinsong Lin , Chang Liu , Fei Liu , Pei Liu , Yun-hui Liu , Wanli Liuchen , Eszter Lukács , Sareena Mann , Miles Mannas , Brett Marinelli , Sabina Martyniak , Francesco Marzola , Lorenzo Mazza , Xueyan Mei , Maria Clara Morais , Luigi Muratore , Chetan Reddy Narayanaswamy , Michał Naskręt , David Navarro-Alarcon , Cyrus Neary , Chi Kit Ng , Christopher Nguan , David Noonan , Ki Hwan Oh , Tom Christian Olesch , Allison M. Okamura , Justin Opfermann , Matteo Pescio , Doan Xuan Viet Pham , Tito Porras , Hongliang Ren , Ariel Rodriguez Jimenez , Ferdinando Rodriguez y Baena , Septimiu E. Salcudean , Asmitha Sathya , Preethi Satish , Lalithkumar Seenivasan , Jiaqi Shao , Yiqing Shen , Yu Sheng , Lucy XiaoYang Shi , Zoe Soulé , Stefanie Speidel , Mingwu Su , Jianhao Su , Idris Sunmola , Kristóf Takács , Yunxi Tang , Patrick Thornycroft , Yu Tian , Jordan Thompson , Mehmet K. Turkcan , Mathias Unberath , Pietro Valdastri , Carlos Vives , Quan Vuong , Martin Wagner , Farong Wang , Wei Wang , Lidian Wang , Chung-Pang Wang , Guankun Wang , Junyi Wang , Erqi Wang , Ziyi Wang , Tanner Watts , Wolfgang Wein , Yimeng Wu , Zijian Wu , Hongjun Wu , Luohong Wu , Jie Ying Wu , Junlin Wu , Victoria Wu , Kaixuan Wu , Mateusz Wójcikowski , Yunye Xiao , Nan Xiao , Wenxuan Xie , Hao Yang , Tianqi Yang , Yinuo Yang , Menglong Ye , Ryan S. Yeung , Nural Yilmaz , Chim Ho Yin , Michael Yip , Rayan Younis , Chenhao Yu , Sayem Nazmuz Zaman , Milos Zefran , Han Zhang , Yuelin Zhang , Yidong Zhang , Yanyong Zhang , Xuyang Zhang , Yameng Zhang , Joyce Zhang , Ning Zhong , Peng Zhou , Haoying Zhou , Xiuli Zuo , Nassir Navab , Mahdi Azizian , Sean D. Huver , Axel Krieger

We propose a novel framework for learning high-level cognitive capabilities in robot manipulation tasks, such as making a smiley face using building blocks. These tasks often involve complex multi-step reasoning, presenting significant…

机器人学 · 计算机科学 2023-05-31 Chuhao Jin , Wenhui Tan , Jiange Yang , Bei Liu , Ruihua Song , Limin Wang , Jianlong Fu

Multimodal Large Language Models (MLLMs) have shown significant advancements, providing a promising future for embodied agents. Existing benchmarks for evaluating MLLMs primarily utilize static images or videos, limiting assessments to…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Zhili Cheng , Yuge Tu , Ran Li , Shiqi Dai , Jinyi Hu , Shengding Hu , Jiahao Li , Yang Shi , Tianyu Yu , Weize Chen , Lei Shi , Maosong Sun

Developing autonomous home robots controlled by natural language has long been a pursuit of humanity. While advancements in large language models (LLMs) and embodied intelligence make this goal closer, several challenges persist: the lack…

机器人学 · 计算机科学 2025-05-16 Dongping Li , Tielong Cai , Tianci Tang , Wenhao Chai , Katherine Rose Driggs-Campbell , Gaoang Wang

The significant advancements in visual understanding and instruction following from Multimodal Large Language Models (MLLMs) have opened up more possibilities for broader applications in diverse and universal human-centric scenarios.…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Keliang Li , Zaifei Yang , Jiahe Zhao , Hongze Shen , Ruibing Hou , Hong Chang , Shiguang Shan , Xilin Chen

In the realm of computer vision and robotics, embodied agents are expected to explore their environment and carry out human instructions. This necessitates the ability to fully understand 3D scenes given their first-person observations and…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Tai Wang , Xiaohan Mao , Chenming Zhu , Runsen Xu , Ruiyuan Lyu , Peisen Li , Xiao Chen , Wenwei Zhang , Kai Chen , Tianfan Xue , Xihui Liu , Cewu Lu , Dahua Lin , Jiangmiao Pang

We present EmbodiedMAE, a unified 3D multi-modal representation for robot manipulation. Current approaches suffer from significant domain gaps between training datasets and robot manipulation tasks, while also lacking model architectures…

机器人学 · 计算机科学 2025-05-16 Zibin Dong , Fei Ni , Yifu Yuan , Yinchuan Li , Jianye Hao

Heterogeneous multi-robot systems (HMRS) have emerged as a powerful approach for tackling complex tasks that single robots cannot manage alone. Current large-language-model-based multi-agent systems (LLM-based MAS) have shown success in…

机器人学 · 计算机科学 2025-02-18 Junting Chen , Checheng Yu , Xunzhe Zhou , Tianqi Xu , Yao Mu , Mengkang Hu , Wenqi Shao , Yikai Wang , Guohao Li , Lin Shao

Multimodal large language models have advanced rapidly, but their adoption in medicine is constrained by limited domain coverage, imperfect modality alignment, and insufficient grounded reasoning. We introduce MedMO, a medical multimodal…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Ankan Deria , Komal Kumar , Adinath Madhavrao Dukre , Eran Segal , Salman Khan , Imran Razzak

Existing Masked Image Modeling methods apply fixed mask patterns to guide the self-supervised training. As those mask patterns resort to different criteria to depict image contents, sticking to a fixed pattern leads to a limited vision cues…

计算机视觉与模式识别 · 计算机科学 2025-04-15 Zhanzhou Feng , Shiliang Zhang

Generalization in embodied AI is hindered by the "seeing-to-doing gap," which stems from data scarcity and embodiment heterogeneity. To address this, we pioneer "pointing" as a unified, embodiment-agnostic intermediate representation,…

机器人学 · 计算机科学 2026-04-07 Yifu Yuan , Haiqin Cui , Yaoting Huang , Yibin Chen , Fei Ni , Zibin Dong , Pengyi Li , Yan Zheng , Hongyao Tang , Jianye Hao

Heterogeneous multirobot systems show great potential in complex tasks requiring coordinated hybrid cooperation. However, existing methods that rely on static or task-specific models often lack generalizability across diverse tasks and…

机器人学 · 计算机科学 2025-10-28 Haokun Liu , Zhaoqi Ma , Yunong Li , Junichiro Sugihara , Yicheng Chen , Jinjie Li , Moju Zhao

Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. However, existing VLA models still face two fundamental challenges: (i) producing precise…

The remarkable progress of Multimodal Large Language Models (MLLMs) has attracted increasing attention to extend them to physical entities like legged robot. This typically requires MLLMs to not only grasp multimodal understanding…

Multimodal large language models (MLLMs) have shown remarkable potential in various domains, yet their application in the medical field is hindered by several challenges. General-purpose MLLMs often lack the specialized knowledge required…

人工智能 · 计算机科学 2025-09-29 Guanghao Zhu , Zhitian Hou , Zeyu Liu , Zhijie Sang , Congkai Xie , Hongxia Yang

Human-robot interaction is increasingly moving toward multi-robot, socially grounded environments. Existing systems struggle to integrate multimodal perception, embodied expression, and coordinated decision-making in a unified framework.…

机器人学 · 计算机科学 2026-03-25 Shaid Hasan , Breenice Lee , Sujan Sarker , Tariq Iqbal

While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on question-answering or multiple-choice formats. These protocols allow models to exploit…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Haozhe Shan , Xiancong Ren , Han Dong , Haoyuan Shi , Yingji Zhang , Jiayu Hu , Yi Zhang , Yong Dai , Bin Shen , Lizhen Qu , Zenglin Xu , Xiaozhu Ju
‹ 上一页 1 2 3 10 下一页 ›