中文
相关论文

相关论文: GRAZE: Grounded Refinement and Motion-Aware Zero-S…

200 篇论文

In this paper, we present Sim-Grasp, a robust 6-DOF two-finger grasping system that integrates advanced language models for enhanced object manipulation in cluttered environments. We introduce the Sim-Grasp-Dataset, which includes 1,550…

机器人学 · 计算机科学 2024-07-18 Juncheng Li , David J. Cappelleri

Visual Place Recognition (VPR) requires robust retrieval of geotagged images despite large appearance, viewpoint, and environmental variation. Prior methods focus on descriptor fine-tuning or fixed sampling strategies yet neglect the…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Shunpeng Chen , Changwei Wang , Rongtao Xu , Xingtian Pei , Yukun Song , Jinzhou Lin , Wenhao Xu , Jingyi Zhang , Li Guo , Shibiao Xu

Handling fragile objects remains a major challenge for robotic manipulation. Tactile sensing and soft robotics can improve delicate object handling, but typically involve high integration complexity or slow response times. We address these…

机器人学 · 计算机科学 2026-02-17 Siqi Shang , Mingyo Seo , Yuke Zhu , Lillian Chin

This paper presents a grounded language-image pre-training (GLIP) model for learning object-level, language-aware, and semantic-rich visual representations. GLIP unifies object detection and phrase grounding for pre-training. The…

Grasp pose detection in cluttered, real-world environments remains a significant challenge due to noisy and incomplete sensory data combined with complex object geometries. This paper introduces Grasp the Graph 2.0 (GtG 2.0) method, a…

机器人学 · 计算机科学 2026-01-12 Ali Rashidi Moghadam , Sayedmohammadreza Rastegari , Mehdi Tale Masouleh , Ahmad Kalhor

We present a model for temporally precise action spotting in videos, which uses a dense set of detection anchors, predicting a detection confidence and corresponding fine-grained temporal displacement for each anchor. We experiment with two…

计算机视觉与模式识别 · 计算机科学 2022-07-13 João V. B. Soares , Avijit Shah , Topojoy Biswas

In this paper, we present a novel method for self-supervised fine-tuning of pose estimation. Leveraging zero-shot pose estimation, our approach enables the robot to automatically obtain training data without manual labeling. After pose…

机器人学 · 计算机科学 2024-12-13 Frederik Hagelskjær

We address the problem of fine-grained action localization from temporally untrimmed web videos. We assume that only weak video-level annotations are available for training. The goal is to use these weak labels to identify temporal segments…

计算机视觉与模式识别 · 计算机科学 2015-08-05 Chen Sun , Sanketh Shetty , Rahul Sukthankar , Ram Nevatia

We address the task of jointly determining what a person is doing and where they are looking based on the analysis of video captured by a headworn camera. To facilitate our research, we first introduce the EGTEA Gaze+ dataset. Our dataset…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Yin Li , Miao Liu , James M. Rehg

Surveillance footage represents a valuable resource and opportunities for conducting gait analysis. However, the typical low quality and high noise levels in such footage can severely impact the accuracy of pose estimation algorithms, which…

计算机视觉与模式识别 · 计算机科学 2024-04-19 Andrei Niculae , Andy Catruna , Adrian Cosma , Daniel Rosner , Emilian Radoi

High-dimensional recordings of dynamical processes are often characterized by a much smaller set of effective variables, evolving on low-dimensional manifolds. Identifying these latent dynamics requires solving two intertwined problems:…

机器学习 · 计算机科学 2026-01-21 Manuel Hinz , Maximilian Mauel , Patrick Seifner , David Berghaus , Kostadin Cvejoski , Ramses J. Sanchez

In this paper, we present a transformer-based architecture, namely TF-Grasp, for robotic grasp detection. The developed TF-Grasp framework has two elaborate designs making it well suitable for visual grasping tasks. The first key design is…

机器人学 · 计算机科学 2022-09-14 Shaochen Wang , Zhangli Zhou , Zhen Kan

Humans, this species expert in grasp detection, can grasp objects by taking into account hand-object positioning information. This work proposes a method to enable a robot manipulator to learn the same, grasping objects in the most optimal…

Although, in the task of grasping via a data-driven method, closed-loop feedback and predicting 6 degrees of freedom (DoF) grasp rather than conventionally used 4DoF top-down grasp are demonstrated to improve performance individually, few…

机器人学 · 计算机科学 2022-06-22 Dongwon Son

Gaze is a valuable means of communication for impaired people with extremely limited motor capabilities. However, robust gaze-based intent recognition in multi-object environments is challenging due to gaze noise, micro-saccades, viewpoint…

机器人学 · 计算机科学 2026-03-09 Yuzhi Lai , Shenghai Yuan , Peizheng Li , Andreas Zell

Grasping is natural for humans. However, it involves complex hand configurations and soft tissue deformation that can result in complicated regions of contact between the hand and the object. Understanding and modeling this contact can…

计算机视觉与模式识别 · 计算机科学 2020-07-21 Samarth Brahmbhatt , Chengcheng Tang , Christopher D. Twigg , Charles C. Kemp , James Hays

In this paper, we study the problem of task-oriented grasp synthesis from partial point cloud data using an eye-in-hand camera configuration. In task-oriented grasp synthesis, a grasp has to be selected so that the object is not lost during…

机器人学 · 计算机科学 2023-09-22 Aditya Patankar , Khiem Phi , Dasharadhan Mahalingam , Nilanjan Chakraborty , IV Ramakrishnan

In this work, we address the limitation of surface fitting-based grasp planning algorithm, which primarily focuses on geometric alignment between the gripper and object surface while overlooking the stability of contact point distribution,…

A representation gap exists between grasp synthesis for rigid and soft grippers. Anygrasp [1] and many other grasp synthesis methods are designed for rigid parallel grippers, and adapting them to soft grippers often fails to capture their…

机器人学 · 计算机科学 2026-02-20 Tanisha Parulekar , Ge Shi , Josh Pinskier , David Howard , Jen Jen Chung

Vision-based grasping of unknown objects in unstructured environments is a key challenge for autonomous robotic manipulation. A practical grasp synthesis system is required to generate a diverse set of 6-DoF grasps from which a…

机器人学 · 计算机科学 2024-11-26 Kuldeep R Barad , Andrej Orsula , Antoine Richard , Jan Dentler , Miguel Olivares-Mendez , Carol Martinez