中文
相关论文

相关论文: Right-Side-Out: Learning Zero-Shot Sim-to-Real Gar…

200 篇论文

Designing controllers that accomplish tasks while guaranteeing safety constraints remains a significant challenge. We often want an agent to perform well in a nominal task, such as environment exploration, while ensuring it can avoid unsafe…

系统与控制 · 电气工程与系统科学 2025-06-04 Azra Begzadić , Nikhil Uday Shinde , Sander Tonkens , Dylan Hirsch , Kaleb Ugalde , Michael C. Yip , Jorge Cortés , Sylvia Herbert

Segment Anything Model (SAM) is an advanced foundational model for image segmentation, which is gradually being applied to remote sensing images (RSIs). Due to the domain gap between RSIs and natural images, traditional methods typically…

计算机视觉与模式识别 · 计算机科学 2025-01-14 Nanqing Liu , Xun Xu , Yongyi Su , Haojie Zhang , Heng-Chao Li

High-fidelity physics simulation is essential for scalable robotic learning, but the sim-to-real gap persists, especially for tasks involving complex, dynamic, and discontinuous interactions like physical contacts. Explicit system…

机器人学 · 计算机科学 2026-01-21 Changwei Jing , Jai Krishna Bandi , Jianglong Ye , Yan Duan , Pieter Abbeel , Xiaolong Wang , Sha Yi

Many problems in image processing and computer vision (e.g. colorization, style transfer) can be posed as 'manipulating' an input image into a corresponding output image given a user-specified guiding signal. A holy-grail solution towards…

计算机视觉与模式识别 · 计算机科学 2017-03-23 Hao Wang , Xiaodan Liang , Hao Zhang , Dit-Yan Yeung , Eric P. Xing

We study active object tracking, where a tracker takes visual observations (i.e., frame sequences) as input and produces the corresponding camera control signals as output (e.g., move forward, turn left, etc.). Conventional methods tackle…

计算机视觉与模式识别 · 计算机科学 2019-02-14 Wenhan Luo , Peng Sun , Fangwei Zhong , Wei Liu , Tong Zhang , Yizhou Wang

To tackle the "reality gap" encountered in Sim-to-Real transfer, this study proposes a diffusion-based framework that minimizes inconsistencies in grasping actions between the simulation settings and realistic environments. The process…

机器人学 · 计算机科学 2024-03-19 Yiwei Li , Zihao Wu , Huaqin Zhao , Tianze Yang , Zhengliang Liu , Peng Shu , Jin Sun , Ramviyas Parasuraman , Tianming Liu

Large-scale image-text pre-trained models enable zero-shot classification and provide consistent accuracy across various data distributions. Nonetheless, optimizing these models in downstream tasks typically requires fine-tuning, which…

计算机视觉与模式识别 · 计算机科学 2024-08-13 Sungyeon Kim , Boseung Jeong , Donghyun Kim , Suha Kwak

We introduce Zero-1-to-3, a framework for changing the camera viewpoint of an object given just a single RGB image. To perform novel view synthesis in this under-constrained setting, we capitalize on the geometric priors that large-scale…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Ruoshi Liu , Rundi Wu , Basile Van Hoorick , Pavel Tokmakov , Sergey Zakharov , Carl Vondrick

This paper presents a learning-based clothing animation method for highly efficient virtual try-on simulation. Given a garment, we preprocess a rich database of physically-based dressed character simulations, for multiple body shapes and…

计算机视觉与模式识别 · 计算机科学 2019-03-19 Igor Santesteban , Miguel A. Otaduy , Dan Casas

Recent advancements in text-guided diffusion models have shown promise for general image editing via inversion techniques, but often struggle to maintain ID and structural consistency in real face editing tasks. To address this limitation,…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Yang Hou , Minggu Wang , Jianjun Zhao

This paper explores the development of UniFolding, a sample-efficient, scalable, and generalizable robotic system for unfolding and folding various garments. UniFolding employs the proposed UFONet neural network to integrate unfolding and…

机器人学 · 计算机科学 2023-11-03 Han Xue , Yutong Li , Wenqiang Xu , Huanyu Li , Dongzhe Zheng , Cewu Lu

Data driven and learning based solutions for modeling dynamic garments have significantly advanced, especially in the context of digital humans. However, existing approaches often focus on modeling garments with respect to a fixed…

图形学 · 计算机科学 2024-07-09 Peizhuo Li , Tuanfeng Y. Wang , Timur Levent Kesdogan , Duygu Ceylan , Olga Sorkine-Hornung

Cloth-changing person re-identification (re-ID) is a new rising research topic that aims at retrieving pedestrians whose clothes are changed. This task is quite challenging and has not been fully studied to date. Current works mainly focus…

计算机视觉与模式识别 · 计算机科学 2021-07-27 Xiujun Shu , Ge Li , Xiao Wang , Weijian Ruan , Qi Tian

Traditional PID controllers have limited adaptability for plasma shape control, and task-specific reinforcement learning (RL) methods suffer from limited generalization and the need for repetitive retraining. To overcome these challenges,…

等离子体物理 · 物理学 2025-10-21 Niannian Wu , Rongpeng Li , Zongyu Yang , Yong Xiao , Ning Wei , Yihang Chen , Bo Li , Zhifeng Zhao , Wulyu Zhong

Sim-to-real transfer remains a fundamental challenge in robot manipulation due to the entanglement of perception and control in end-to-end learning. We present a decoupled framework that learns each component where it is most reliable:…

机器人学 · 计算机科学 2025-10-01 Jialei Huang , Zhaoheng Yin , Yingdong Hu , Shuo Wang , Xingyu Lin , Yang Gao

Reinforcement learning (RL) algorithms can enable high-maneuverability in unmanned aerial vehicles (MAVs), but transferring them from simulation to real-world use is challenging. Variable-pitch propeller (VPP) MAVs offer greater agility,…

机器人学 · 计算机科学 2025-04-11 Zhikun Wang , Shiyu Zhao

Image-based virtual try-on, widely used in online shopping, aims to generate images of a naturally dressed person conditioned on certain garments, providing significant research and commercial potential. A key challenge of try-on is to…

计算机视觉与模式识别 · 计算机科学 2024-11-18 Hanzhong Guo , Jianfeng Zhang , Cheng Zou , Jun Li , Meng Wang , Ruxue Wen , Pingzhong Tang , Jingdong Chen , Ming Yang

The 2D virtual try-on task has recently attracted a great interest from the research community, for its direct potential applications in online shopping as well as for its inherent and non-addressed scientific challenges. This task requires…

计算机视觉与模式识别 · 计算机科学 2020-07-30 Thibaut Issenhuth , Jérémie Mary , Clément Calauzènes

We present a system for non-prehensile manipulation that require a significant number of contact mode transitions and the use of environmental contacts to successfully manipulate an object to a target location. Our method is based on deep…

机器人学 · 计算机科学 2023-09-07 Minchan Kim , Junhyek Han , Jaehyung Kim , Beomjoon Kim

Methods for object detection and segmentation often require abundant instance-level annotations for training, which are time-consuming and expensive to collect. To address this, the task of zero-shot object detection (or segmentation) aims…

计算机视觉与模式识别 · 计算机科学 2023-02-16 Siddhesh Khandelwal , Anirudth Nambirajan , Behjat Siddiquie , Jayan Eledath , Leonid Sigal