中文
相关论文

相关论文: Transformer Driven Visual Servoing for Fabric Text…

200 篇论文

Virtual try-on aims to synthesize a realistic image of a person wearing a target garment, but accurately modeling garment-body correspondence remains a persistent challenge, especially under pose and appearance variation. In this paper, we…

图形学 · 计算机科学 2025-11-06 Seungyong Lee , Jeong-gi Kwak

This paper considers image-based virtual try-on, which renders an image of a person wearing a curated garment, given a pair of images depicting the person and the garment, respectively. Previous works adapt existing exemplar-based…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Yisol Choi , Sangkyung Kwak , Kyungmin Lee , Hyungwon Choi , Jinwoo Shin

Prior methods for controlling image generation are limited in their ability to be taught new tasks. In contrast, vision-language models, or VLMs, can learn tasks in-context and produce the correct outputs for a given input. We propose a…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Grace Luo , Jonathan Granskog , Aleksander Holynski , Trevor Darrell

We have seen much recent progress in rigid object manipulation, but interaction with deformable objects has notably lagged behind. Due to the large configuration space of deformable objects, solutions using traditional modelling approaches…

机器人学 · 计算机科学 2018-10-09 Jan Matas , Stephen James , Andrew J. Davison

In physics-based cloth animation, rich folds and detailed wrinkles are achieved at the cost of expensive computational resources and huge labor tuning. Data-driven techniques make efforts to reduce the computation significantly by a…

计算机视觉与模式识别 · 计算机科学 2021-02-24 Lan Chen , Lin Gao , Jie Yang , Shibiao Xu , Juntao Ye , Xiaopeng Zhang , Yu-Kun Lai

The swiftly expanding retail sector is increasingly adopting autonomous mobile robots empowered by artificial intelligence and machine learning algorithms to gain an edge in the competitive market. However, these autonomous robots encounter…

Precise reconstruction and manipulation of the crumpled cloths is challenging due to the high dimensionality of cloth models, as well as the limited observation at self-occluded regions. We leverage the recent progress in the field of…

机器人学 · 计算机科学 2024-05-16 Wenbo Wang , Gen Li , Miguel Zamora , Stelian Coros

We consider the problem of grasping deformable objects with soft shells using a robotic gripper. Such objects have a center-of-mass that changes dynamically and are fragile so prone to burst. Thus, it is difficult for robots to generate…

机器人学 · 计算机科学 2025-10-14 Yonghyun Lee , Sungeun Hong , Min-gu Kim , Gyeonghwan Kim , Changjoo Nam

Deepfakes, which employ GAN to produce highly realistic facial modification, are widely regarded as the prevailing method. Traditional CNN have been able to identify bogus media, but they struggle to perform well on different datasets and…

计算机视觉与模式识别 · 计算机科学 2025-10-24 Deepak Dagar , Dinesh Kumar Vishwakarma

The Diffusion model has a strong ability to generate wild images. However, the model can just generate inaccurate images with the guidance of text, which makes it very challenging to directly apply the text-guided generative model for…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Shufang Zhang , Minxue Ni , Lei Wang , Wenxin Ding , Shuai Chen , Yuhong Liu

Manipulating deformable objects in robotic cells is often costly and not widely accessible. However, the use of localized pneumatic gripping systems can enhance accessibility. Current methods that use pneumatic grippers to handle deformable…

机器人学 · 计算机科学 2025-01-10 Roman Mykhailyshyn , Jonathan Lee , Mykhailo Mykhailyshyn , Kensuke Harada , Ann Majewicz Fey

Robotic automation in surgery requires precise tracking of surgical tools and mapping of deformable tissue. Previous works on surgical perception frameworks require significant effort in developing features for surgical tool and tissue…

机器人学 · 计算机科学 2021-03-26 Jingpei Lu , Ambareesh Jayakumari , Florian Richter , Yang Li , Michael C. Yip

Visual servoing techniques guide robotic motion using visual information to accomplish manipulation tasks, requiring high precision and robustness against noise. Traditional methods often require prior knowledge and are susceptible to…

机器人学 · 计算机科学 2026-02-24 Haoyu Zhang , Yang Liu , Yimu Jiang , Weiyang Lin , Chao Ye

The recent trend in multiple object tracking (MOT) is heading towards leveraging deep learning to boost the tracking performance. In this paper, we propose a novel solution named TransSTAM, which leverages Transformer to effectively model…

计算机视觉与模式识别 · 计算机科学 2022-06-01 Peng Dai , Yiqiang Feng , Renliang Weng , Changshui Zhang

Remote sensing change detection between bi-temporal images receives growing concentration from researchers. However, comparing two bi-temporal images for detecting changes is challenging, as they demonstrate different appearances. In this…

计算机视觉与模式识别 · 计算机科学 2023-10-04 Luyi Qiu , Xiaofeng Zhang , ChaoChen Gu , and ShanYing Zhu

Document image dewarping remains a challenging task in the deep learning era. While existing methods have improved by leveraging text line awareness, they typically focus only on a single horizontal dimension. In this paper, we propose a…

计算机视觉与模式识别 · 计算机科学 2026-03-05 Heng Li , Xiangping Wu , Qingcai Chen

This paper describes the integration of weighted delay-and-sum beamforming with speech source localization using image processing and robot head visual servoing for source tracking. We take into consideration the fact that the directivity…

音频与语音处理 · 电气工程与系统科学 2019-06-19 José Novoa , Rodrigo Mahu , Alejandro Díaz , Jorge Wuth , Richard Stern , Nestor Becerra Yoma

Image matting aims to predict alpha values of elaborate uncertainty areas of natural images, like hairs, smoke, and spider web. However, existing methods perform poorly when faced with highly transparent foreground objects due to the large…

计算机视觉与模式识别 · 计算机科学 2023-03-14 Huanqia Cai , Fanglei Xue , Lele Xu , Lili Guo

Virtual Try-Off (VTOFF) is a challenging multimodal image generation task that aims to synthesize high-fidelity flat-lay garments under complex geometric deformation and rich high-frequency textures. Existing methods often rely on…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Yihan Zhu , Mengying Ge

Empowered by deep learning, recent methods for material capture can estimate a spatially-varying reflectance from a single photograph. Such lightweight capture is in stark contrast with the tens or hundreds of pictures required by…

图形学 · 计算机科学 2019-06-28 Valentin Deschaintre , Miika Aittala , Fredo Durand , George Drettakis , Adrien Bousseau