English
Related papers

Related papers: SK-Adapter: Skeleton-Based Structural Control for …

200 papers

Segment anything model (SAM) demonstrates strong generalization ability on natural image segmentation. However, its direct adaptation in medical image segmentation tasks shows significant performance drops. It also requires an excessive…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Heng Guo , Jianfeng Zhang , Jiaxing Huang , Tony C. W. Mok , Dazhou Guo , Ke Yan , Le Lu , Dakai Jin , Minfeng Xu

Recent 3D generative models have achieved remarkable performance in synthesizing high resolution photorealistic images with view consistency and detailed 3D shapes, but training them for diverse domains is challenging since it requires…

Computer Vision and Pattern Recognition · Computer Science 2023-04-03 Gwanghyun Kim , Se Young Chun

Surface cracks in infrastructure can lead to severe deterioration and expensive maintenance if not efficiently repaired. Manual repair methods are labor-intensive, time-consuming, and imprecise. While advancements in robotic perception and…

Robotics · Computer Science 2025-08-13 Joshua Genova , Eric Cabrera , Vedhus Hoskere

Objects with complex structures pose significant challenges to existing instance segmentation methods that rely on boundary or affinity maps, which are vulnerable to small errors around contacting pixels that cause noticeable connectivity…

Computer Vision and Pattern Recognition · Computer Science 2023-10-10 Zudi Lin , Donglai Wei , Aarush Gupta , Xingyu Liu , Deqing Sun , Hanspeter Pfister

Generative modeling of 3D human bodies have been studied extensively in computer vision. The core is to design a compact latent representation that is both expressive and semantically interpretable, yet existing approaches struggle to…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 Haorui Ji , Rong Wang , Taojun Lin , Hongdong Li

Recent advancements in subject-driven image generation have led to zero-shot generation, yet precise selection and focus on crucial subject representations remain challenging. Addressing this, we introduce the SSR-Encoder, a novel…

Computer Vision and Pattern Recognition · Computer Science 2024-03-15 Yuxuan Zhang , Yiren Song , Jiaming Liu , Rui Wang , Jinpeng Yu , Hao Tang , Huaxia Li , Xu Tang , Yao Hu , Han Pan , Zhongliang Jing

Point cloud registration is fundamental in 3D vision applications, including autonomous driving, robotics, and medical imaging, where precise alignment of multiple point clouds is essential for accurate environment reconstruction. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Yongqiang Wang , Weigang Li , Wenping Liu , Zhiqiang Tian , Jinling Li

Many robotic tasks involving some form of 3D visual perception greatly benefit from a complete knowledge of the working environment. However, robots often have to tackle unstructured environments and their onboard visual sensors can only…

Computer Vision and Pattern Recognition · Computer Science 2022-10-24 Andrea Rosasco , Stefano Berti , Fabrizio Bottarel , Michele Colledanchise , Lorenzo Natale

We propose a new generative model for 3D garment deformations that enables us to learn, for the first time, a data-driven method for virtual try-on that effectively addresses garment-body collisions. In contrast to existing methods that…

Computer Vision and Pattern Recognition · Computer Science 2021-05-14 Igor Santesteban , Nils Thuerey , Miguel A. Otaduy , Dan Casas

Automatically estimating 3D skeleton, shape, camera viewpoints, and part articulation from sparse in-the-wild image ensembles is a severely under-constrained and challenging problem. Most prior methods rely on large-scale image datasets,…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Chun-Han Yao , Wei-Chih Hung , Yuanzhen Li , Michael Rubinstein , Ming-Hsuan Yang , Varun Jampani

Numerous prior studies predominantly emphasize constructing relation vectors for individual neighborhood points and generating dynamic kernels for each vector and embedding these into high-dimensional spaces to capture implicit local…

Computer Vision and Pattern Recognition · Computer Science 2024-04-24 Shuofeng Sun , Yongming Rao , Jiwen Lu , Haibin Yan

Creating 3D head avatars is a significant yet challenging task for many applicated scenarios. Previous studies have set out to learn 3D human head generative models using massive 2D image data. Although these models are highly generalizable…

Computer Vision and Pattern Recognition · Computer Science 2024-10-03 Yiyu Zhuang , Yuxiao He , Jiawei Zhang , Yanwen Wang , Jiahe Zhu , Yao Yao , Siyu Zhu , Xun Cao , Hao Zhu

Skeleton-based action recognition is widely utilized in sensor systems including human-computer interaction and intelligent surveillance. Nevertheless, current sensor devices typically generate sparse skeleton data as discrete coordinates,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Yuhan Chen , Yicui Shi , Guofa Li , Liping Zhang , Jie Li , Jiaxin Gao , Wenbo Chu

Efficient, accurate and low-cost estimation of human skeletal information is crucial for a range of applications such as biology education and human-computer interaction. However, current simple skeleton models, which are typically based on…

Computer Vision and Pattern Recognition · Computer Science 2024-09-04 Zhiheng Peng , Kai Zhao , Xiaoran Chen , Li Ma , Siyu Xia , Changjie Fan , Weijian Shang , Wei Jing

With the overwhelming trend of mask image modeling led by MAE, generative pre-training has shown a remarkable potential to boost the performance of fundamental models in 2D vision. However, in 3D vision, the over-reliance on…

Computer Vision and Pattern Recognition · Computer Science 2023-09-08 Ziyi Wang , Xumin Yu , Yongming Rao , Jie Zhou , Jiwen Lu

Semi-autonomous prosthesis controllers based on computer vision improve performance while reducing cognitive effort. However, controllers relying on full-depth data face challenges in being deployed as embedded prosthesis controllers due to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-28 Miguel Nobre Castro , Strahinja Dosen

Capitalizing on large pre-trained models for various downstream tasks of interest have recently emerged with promising performance. Due to the ever-growing model size, the standard full fine-tuning based task adaptation strategy becomes…

Computer Vision and Pattern Recognition · Computer Science 2022-10-14 Junting Pan , Ziyi Lin , Xiatian Zhu , Jing Shao , Hongsheng Li

The last several years have seen significant progress in using depth cameras for tracking articulated objects such as human bodies, hands, and robotic manipulators. Most approaches focus on tracking skeletal parameters of a fixed shape…

Computer Vision and Pattern Recognition · Computer Science 2017-11-23 Aaron Walsman , Weilin Wan , Tanner Schmidt , Dieter Fox

Estimating 3D geometry from monocular colonoscopy images is challenging due to non-Lambertian surfaces, moving light sources, and large textureless regions. While recent 3D geometric foundation models eliminate the need for multi-stage…

Image and Video Processing · Electrical Eng. & Systems 2025-12-01 Zhiyi Jiang , Yifu Wang , Xuelian Cheng , Zongyuan Ge

In Natural Language (NL) applications, there is often a mismatch between what the NL interface is capable of interpreting and what a lay user knows how to express. This work describes a novel natural language interface that reduces this…

Computation and Language · Computer Science 2020-12-14 Clifton McFate , Aditya Kalyanpur , Dave Ferrucci , Andrea Bradshaw , Ariel Diertani , David Melville , Lori Moon