English
Related papers

Related papers: Pose-Free Generalizable Rendering Transformer

200 papers

Generating a photorealistic image with intended human pose is a promising yet challenging research topic for many applications such as smart photo editing, movie making, virtual try-on, and fashion display. In this paper, we present a novel…

Computer Vision and Pattern Recognition · Computer Science 2019-09-19 Wei Sun , Jawadul H. Bappy , Shanglin Yang , Yi Xu , Tianfu Wu , Hui Zhou

While object reconstruction has made great strides in recent years, current methods typically require densely captured images and/or known camera poses, and generalize poorly to novel object categories. To step toward object reconstruction…

Computer Vision and Pattern Recognition · Computer Science 2024-01-29 Hanwen Jiang , Zhenyu Jiang , Kristen Grauman , Yuke Zhu

We address the task of multi-view image-to-image translation for person image generation. The goal is to synthesize photo-realistic multi-view images with pose-consistency across all views. Our proposed end-to-end framework is based on a…

Computer Vision and Pattern Recognition · Computer Science 2021-04-14 Idit Diamant , Oranit Dror , Hai Victor Habi , Arnon Netzer

This paper proposes a new generative adversarial network for pose transfer, i.e., transferring the pose of a given person to a target pose. The generator of the network comprises a sequence of Pose-Attentional Transfer Blocks that each…

Computer Vision and Pattern Recognition · Computer Science 2019-05-14 Zhen Zhu , Tengteng Huang , Baoguang Shi , Miao Yu , Bofei Wang , Xiang Bai

Existing automatic approaches for 3D virtual character motion synthesis supporting scene interactions do not generalise well to new objects outside training distributions, even when trained on extensive motion capture datasets with diverse…

Computer Vision and Pattern Recognition · Computer Science 2024-02-16 Wanyue Zhang , Rishabh Dabral , Thomas Leimkühler , Vladislav Golyanik , Marc Habermann , Christian Theobalt

We introduce GNeRF, a framework to marry Generative Adversarial Networks (GAN) with Neural Radiance Field (NeRF) reconstruction for the complex scenarios with unknown and even randomly initialized camera poses. Recent NeRF-based advances…

Computer Vision and Pattern Recognition · Computer Science 2021-08-19 Quan Meng , Anpei Chen , Haimin Luo , Minye Wu , Hao Su , Lan Xu , Xuming He , Jingyi Yu

This work targets at using a general deep learning framework to synthesize free-viewpoint images of arbitrary human performers, only requiring a sparse number of camera views as inputs and skirting per-case fine-tuning. The large variation…

Computer Vision and Pattern Recognition · Computer Science 2022-04-26 Wei Cheng , Su Xu , Jingtan Piao , Chen Qian , Wayne Wu , Kwan-Yee Lin , Hongsheng Li

Recent advances in diffusion-based generative models have shown incredible promise for zero shot image-to-image translation and editing. Most of these approaches work by combining or replacing network-specific features used in the…

Computer Vision and Pattern Recognition · Computer Science 2025-10-06 Zeqi Gu , Ethan Yang , Abe Davis

While learning-based models hold great promise for MRI reconstruction, single-site models trained on limited local datasets often show poor generalization. This has motivated collaborative training across institutions via federated learning…

Image and Video Processing · Electrical Eng. & Systems 2025-11-10 Valiyeh A. Nezhad , Gokberk Elmas , Bilal Kabas , Fuat Arslan , Emine U. Saritas , Tolga Çukur

Feature extraction, matching, structure from motion (SfM), and novel view synthesis (NVS) have traditionally been treated as separate problems with independent optimization objectives. We present GloSplat, a framework that performs…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Tianyu Xiong , Rui Li , Linjie Li , Jiaqi Yang

3D Gaussian Splatting (3DGS) has demonstrated remarkable real-time performance in novel view synthesis, yet its effectiveness relies heavily on dense multi-view inputs with precisely known camera poses, which are rarely available in…

Computer Vision and Pattern Recognition · Computer Science 2025-08-22 Zongqi He , Hanmin Li , Kin-Chung Chan , Yushen Zuo , Hao Xie , Zhe Xiao , Jun Xiao , Kin-Man Lam

The creation of 2D realistic facial images and 3D face shapes using generative networks has been a hot topic in recent years. Existing face generators exhibit exceptional performance on faces in small to medium poses (with respect to…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Yiqian Wu , Jing Zhang , Hongbo Fu , Xiaogang Jin

Some deep learning-based point cloud registration methods struggle with zero-shot generalization, often requiring dataset-specific hyperparameter tuning or retraining for new environments. We identify three critical limitations: (a) fixed…

Computer Vision and Pattern Recognition · Computer Science 2026-01-07 Hyungtae Lim , Minkyun Seo , Luca Carlone , Jaesik Park

A new passive approach called Generalized Scene Reconstruction (GSR) enables "generalized scenes" to be effectively reconstructed. Generalized scenes are defined to be "boundless" spaces that include non-Lambertian, partially transmissive,…

Computer Vision and Pattern Recognition · Computer Science 2018-05-28 John K. Leffingwell , Donald J. Meagher , Khan W. Mahmud , Scott Ackerson

The task of face reenactment is to transfer the head motion and facial expressions from a driving video to the appearance of a source image, which may be of a different person (cross-reenactment). Most existing methods are CNN-based and…

Computer Vision and Pattern Recognition · Computer Science 2024-06-11 Andre Rochow , Max Schwarz , Sven Behnke

The inherent challenge of image fusion lies in capturing the correlation of multi-source images and comprehensively integrating effective information from different sources. Most existing techniques fail to perform dynamic image fusion…

Computer Vision and Pattern Recognition · Computer Science 2024-11-06 Bing Cao , Yinan Xia , Yi Ding , Changqing Zhang , Qinghua Hu

Immersive scene generation, notably panorama creation, benefits significantly from the adaptation of large pre-trained text-to-image (T2I) models for multi-view image generation. Due to the high cost of acquiring multi-view images,…

Computer Vision and Pattern Recognition · Computer Science 2024-08-06 Aoming Liu , Zhong Li , Zhang Chen , Nannan Li , Yi Xu , Bryan A. Plummer

We propose a method at the intersection of Computer Vision and Computer Graphics fields, which automatically generates RGBD images using neural networks, based on previously seen and synchronized video, depth and pose signals. Since the…

Computer Vision and Pattern Recognition · Computer Science 2020-07-15 Mihai Cristian Pîrvu

The vast majority of Shape-from-Polarization (SfP) methods work under the oversimplified assumption of using orthographic cameras. Indeed, it is still not well understood how to project the Stokes vectors when the incoming rays are not…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Mara Pistellato , Filippo Bergamasco

Generalizable 3D Gaussian splitting (3DGS) can reconstruct new scenes from sparse-view observations in a feed-forward inference manner, eliminating the need for scene-specific retraining required in conventional 3DGS. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Zhiyuan Min , Yawei Luo , Jianwen Sun , Yi Yang