中文
相关论文

相关论文: 3D-GOI: 3D GAN Omni-Inversion for Multifaceted and…

200 篇论文

Differentiable rendering has paved the way to training neural networks to perform "inverse graphics" tasks such as predicting 3D geometry from monocular photographs. To train high performing models, most of the current approaches rely on…

计算机视觉与模式识别 · 计算机科学 2021-04-22 Yuxuan Zhang , Wenzheng Chen , Huan Ling , Jun Gao , Yinan Zhang , Antonio Torralba , Sanja Fidler

Recent advancements in large-scale text-to-image diffusion models have enabled many applications in image editing. However, none of these methods have been able to edit the layout of single existing images. To address this gap, we propose…

计算机视觉与模式识别 · 计算机科学 2023-06-23 Zhiyuan Zhang , Zhitong Huang , Jing Liao

In this paper, we introduce Geometry-Inverse-Meet-Pixel-Insert, short for GEO, an exceptionally versatile image editing technique designed to cater to customized user requirements at both local and global scales. Our approach seamlessly…

计算机视觉与模式识别 · 计算机科学 2024-09-19 Yan Zheng , Lemeng Wu

In this paper, we study the problem of 3D scene geometry decomposition and manipulation from 2D views. By leveraging the recent implicit neural representation techniques, particularly the appealing neural radiance fields, we introduce an…

计算机视觉与模式识别 · 计算机科学 2023-03-13 Bing Wang , Lu Chen , Bo Yang

The objective of person re-identification (re-ID) is to retrieve a person's images from an image gallery, given a single instance of the person of interest. Despite several advancements, learning discriminative identity-sensitive and…

计算机视觉与模式识别 · 计算机科学 2021-06-02 Arnab Karmakar , Deepak Mishra

Recently, AI-manipulated face techniques have developed rapidly and constantly, which has raised new security issues in society. Although existing detection methods consider different categories of fake faces, the performance on detecting…

计算机视觉与模式识别 · 计算机科学 2024-10-30 Yang Yu , Rongrong Ni , Yao Zhao

Previous portrait image generation methods roughly fall into two categories: 2D GANs and 3D-aware GANs. 2D GANs can generate high fidelity portraits but with low view consistency. 3D-aware GAN methods can maintain view consistency but their…

计算机视觉与模式识别 · 计算机科学 2022-03-22 Jingxiang Sun , Xuan Wang , Yong Zhang , Xiaoyu Li , Qi Zhang , Yebin Liu , Jue Wang

We introduce InseRF, a novel method for generative object insertion in the NeRF reconstructions of 3D scenes. Based on a user-provided textual description and a 2D bounding box in a reference viewpoint, InseRF generates new objects in 3D…

计算机视觉与模式识别 · 计算机科学 2024-01-11 Mohamad Shahbazi , Liesbeth Claessens , Michael Niemeyer , Edo Collins , Alessio Tonioni , Luc Van Gool , Federico Tombari

Large-scale text-to-image models enable a wide range of image editing techniques, using text prompts or even spatial controls. However, applying these editing methods to multi-view images depicting a single scene leads to 3D-inconsistent…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Or Patashnik , Rinon Gal , Daniel Cohen-Or , Jun-Yan Zhu , Fernando De la Torre

Image inversion is a fundamental task in generative models, aiming to map images back to their latent representations to enable downstream applications such as editing, restoration, and style transfer. This paper provides a comprehensive…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Yinan Chen , Jiangning Zhang , Yali Bi , Xiaobin Hu , Teng Hu , Zhucun Xue , Ran Yi , Yong Liu , Ying Tai

Reconstructing objects from posed images is a crucial and complex task in computer graphics and computer vision. While NeRF-based neural reconstruction methods have exhibited impressive reconstruction ability, they tend to be…

计算机视觉与模式识别 · 计算机科学 2024-10-18 Shuichang Lai , Letian Huang , Jie Guo , Kai Cheng , Bowen Pan , Xiaoxiao Long , Jiangjing Lyu , Chengfei Lv , Yanwen Guo

Close-up facial images captured at short distances often suffer from perspective distortion, resulting in exaggerated facial features and unnatural/unattractive appearances. We propose a simple yet effective method for correcting…

计算机视觉与模式识别 · 计算机科学 2023-12-11 Zhixiang Wang , Yu-Lun Liu , Jia-Bin Huang , Shin'ichi Satoh , Sizhuo Ma , Gurunandan Krishnan , Jian Wang

Generative Adversarial Network (GAN) inversion have demonstrated excellent performance in image inpainting that aims to restore lost or damaged image texture using its unmasked content. Previous GAN inversion-based methods usually utilize…

计算机视觉与模式识别 · 计算机科学 2025-04-18 Libo Zhang , Yongsheng Yu , Jiali Yao , Heng Fan

Monocular 3D clothed human reconstruction aims to generate a complete and realistic textured 3D avatar from a single image. Existing methods are commonly trained under multi-view supervision with annotated geometric priors, and during…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Nanjie Yao , Gangjian Zhang , Wenhao Shen , Jian Shu , Yu Feng , Hao Wang

Current methods commonly utilize three-branch structures of inversion, reconstruction, and editing, to tackle consistent image editing task. However, these methods lack control over the generation position of the edited object and have…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Pengfei Jiang , Mingbao Lin , Fei Chao

Image inpainting is an old problem in computer vision that restores occluded regions and completes damaged images. In the case of facial image inpainting, most of the methods generate only one result for each masked image, even though there…

计算机视觉与模式识别 · 计算机科学 2023-01-23 Dongsik Yoon , Jeong-gi Kwak , Yuanming Li , David Han , Hanseok Ko

Despite the demonstrated editing capacity in the latent space of a pretrained GAN model, inverting real-world images is stuck in a dilemma that the reconstruction cannot be faithful to the original input. The main reason for this is that…

计算机视觉与模式识别 · 计算机科学 2022-07-19 Haorui Song , Yong Du , Tianyi Xiang , Junyu Dong , Jing Qin , Shengfeng He

In recent years, text-guided image manipulation has gained increasing attention in the multimedia and computer vision community. The input to conditional image generation has evolved from image-only to multimodality. In this paper, we study…

计算机视觉与模式识别 · 计算机科学 2021-11-30 Tianhao Zhang , Hung-Yu Tseng , Lu Jiang , Weilong Yang , Honglak Lee , Irfan Essa

Corner cases are crucial for training and validating autonomous driving systems, yet collecting them from the real world is often costly and hazardous. Editing objects within captured sensor data offers an effective alternative for…

计算机视觉与模式识别 · 计算机科学 2025-08-29 Jiusi Li , Jackson Jiang , Jinyu Miao , Miao Long , Tuopu Wen , Peijin Jia , Shengxiang Liu , Chunlei Yu , Maolin Liu , Yuzhan Cai , Kun Jiang , Mengmeng Yang , Diange Yang

Recent improvements to Generative Adversarial Networks (GANs) have made it possible to generate realistic images in high resolution based on natural language descriptions such as image captions. Furthermore, conditional GANs allow us to…

计算机视觉与模式识别 · 计算机科学 2019-01-04 Tobias Hinz , Stefan Heinrich , Stefan Wermter