中文
相关论文

相关论文: MV-CoLight: Efficient Object Compositing with Cons…

200 篇论文

Generating high-quality 3D assets from a given image is highly desirable in various applications such as AR/VR. Recent advances in single-image 3D generation explore feed-forward models that learn to infer the 3D model of an object without…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Yongwei Chen , Tengfei Wang , Tong Wu , Xingang Pan , Kui Jia , Ziwei Liu

Recent 3D-aware head generative models based on 3D Gaussian Splatting achieve real-time, photorealistic and view-consistent head synthesis. However, a fundamental limitation persists: the deep entanglement of illumination and intrinsic…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Yating Wang , Yuan Sun , Xuan Wang , Ran Yi , Boyao Zhou , Yipengjing Sun , Hongyu Liu , Yinuo Wang , Lizhuang Ma

Existing deep learning-based low-light enhancement methods are typically trained on limited datasets with single enhancement targets, which restricts their generalization ability and controllability in real-world applications. To overcome…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Yufeng Yang , Jianzhuang Liu , Jisheng Chu , Yuqi Peng , Xianfang Zeng , Jiancheng Huang , Shifeng Chen

4D radar has received significant attention in autonomous driving thanks to its robustness under adverse weathers. Due to the sparse points and noisy measurements of the 4D radar, most of the research finish the 3D object detection task by…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Hanzhi Zhong , Zhiyu Xiang , Ruoyu Xu , Jingyun Fu , Peng Xu , Shaohong Wang , Zhihao Yang , Tianyu Pu , Eryun Liu

In this work, we present a conceptually simple yet effective framework for cross-modality 3D object detection, named voxel field fusion. The proposed approach aims to maintain cross-modality consistency by representing and fusing augmented…

计算机视觉与模式识别 · 计算机科学 2022-06-01 Yanwei Li , Xiaojuan Qi , Yukang Chen , Liwei Wang , Zeming Li , Jian Sun , Jiaya Jia

In computer vision, it is well-known that a lack of data diversity will impair model performance. In this study, we address the challenges of enhancing the dataset diversity problem in order to benefit various downstream tasks such as…

计算机视觉与模式识别 · 计算机科学 2024-08-02 Yuhang Li , Xin Dong , Chen Chen , Weiming Zhuang , Lingjuan Lyu

Despite the impressive text-to-image (T2I) synthesis capabilities of diffusion models, they often struggle to understand compositional relationships between objects and attributes, especially in complex settings. Existing solutions have…

计算机视觉与模式识别 · 计算机科学 2025-04-29 Evans Xu Han , Linghao Jin , Xiaofeng Liu , Paul Pu Liang

Current methods for 3D scene reconstruction from sparse posed images employ intermediate 3D representations such as neural fields, voxel grids, or 3D Gaussians, to achieve multi-view consistent scene appearance and geometry. In this paper…

计算机视觉与模式识别 · 计算机科学 2025-02-03 Vitor Guizilini , Muhammad Zubair Irshad , Dian Chen , Greg Shakhnarovich , Rares Ambrus

Reconstructing 3D assets from images, known as inverse rendering (IR), remains a challenging task due to its ill-posed nature. 3D Gaussian Splatting (3DGS) has demonstrated impressive capabilities for novel view synthesis (NVS) tasks.…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Hanxiao Sun , YuPeng Gao , Jin Xie , Jian Yang , Beibei Wang

Image composition refers to inserting a foreground object into a background image to obtain a composite image. In this work, we focus on generating plausible shadows for the inserted foreground object to make the composite image more…

计算机视觉与模式识别 · 计算机科学 2024-01-05 Xinhao Tao , Junyan Cao , Yan Hong , Li Niu

Modern computer vision algorithms have brought significant advancement to 3D geometry reconstruction. However, illumination and material reconstruction remain less studied, with current approaches assuming very simplified models for…

计算机视觉与模式识别 · 计算机科学 2019-03-19 Dejan Azinović , Tzu-Mao Li , Anton Kaplanyan , Matthias Nießner

Mixed Reality scene relighting, where virtual changes to lighting conditions realistically interact with physical objects, producing authentic illumination and shadows, can be used in a variety of applications. One such application in real…

图形学 · 计算机科学 2025-08-22 Hanwen Zhao , John Akers , Baback Elmieh , Ira Kemelmacher-Shlizerman

Recent works in inverse rendering have shown promise in using multi-view images of an object to recover shape, albedo, and materials. However, the recovered components often fail to render accurately under new lighting conditions due to the…

计算机视觉与模式识别 · 计算机科学 2025-04-03 Yehonathan Litman , Or Patashnik , Kangle Deng , Aviral Agrawal , Rushikesh Zawar , Fernando De la Torre , Shubham Tulsiani

Video camouflaged object detection (VCOD) is challenging due to dynamic environments. Existing methods face two main issues: (1) SAM-based methods struggle to separate camouflaged object edges due to model freezing, and (2) MLLM-based…

计算机视觉与模式识别 · 计算机科学 2025-09-09 Hua Zhang , Changjiang Luo , Ruoyu Chen

High-quality image acquisition in real-world environments remains challenging due to complex illumination variations and inherent limitations of camera imaging pipelines. These issues are exacerbated in multi-view capture, where differences…

计算机视觉与模式识别 · 计算机科学 2026-02-23 Ziteng Cui , Shuhong Liu , Xiaoyu Dong , Xuangeng Chu , Lin Gu , Ming-Hsuan Yang , Tatsuya Harada

We introduce a framework that enables both multi-view character consistency and 3D camera control in video diffusion models through a novel customization data pipeline. We train the character consistency component with recorded volumetric…

Gaussian Splatting has become a popular technique for various 3D Computer Vision tasks, including novel view synthesis, scene reconstruction, and dynamic scene rendering. However, the challenge of natural-looking object insertion, where the…

计算机视觉与模式识别 · 计算机科学 2026-05-04 Vsevolod Skorokhodov , Nikita Durasov , Pascal Fua

Generating images from text involving complex and novel object arrangements remains a significant challenge for current text-to-image (T2I) models. Although prior layout-based methods improve object arrangements using spatial constraints…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Zeeshan Khan , Shizhe Chen , Cordelia Schmid

We propose RelitLRM, a Large Reconstruction Model (LRM) for generating high-quality Gaussian splatting representations of 3D objects under novel illuminations from sparse (4-8) posed images captured under unknown static lighting. Unlike…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Tianyuan Zhang , Zhengfei Kuang , Haian Jin , Zexiang Xu , Sai Bi , Hao Tan , He Zhang , Yiwei Hu , Milos Hasan , William T. Freeman , Kai Zhang , Fujun Luan

Infrared-visible image fusion aims to integrate infrared and visible information into a single fused image. Existing 2D fusion methods focus on fusing images from fixed camera viewpoints, neglecting a comprehensive understanding of complex…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Chao Yang , Deshui Miao , Chao Tian , Guoqing Zhu , Yameng Gu , Zhenyu He