English
Related papers

Related papers: Under One Sun: Multi-Object Generative Perception …

200 papers

This paper introduces a versatile paradigm for integrating multi-view reflectance (optional) and normal maps acquired through photometric stereo. Our approach employs a pixel-wise joint re-parameterization of reflectance and normal,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-01 Baptiste Brument , Robin Bruneau , Yvain Quéau , Jean Mélou , François Bernard Lauze , Jean-Denis , Jean-Denis Durou , Lilian Calvet

We present a novel differentiable rendering framework for joint geometry, material, and lighting estimation from multi-view images. In contrast to previous methods which assume a simplified environment map or co-located flashlights, in this…

Computer Vision and Pattern Recognition · Computer Science 2023-03-31 Jingyang Zhang , Yao Yao , Shiwei Li , Jingbo Liu , Tian Fang , David McKinnon , Yanghai Tsin , Long Quan

Metasurface-generated holography has emerged as a promising route for fully reproducing vivid scenes by manipulating the optical properties of light using ultra-compact devices. However, achieving multiple holographic images using a single…

3D object generation from a single image involves estimating the full 3D geometry and texture of unseen views from an unposed RGB image captured in the wild. Accurately reconstructing an object's complete 3D structure and texture has…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Hritam Basak , Hadi Tabatabaee , Shreekant Gayaka , Ming-Feng Li , Xin Yang , Cheng-Hao Kuo , Arnie Sen , Min Sun , Zhaozheng Yin

A fundamental problem in computer vision is that of inferring the intrinsic, 3D structure of the world from flat, 2D images of that world. Traditional methods for recovering scene properties such as shape, reflectance, or illumination rely…

Computer Vision and Pattern Recognition · Computer Science 2020-10-09 Jonathan T. Barron , Jitendra Malik

Recent advances in 3D scene generation produce visually appealing output, but current representations hinder artists' workflows that require modifiable 3D textured mesh scenes for visual effects and game development. Despite significant…

Computer Vision and Pattern Recognition · Computer Science 2025-12-22 Tobias Sautter , Jan-Niklas Dihlmann , Hendrik P. A. Lensch

The reflectance field of a face describes the reflectance properties responsible for complex lighting effects including diffuse, specular, inter-reflection and self shadowing. Most existing methods for estimating the face reflectance from a…

Computer Vision and Pattern Recognition · Computer Science 2020-08-25 Mallikarjun B R. , Ayush Tewari , Tae-Hyun Oh , Tim Weyrich , Bernd Bickel , Hans-Peter Seidel , Hanspeter Pfister , Wojciech Matusik , Mohamed Elgharib , Christian Theobalt

A unified diffusion framework for multi-modal generation and understanding has the transformative potential to achieve seamless and controllable image diffusion and other cross-modal tasks. In this paper, we introduce MMGen, a unified…

Computer Vision and Pattern Recognition · Computer Science 2025-03-27 Jiepeng Wang , Zhaoqing Wang , Hao Pan , Yuan Liu , Dongdong Yu , Changhu Wang , Wenping Wang

The reflection superposition phenomenon is complex and widely distributed in the real world, which derives various simplified linear and nonlinear formulations of the problem. In this paper, based on the investigation of the weaknesses of…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Qiming Hu , Xiaojie Guo

Reconstructing the shape and appearance of real-world objects using measured 2D images has been a long-standing problem in computer vision. In this paper, we introduce a new analysis-by-synthesis technique capable of producing high-quality…

Computer Vision and Pattern Recognition · Computer Science 2021-06-28 Fujun Luan , Shuang Zhao , Kavita Bala , Zhao Dong

We propose a method to detect and reconstruct multiple 3D objects from a single RGB image. The key idea is to optimize for detection, alignment and shape jointly over all objects in the RGB image, while focusing on realistic and physically…

Computer Vision and Pattern Recognition · Computer Science 2021-06-23 Francis Engelmann , Konstantinos Rematas , Bastian Leibe , Vittorio Ferrari

We present Im2SurfTex, a method that generates textures for input 3D shapes by learning to aggregate multi-view image outputs produced by 2D image diffusion models onto the shapes' texture space. Unlike existing texture generation…

Graphics · Computer Science 2025-12-11 Yiangos Georgiou , Marios Loizou , Melinos Averkiou , Evangelos Kalogerakis

Polygon representation learning is essential for diverse applications, encompassing tasks such as shape coding, building pattern classification, and geographic question answering. While recent years have seen considerable advancements in…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Dazhou Yu , Yuntong Hu , Yun Li , Liang Zhao

Despite remarkable progress in image translation, the complex scene with multiple discrepant objects remains a challenging problem. The translated images have low fidelity and tiny objects in fewer details causing unsatisfactory performance…

Computer Vision and Pattern Recognition · Computer Science 2022-12-26 Liyun Zhang , Photchara Ratsamee , Bowen Wang , Zhaojie Luo , Yuki Uranishi , Manabu Higashida , Haruo Takemura

We present a unified and compact scene representation for robotics, where each object in the scene is depicted by a latent code capturing geometry and appearance. This representation can be decoded for various tasks such as novel view…

Object-centric reconstruction seeks to recover the 3D structure of a scene through composition of independent objects. While this independence can simplify modeling, it discards strong signals that could improve reconstruction, notably…

Computer Vision and Pattern Recognition · Computer Science 2026-03-30 Qirui Wu , Yawar Siddiqui , Duncan Frost , Samir Aroudj , Armen Avetisyan , Richard Newcombe , Angel X. Chang , Jakob Engel , Henry Howard-Jenkins

We present InvRGB+L, a novel inverse rendering model that reconstructs large, relightable, and dynamic scenes from a single RGB+LiDAR sequence. Conventional inverse graphics methods rely primarily on RGB observations and use LiDAR mainly…

Computer Vision and Pattern Recognition · Computer Science 2025-07-24 Xiaoxue Chen , Bhargav Chandaka , Chih-Hao Lin , Ya-Qin Zhang , David Forsyth , Hao Zhao , Shenlong Wang

Lighting has a strong influence on visual appearance, yet understanding and representing lighting in images remains notoriously difficult. Various lighting representations exist, such as environment maps, irradiance, spherical harmonics, or…

Computer Vision and Pattern Recognition · Computer Science 2026-03-05 Zitian Zhang , Iliyan Georgiev , Michael Fischer , Yannick Hold-Geoffroy , Jean-François Lalonde , Valentin Deschaintre

We propose spatial polarization multiplexing (SPM) for joint sensing of shape and reflectance of a static or dynamic deformable object, which is also invisible to the naked eye. Past structured-light methods are limited to shape acquisition…

Computer Vision and Pattern Recognition · Computer Science 2026-01-07 Tomoki Ichikawa , Ryo Kawahara , Ko Nishino

Multipath effects significantly influence the quality of microwave imaging in highly reflective environments, while the physical measurement aperture size constrains resolution. It is shown that by exploiting multipath reflections, improved…

Image and Video Processing · Electrical Eng. & Systems 2026-05-06 Quanfeng Wang , Mei Song Tong , Thomas F. Eibert