中文
相关论文

相关论文: HumanSplat: Generalizable Single-Image Human Gauss…

200 篇论文

We present Splat-SAP, a feed-forward approach to render novel views of human-centered scenes from binocular cameras with large sparsity. Gaussian Splatting has shown its promising potential in rendering tasks, but it typically necessitates…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Boyao Zhou , Shunyuan Zheng , Zhanfeng Liao , Zihan Ma , Hanzhang Tu , Boning Liu , Yebin Liu

Reconstructing intricate, ever-changing environments remains a central ambition in computer vision, yet existing solutions often crumble before the complexity of real-world dynamics. We present DynaSplat, an approach that extends Gaussian…

计算机视觉与模式识别 · 计算机科学 2025-06-12 Junli Deng , Ping Shi , Qipei Li , Jinyang Guo

Recently, generalizable feed-forward methods based on 3D Gaussian Splatting have gained significant attention for their potential to reconstruct 3D scenes using finite resources. These approaches create a 3D radiance field, parameterized by…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Wonseok Roh , Hwanhee Jung , Jong Wook Kim , Seunggwan Lee , Innfarn Yoo , Andreas Lugmayr , Seunggeun Chi , Karthik Ramani , Sangpil Kim

We introduce pixelSplat, a feed-forward model that learns to reconstruct 3D radiance fields parameterized by 3D Gaussian primitives from pairs of images. Our model features real-time and memory-efficient rendering for scalable training as…

计算机视觉与模式识别 · 计算机科学 2024-04-08 David Charatan , Sizhe Li , Andrea Tagliasacchi , Vincent Sitzmann

The efficient spatial allocation of primitives serves as the foundation of 3D Gaussian Splatting, as it directly dictates the synergy between representation compactness, reconstruction speed, and rendering fidelity. Previous solutions,…

计算机视觉与模式识别 · 计算机科学 2026-04-20 Roni Itkin , Noam Issachar , Yehonatan Keypur , Xingyu Chen , Anpei Chen , Sagie Benaim

While existing feed-forward Gaussian splatting models offer computational efficiency and can generalize to sparse view settings, their performance is fundamentally constrained by relying on a single forward pass for inference. We propose…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Haofei Xu , Daniel Barath , Andreas Geiger , Marc Pollefeys

Reconstructing dynamic humans together with static scenes from monocular videos remains difficult, especially under fast motion, where RGB frames suffer from motion blur. Event cameras exhibit distinct advantages, e.g., microsecond temporal…

计算机视觉与模式识别 · 计算机科学 2025-09-24 Xiaoting Yin , Hao Shi , Kailun Yang , Jiajun Zhai , Shangwei Guo , Lin Wang , Kaiwei Wang

Accurate 3D reconstruction in degraded imaging conditions remains a key challenge in photogrammetry and neural rendering. In underwater environments, spatially varying visibility caused by scattering, attenuation, and sparse observations…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Zhuodong Jiang , Haoran Wang , Guoxi Huang , Brett Seymour , Nantheera Anantrasirichai

The semantic synthesis of unseen scenes from multiple viewpoints is crucial for research in 3D scene understanding. Current methods are capable of rendering novel-view images and semantic maps by reconstructing generalizable Neural Radiance…

图形学 · 计算机科学 2025-05-09 Feng Xiao , Hongbin Xu , Wanlin Liang , Wenxiong Kang

In this work, we tackle the task of learning 3D human Gaussians from a single image, focusing on recovering detailed appearance and geometry including unobserved regions. We introduce a single-view generalizable Human Gaussian Model (HGM),…

计算机视觉与模式识别 · 计算机科学 2025-05-09 Jinnan Chen , Chen Li , Jianfeng Zhang , Lingting Zhu , Buzhen Huang , Hanlin Chen , Gim Hee Lee

We introduce HART, a unified framework for sparse-view human reconstruction. Given a small set of uncalibrated RGB images of a person as input, it outputs a watertight clothed mesh, the aligned SMPL-X body mesh, and a Gaussian-splat…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Xiyi Chen , Shaofei Wang , Marko Mihajlovic , Taewon Kang , Sergey Prokudin , Ming Lin

A few recent works explored incorporating geometric priors to regularize the optimization of Gaussian splatting, further improving its performance. However, those early studies mainly focused on the use of low-order geometric priors (e.g.,…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Yangming Li , Chaoyu Liu , Lihao Liu , Simon Masnou , Carola-Bibiane Schönlieb

We propose PoseGaussian, a pose-guided Gaussian Splatting framework for high-fidelity human novel view synthesis. Human body pose serves a dual purpose in our design: as a structural prior, it is fused with a color encoder to refine depth…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Ju Shen , Chen Chen , Tam V. Nguyen , Vijayan K. Asari

Recent progress in feed-forward 3D Gaussian Splatting (3DGS) has notably improved rendering quality. However, the spatially uniform and highly redundant 3DGS map generated by previous feed-forward 3DGS methods limits their integration into…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Zicheng Zhang , Xiangting Meng , Ke Wu , Wenchao Ding

This work addresses the problem of real-time rendering of photorealistic human body avatars learned from multi-view videos. While the classical approaches to model and render virtual humans generally use a textured mesh, recent research has…

计算机视觉与模式识别 · 计算机科学 2024-03-29 Arthur Moreau , Jifei Song , Helisa Dhamo , Richard Shaw , Yiren Zhou , Eduardo Pérez-Pellitero

We aim to address sparse-view reconstruction of a 3D scene by leveraging priors from large-scale vision models. While recent advancements such as 3D Gaussian Splatting (3DGS) have demonstrated remarkable successes in 3D reconstruction,…

计算机视觉与模式识别 · 计算机科学 2025-07-29 Hanyang Yu , Xiaoxiao Long , Ping Tan

Single-view 3D human reconstruction has garnered significant attention in recent years. Despite numerous advancements, prior research has concentrated on reconstructing 3D models from clear, close-up images of individual subjects, often…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Yizheng Song , Yiyu Zhuang , Qipeng Xu , Haixiang Wang , Jiahe Zhu , Jing Tian , Siyu Zhu , Hao Zhu

Feed-forward 3D reconstruction from sparse, low-resolution (LR) images is a crucial capability for real-world applications, such as autonomous driving and embodied AI. However, existing methods often fail to recover fine texture details.…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Xinyuan Hu , Changyue Shi , Chuxiao Yang , Minghao Chen , Jiajun Ding , Tao Wei , Chen Wei , Zhou Yu , Min Tan

Feed-forward 3D reconstruction offers substantial runtime advantages over per-scene optimization, which remains slow at inference and often fragile under sparse views. However, existing feed-forward methods still have potential for further…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Tianyu Chen , Wei Xiang , Kang Han , Yu Lu , Di Wu , Gaowen Liu , Ramana Rao Kompella

Recent advances in generative AI have accelerated the production of ultra-high-resolution visual content, posing significant challenges for efficient compression and real-time decoding on end-user devices. Inspired by 3D Gaussian Splatting,…

计算机视觉与模式识别 · 计算机科学 2026-01-13 Linfei Li , Lin Zhang , Zhong Wang , Ying Shen