中文
相关论文

相关论文: Physically Aware 360$^\circ$ View Generation from …

200 篇论文

This is a technical report on the 360-degree panoramic image generation task based on diffusion models. Unlike ordinary 2D images, 360-degree panoramic images capture the entire $360^\circ\times 180^\circ$ field of view. So the rightmost…

计算机视觉与模式识别 · 计算机科学 2023-11-23 Mengyang Feng , Jinlin Liu , Miaomiao Cui , Xuansong Xie

Medical image synthesis is crucial for alleviating data scarcity and privacy constraints. However, fine-tuning general text-to-image (T2I) models remains challenging, mainly due to the significant modality gap between complex visual details…

计算机视觉与模式识别 · 计算机科学 2026-03-12 Xin Huang , Junjie Liang , Qingshan Hou , Peng Cao , Jinzhu Yang , Xiaoli Liu , Osmar R. Zaiane

Recent advancements in computer vision have led to a renewed interest in developing assistive technologies for individuals with visual impairments. Although extensive research has been conducted in the field of computer vision-based…

计算机视觉与模式识别 · 计算机科学 2024-11-19 Inpyo Song , Sanghyeon Lee , Minjun Joo , Jangwon Lee

We introduce DiffPhysCam, a differentiable camera simulator designed to support robotics and embodied AI applications by enabling gradient-based optimization in visual perception pipelines. Generating synthetic images that closely mimic…

图形学 · 计算机科学 2025-08-13 Bo-Hsun Chen , Nevindu M. Batagoda , Dan Negrut

Recent implicit neural representations have shown great results for novel view synthesis. However, existing methods require expensive per-scene optimization from many views hence limiting their application to real-world unbounded urban…

计算机视觉与模式识别 · 计算机科学 2023-08-25 Muhammad Zubair Irshad , Sergey Zakharov , Katherine Liu , Vitor Guizilini , Thomas Kollar , Adrien Gaidon , Zsolt Kira , Rares Ambrus

Virtual reality and augmented reality (XR) bring increasing demand for 3D content. However, creating high-quality 3D content requires tedious work that a human expert must do. In this work, we study the challenging task of lifting a single…

计算机视觉与模式识别 · 计算机科学 2023-04-04 Dejia Xu , Yifan Jiang , Peihao Wang , Zhiwen Fan , Yi Wang , Zhangyang Wang

We aim to obtain an interpretable, expressive, and disentangled scene representation that contains comprehensive structural and textural information for each object. Previous scene representations learned by neural networks are often…

计算机视觉与模式识别 · 计算机科学 2018-12-19 Shunyu Yao , Tzu Ming Harry Hsu , Jun-Yan Zhu , Jiajun Wu , Antonio Torralba , William T. Freeman , Joshua B. Tenenbaum

This paper addresses the limitations of neural rendering-based multi-view surface reconstruction methods, which require an additional mesh extraction step that is inconvenient and would produce poor-quality surfaces with mesh aliasing,…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Qitong Zhang , Jieqing Feng

Point scene understanding is a challenging task to process real-world scene point cloud, which aims at segmenting each object, estimating its pose, and reconstructing its mesh simultaneously. Recent state-of-the-art method first segments…

计算机视觉与模式识别 · 计算机科学 2024-03-26 Xiaoxuan Yu , Hao Wang , Weiming Li , Qiang Wang , Soonyong Cho , Younghun Sung

While the proposal of the Tri-plane representation has advanced the development of the 3D-aware image generative models, problems rooted in its inherent structure, such as multi-face artifacts caused by sharing the same features in…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Ru Jia , Xiaozhuang Ma , Jianji Wang , Nanning Zheng

360 images, with a field-of-view (FoV) of 180x360, provide immersive and realistic environments for emerging virtual reality (VR) applications, such as virtual tourism, where users desire to create diverse panoramic scenes from a narrow FoV…

计算机视觉与模式识别 · 计算机科学 2024-01-22 Hao Ai , Zidong Cao , Haonan Lu , Chen Chen , Jian Ma , Pengyuan Zhou , Tae-Kyun Kim , Pan Hui , Lin Wang

Generating complete digital twins from videos requires precise camera control, global scene coverage, and strict spatial-temporal consistency constraints that remain challenging for perspective video generators due to their limited field of…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Ting-Hsuan Chen , Ying-Huan Chen , Tao Tu , Jie-Ying Lee , Cho-Ying Wu , Fangzhou Lin , Hengyuan Zhang , David Paz , Xinyu Huang , Yuliang Guo , Yu-Lun Liu , Yue Wang , Liu Ren

Neural Radiance Fields (NeRF) coupled with GANs represent a promising direction in the area of 3D reconstruction from a single view, owing to their ability to efficiently model arbitrary topologies. Recent work in this area, however, has…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Dario Pavllo , David Joseph Tan , Marie-Julie Rakotosaona , Federico Tombari

Endoluminal endoscopic procedures are essential for diagnosing colorectal cancer and other severe conditions in the digestive tract, urogenital system, and airways. 3D reconstruction and novel-view synthesis from endoscopic images are…

计算机视觉与模式识别 · 计算机科学 2025-07-10 Joanna Kaleta , Weronika Smolak-Dyżewska , Dawid Malarz , Diego Dall'Alba , Przemysław Korzeniowski , Przemysław Spurek

In robot-assisted minimally invasive surgery, high-fidelity dynamic endoscopic scene reconstruction and simulation are crucial to enhancing downstream tasks and advancing surgical outcomes. However, existing methods primarily focus on…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Changjing Liu , Yiming Huang , Long Bai , Beilei Cui , Hongliang Ren

Novel view synthesis from monocular videos of dynamic scenes with unknown camera poses remains a fundamental challenge in computer vision and graphics. While recent advances in 3D representations such as Neural Radiance Fields (NeRF) and 3D…

计算机视觉与模式识别 · 计算机科学 2025-11-10 Mengqi Guo , Bo Xu , Yanyan Li , Gim Hee Lee

We introduce 4DGS360, a diffusion-free framework for 360$^{\circ}$ dynamic object reconstruction from casual monocular video. Existing methods often fail to reconstruct consistent 360$^{\circ}$ geometry, as their heavy reliance on 2D-native…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Jae Won Jang , Yeonjin Chang , Wonsik Shin , Juhwan Cho , Nojun Kwak

Underwater scene reconstruction is essential for immersive exploration of aquatic environments, yet remains challenging due to complex participating-media effects such as absorption and scattering, as well as the limited field of view (FoV)…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Jiangbei Hu , Weichao Song , Shibo Yu , Mohan Wang , Zihan Yi , Rui Wu , Mingkang Xiang , Na Lei , Shengfa Wang , Zhongxuan Luo , Ying He

Disentangling anatomical and contrast information from medical images has gained attention recently, demonstrating benefits for various image analysis tasks. Current methods learn disentangled representations using either paired multi-modal…

图像与视频处理 · 电气工程与系统科学 2022-05-11 Lianrui Zuo , Yihao Liu , Yuan Xue , Shuo Han , Murat Bilgel , Susan M. Resnick , Jerry L. Prince , Aaron Carass

Despite recent advances in single-object front-facing inpainting using NeRF and 3D Gaussian Splatting (3DGS), inpainting in complex 360{\deg} scenes remains largely underexplored. This is primarily due to three key challenges: (i)…

计算机视觉与模式识别 · 计算机科学 2025-11-11 Shaoxiang Wang , Shihong Zhang , Christen Millerdurai , Rüdiger Westermann , Didier Stricker , Alain Pagani