中文
相关论文

相关论文: Gimbal360: Differentiable Auto-Leveling for Canoni…

200 篇论文

Feed-forward 3D Gaussian Splatting (3DGS) has shown great promise for real-time novel view synthesis, but its application to panoramic imagery remains challenging. Existing methods often rely on multi-view cost volumes for geometric…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Qiwei Wang , Xianghui Ze , Jingyi Yu , Yujiao Shi

Obstacle avoidance in unmanned aerial vehicles (UAVs), as a fundamental capability, has gained increasing attention with the growing focus on spatial intelligence. However, current obstacle-avoidance methods mainly depend on limited…

机器人学 · 计算机科学 2026-03-09 Xiangkai Zhang , Dizhe Zhang , WenZhuo Cao , Zhaoliang Wan , Yingjie Niu , Lu Qi , Xu Yang , Zhiyong Liu

Recent illumination estimation methods have focused on enhancing the resolution and improving the quality and diversity of the generated textures. However, few have explored tailoring the neural network architecture to the Equirectangular…

计算机视觉与模式识别 · 计算机科学 2024-10-18 Jack Hilliard , Adrian Hilton , Jean-Yves Guillemaut

We present PanoPlane, an approach for high-fidelity sparse-view indoor novel view synthesis that reconstructs closed room geometry via panoramic scene completion. Unlike perspective-based methods that generate training views from limited…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Adil Qureshi , Dongki Jung , Jaehoon Choi , Dinesh Manocha

3D scene reconstruction is fundamental for spatial intelligence applications such as AR, robotics, and digital twins. Traditional multi-view stereo struggles with sparse viewpoints or low-texture regions, while neural rendering approaches,…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Jiaqi Yao , Zhongmiao Yan , Jingyi Xu , Songpengcheng Xia , Yan Xiang , Ling Pei

The rapid advancement of diffusion models holds the promise of revolutionizing the application of VR and AR technologies, which typically require scene-level 4D assets for user experience. Nonetheless, existing diffusion models…

计算机视觉与模式识别 · 计算机科学 2025-05-14 Haiyang Zhou , Wangbo Yu , Jiawen Guan , Xinhua Cheng , Yonghong Tian , Li Yuan

3D-aware GANs aim to synthesize realistic 3D scenes such that they can be rendered in arbitrary perspectives to produce images. Although previous methods produce realistic images, they suffer from unstable training or degenerate solutions…

计算机视觉与模式识别 · 计算机科学 2023-11-13 Minjung Shin , Yunji Seo , Jeongmin Bae , Young Sun Choi , Hyunsu Kim , Hyeran Byun , Youngjung Uh

Image generation models trained on large datasets can synthesize high-quality images but often produce spatially inconsistent and distorted images due to limited information about the underlying structures and spatial layouts. In this work,…

计算机视觉与模式识别 · 计算机科学 2025-11-27 Hyundo Lee , Suhyung Choi , Inwoo Hwang , Byoung-Tak Zhang

Object-level manipulation, relocating or reorienting objects in images or videos while preserving scene realism, is central to film post-production, AR, and creative editing. Yet existing methods struggle to jointly achieve three core…

计算机视觉与模式识别 · 计算机科学 2026-02-13 Penghui Ruan , Bojia Zi , Xianbiao Qi , Youze Huang , Rong Xiao , Pichao Wang , Jiannong Cao , Yuhui Shi

Disentangling visual layers in real-world images is a persistent challenge in vision and graphics, as such layers often involve non-linear and globally coupled interactions, including shading, reflection, and perspective distortion. In this…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Zheng Gu , Min Lu , Zhida Sun , Dani Lischinski , Daniel Cohen-Or , Hui Huang

Panoramic image stitching provides a unified, wide-angle view of a scene that extends beyond the camera's field of view. Stitching frames of a panning video into a panoramic photograph is a well-understood problem for stationary scenes, but…

计算机视觉与模式识别 · 计算机科学 2024-10-29 Jingwei Ma , Erika Lu , Roni Paiss , Shiran Zada , Aleksander Holynski , Tali Dekel , Brian Curless , Michael Rubinstein , Forrester Cole

In this paper, we address panoramic semantic segmentation which is under-explored due to two critical challenges: (1) image distortions and object deformations on panoramas; (2) lack of semantic annotations in the 360{\deg} imagery. To…

计算机视觉与模式识别 · 计算机科学 2024-06-03 Jiaming Zhang , Kailun Yang , Hao Shi , Simon Reiß , Kunyu Peng , Chaoxiang Ma , Haodong Fu , Philip H. S. Torr , Kaiwei Wang , Rainer Stiefelhagen

Geometric alignment appears in a variety of applications, ranging from domain adaptation, optimal transport, and normalizing flows in machine learning; optical flow and learned augmentation in computer vision and deformable registration…

计算机视觉与模式识别 · 计算机科学 2021-10-27 Steffen Czolbe , Aasa Feragen , Oswin Krause

Visual servoing, the method of controlling robot motion through feedback from visual sensors, has seen significant advancements with the integration of optical flow-based methods. However, its application remains limited by inherent…

The richness of natural images makes the quest for optimal representations in image processing and computer vision challenging. The latter observation has not prevented the design of image representations, which trade off between efficiency…

计算机视觉与模式识别 · 计算机科学 2011-10-27 Laurent Jacques , Laurent Duval , Caroline Chaux , Gabriel Peyré

Driven by the demand for spatial intelligence and holistic scene perception, omnidirectional images (ODIs), which provide a complete 360\textdegree{} field of view, are receiving growing attention across diverse applications such as virtual…

计算机视觉与模式识别 · 计算机科学 2025-09-10 Xin Lin , Xian Ge , Dizhe Zhang , Zhaoliang Wan , Xianshun Wang , Xiangtai Li , Wenjie Jiang , Bo Du , Dacheng Tao , Ming-Hsuan Yang , Lu Qi

The panorama image can simultaneously demonstrate complete information of the surrounding environment and has many advantages in virtual tourism, games, robotics, etc. However, the progress of panorama depth estimation cannot completely…

计算机视觉与模式识别 · 计算机科学 2022-12-06 Qingsong Yan , Qiang Wang , Kaiyong Zhao , Bo Li , Xiaowen Chu , Fei Deng

Predicting panoramic indoor lighting from a single perspective image is a fundamental but highly ill-posed problem in computer vision and graphics. To achieve locale-aware and robust prediction, this problem can be decomposed into three…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Jiayang Bai , Zhen He , Shan Yang , Jie Guo , Zhenyu Chen , Yan Zhang , Yanwen Guo

Single-channel 3D reconstruction is widely used in fields such as robotics and medical imaging. While these methods are good at reconstructing 3D geometry, their outputs are typically uncolored 3D models, making 3D colorization necessary…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Yeonjin Chang , Juhwan Cho , Seunghyeon Seo , Wonsik Shin , Nojun Kwak

Image tiling -- the seamless connection of disparate images to create a coherent visual field -- is crucial for applications such as texture creation, video game asset development, and digital art. Traditionally, tiles have been constructed…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Or Madar , Ohad Fried