中文
相关论文

相关论文: RenderWorld: World Model with Self-Supervised 3D L…

200 篇论文

Neural rendering, particularly 3D Gaussian Splatting (3DGS), has evolved rapidly and become a key component for building world models. However, existing viewer solutions remain fragmented, heavy, or constrained by legacy pipelines,…

Learning world models can teach an agent how the world works in an unsupervised manner. Even though it can be viewed as a special case of sequence modeling, progress for scaling world models on robotic applications such as autonomous…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Lunjun Zhang , Yuwen Xiong , Ze Yang , Sergio Casas , Rui Hu , Raquel Urtasun

We present Exact Volumetric Ellipsoid Rendering (EVER), a method for real-time differentiable emission-only volume rendering. Unlike recent rasterization based approach by 3D Gaussian Splatting (3DGS), our primitive based representation…

计算机视觉与模式识别 · 计算机科学 2025-07-25 Alexander Mai , Peter Hedman , George Kopanas , Dor Verbin , David Futschik , Qiangeng Xu , Falko Kuester , Jonathan T. Barron , Yinda Zhang

Safe and scalable deployment of end-to-end (E2E) autonomous driving requires extensive and diverse data, particularly safety-critical events. Existing data are mostly generated from simulators with a significant sim-to-real gap or collected…

机器人学 · 计算机科学 2025-09-18 Jiawei Wang , Haowei Sun , Xintao Yan , Shuo Feng , Jun Gao , Henry X. Liu

Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D segmentation. However, the supervision signals from foundation models are not…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Tsuheng Hsu , Guiyu Liu , Juho Kannala , Janne Heikkilä

In this work, we consider the problem of learning end to end perception to control for ground vehicles solely from aerial imagery. Photogrammetric simulators allow the synthesis of novel views through the transformation of pre-generated…

机器人学 · 计算机科学 2024-10-21 Varun Murali , Guy Rosman , Sertac Karaman , Daniela Rus

End-to-end autonomous driving has gained significant attention for its potential to learn robust behavior in interactive scenarios and scale with data. Popular architectures often build on separate modules for perception and planning…

机器人学 · 计算机科学 2026-03-17 David Holtz , Niklas Hanselmann , Simon Doll , Marius Cordts , Bernt Schiele

In this paper, we propose a Neural Radiance Fields (NeRF) based framework, referred to as Novel View Synthesis Framework (NVSF). It jointly learns the implicit neural representation of space and time-varying scene for both LiDAR and Camera.…

计算机视觉与模式识别 · 计算机科学 2025-06-26 Gaurav Sharma , Ravi Kothari , Josef Schmid

Accurate and realistic 3D scene reconstruction enables the lifelike creation of autonomous driving simulation environments. With advancements in 3D Gaussian Splatting (3DGS), previous studies have applied it to reconstruct complex dynamic…

计算机视觉与模式识别 · 计算机科学 2025-02-21 Yedong Shen , Xinran Zhang , Yifan Duan , Shiqi Zhang , Heng Li , Yilong Wu , Jianmin Ji , Yanyong Zhang

Steering estimation is a critical task in autonomous driving, traditionally relying on 2D image-based models. In this work, we explore the advantages of incorporating 3D spatial information through hybrid architectures that combine 3D…

计算机视觉与模式识别 · 计算机科学 2025-03-24 Fouad Makiyeh , Huy-Dung Nguyen , Patrick Chareyre , Ramin Hasani , Marc Blanchon , Daniela Rus

3D Gaussian Splatting has emerged as a powerful representation of geometry and appearance for RGB-only dense Simultaneous Localization and Mapping (SLAM), as it provides a compact dense map representation while enabling efficient and…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Erik Sandström , Keisuke Tateno , Michael Oechsle , Michael Niemeyer , Luc Van Gool , Martin R. Oswald , Federico Tombari

Autonomous parking is a crucial task in the intelligent driving field. Traditional parking algorithms are usually implemented using rule-based schemes. However, these methods are less effective in complex parking scenarios due to the…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Changze Li , Ziheng Ji , Zhe Chen , Tong Qin , Ming Yang

This paper aims to tackle the problem of modeling dynamic urban streets for autonomous driving scenes. Recent methods extend NeRF by incorporating tracked vehicle poses to animate vehicles, enabling photo-realistic view synthesis of dynamic…

计算机视觉与模式识别 · 计算机科学 2024-08-20 Yunzhi Yan , Haotong Lin , Chenxu Zhou , Weijie Wang , Haiyang Sun , Kun Zhan , Xianpeng Lang , Xiaowei Zhou , Sida Peng

This paper presents an effective solution for view extrapolation in autonomous driving scenarios. Recent approaches focus on generating shifted novel view images from given viewpoints using diffusion models. However, these methods heavily…

计算机视觉与模式识别 · 计算机科学 2026-01-01 Yuang Jia , Jinlong Wang , Jiayi Zhao , Chunlam Li , Shunzhou Wang , Wei Gao

Recent advancements in 3D Gaussian Splatting have significantly improved the efficiency and quality of dense semantic SLAM. However, previous methods are generally constrained by limited-category pre-trained classifiers and implicit…

计算机视觉与模式识别 · 计算机科学 2025-03-04 Dianyi Yang , Yu Gao , Xihan Wang , Yufeng Yue , Yi Yang , Mengyin Fu

In this paper a semi-supervised deep framework is proposed for the problem of 3D shape inverse rendering from a single 2D input image. The main structure of proposed framework consists of unsupervised pre-trained components which…

计算机视觉与模式识别 · 计算机科学 2017-11-17 Shima Kamyab , S. Zohreh Azimifar

3D-consistent image generation from a single 2D semantic label is an important and challenging research topic in computer graphics and computer vision. Although some related works have made great progress in this field, most of the existing…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Bo Li , Yi-ke Li , Zhi-fen He , Bin Liu , Yun-Kun Lai

An accurate understanding of a self-driving vehicle's surrounding environment is crucial for its navigation system. To enhance the effectiveness of existing algorithms and facilitate further research, it is essential to provide…

计算机视觉与模式识别 · 计算机科学 2023-11-14 Abtin Mahyar , Hossein Motamednia , Dara Rahmati

3D semantic occupancy prediction is crucial for autonomous driving. While multi-modal fusion improves accuracy over vision-only methods, it typically relies on computationally expensive dense voxel or BEV tensors. We present Gau-Occ, a…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Chengxin Lv , Yihui Li , Hongyu Yang , YunHong Wang

Forecasting the evolution of 3D scenes and generating unseen scenarios via occupancy-based world models offers substantial potential for addressing corner cases in autonomous driving systems. While tokenization has revolutionized image and…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Zhimin Liao , Ping Wei , Ruijie Zhang , Shuaijia Chen , Haoxuan Wang , Ziyang Ren
‹ 上一页 1 8 9 10 下一页 ›