English
Related papers

Related papers: SGAT4PASS: Spherical Geometry-Aware Transformer fo…

200 papers

Recent progress in self-supervised representation learning has resulted in models that are capable of extracting image features that are not only effective at encoding image level, but also pixel-level, semantics. These features have been…

Computer Vision and Pattern Recognition · Computer Science 2024-07-08 Octave Mariotti , Oisin Mac Aodha , Hakan Bilen

Recent advances in vision tasks (e.g., segmentation) highly depend on the availability of large-scale real-world image annotations obtained by cumbersome human labors. Moreover, the perception performance often drops significantly for new…

Computer Vision and Pattern Recognition · Computer Science 2018-07-17 Peilun Li , Xiaodan Liang , Daoyuan Jia , Eric P. Xing

Recent advances in diffusion models have demonstrated exceptional capabilities in image and video generation, further improving the effectiveness of 4D synthesis. Existing 4D generation methods can generate high-quality 4D objects or scenes…

Computer Vision and Pattern Recognition · Computer Science 2024-10-10 Bohan Zeng , Ling Yang , Siyu Li , Jiaming Liu , Zixiang Zhang , Juanxi Tian , Kaixin Zhu , Yongzhen Guo , Fu-Yun Wang , Minkai Xu , Stefano Ermon , Wentao Zhang

The ability of scene understanding has sparked active research for panoramic image semantic segmentation. However, the performance is hampered by distortion of the equirectangular projection (ERP) and a lack of pixel-wise annotations. For…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Xu Zheng , Jinjing Zhu , Yexin Liu , Zidong Cao , Chong Fu , Lin Wang

4D panoptic segmentation is a challenging but practically useful task that requires every point in a LiDAR point-cloud sequence to be assigned a semantic class label, and individual objects to be segmented and tracked over time. Existing…

Computer Vision and Pattern Recognition · Computer Science 2023-11-21 Ali Athar , Enxu Li , Sergio Casas , Raquel Urtasun

Deep learning has achieved remarkable results in 3D shape analysis by learning global shape features from the pixel-level over multiple views. Previous methods, however, compute low-level features for entire views without considering…

Computer Vision and Pattern Recognition · Computer Science 2019-05-21 Zhizhong Han , Xinhai Liu , Yu-Shen Liu , Matthias Zwicker

Existing offline feed-forward methods for joint scene understanding and reconstruction on long image streams often repeatedly perform global computation over an ever-growing set of past observations, causing runtime and GPU memory to…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Renhe Zhang , Yuyang Tan , Jingyu Gong , Zhizhong Zhang , Lizhuang Ma , Yuan Xie , Xin Tan

Common fully glazed facades and transparent objects present architectural barriers and impede the mobility of people with low vision or blindness, for instance, a path detected behind a glass door is inaccessible unless it is correctly…

Computer Vision and Pattern Recognition · Computer Science 2021-08-23 Jiaming Zhang , Kailun Yang , Angela Constantinescu , Kunyu Peng , Karin Müller , Rainer Stiefelhagen

3D point clouds are rich in geometric structure information, while 2D images contain important and continuous texture information. Combining 2D information to achieve better 3D semantic segmentation has become mainstream in 3D scene…

Computer Vision and Pattern Recognition · Computer Science 2022-12-14 Chaolong Yang , Yuyao Yan , Weiguang Zhao , Jianan Ye , Xi Yang , Amir Hussain , Kaizhu Huang

Visual navigation for autonomous agents is a core task in the fields of computer vision and robotics. Learning-based methods, such as deep reinforcement learning, have the potential to outperform the classical solutions developed for this…

Computer Vision and Pattern Recognition · Computer Science 2021-03-23 Zachary Seymour , Kowshik Thopalli , Niluthpol Mithun , Han-Pang Chiu , Supun Samarasekera , Rakesh Kumar

Camera-based 3D semantic scene completion (SSC) provides dense geometric and semantic perception for autonomous driving and robotic navigation. However, existing methods rely on a coupled encoder to deliver both semantic and geometric…

Computer Vision and Pattern Recognition · Computer Science 2026-01-16 Shiyuan Chen , Wei Sui , Bohao Zhang , Zeyd Boukhers , John See , Cong Yang

Accurate 3D reconstruction in degraded imaging conditions remains a key challenge in photogrammetry and neural rendering. In underwater environments, spatially varying visibility caused by scattering, attenuation, and sparse observations…

Computer Vision and Pattern Recognition · Computer Science 2026-04-23 Zhuodong Jiang , Haoran Wang , Guoxi Huang , Brett Seymour , Nantheera Anantrasirichai

Accurate 3D scene representation and panoptic understanding are essential for applications such as virtual reality, robotics, and autonomous driving. However, challenges persist with existing methods, including precise 2D-to-3D mapping,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-08 Shenghao Li

The growing use of wide angle image capture devices and the need for fast and accurate image analysis in computer visions have enforced the need for dedicated under-representation approaches. Most recent decomposition methods segment an…

Computer Vision and Pattern Recognition · Computer Science 2025-09-25 Rémi Giraud , Rodrigo Borba Pinheiro , Yannick Berthoumieu

Segment matching is an important intermediate task in computer vision that establishes correspondences between semantically or geometrically coherent regions across images. Unlike keypoint matching, which focuses on localized features,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-27 Rohit Jayanti , Swayam Agrawal , Vansh Garg , Siddharth Tourani , Muhammad Haris Khan , Sourav Garg , Madhava Krishna

The goal of the Semantic Scene Completion (SSC) task is to simultaneously predict a completed 3D voxel representation of volumetric occupancy and semantic labels of objects in the scene from a single-view observation. Since the…

Computer Vision and Pattern Recognition · Computer Science 2020-04-01 Xiaokang Chen , Kwan-Yee Lin , Chen Qian , Gang Zeng , Hongsheng Li

Existing panoramic depth estimation methods based on convolutional neural networks (CNNs) focus on removing panoramic distortions, failing to perceive panoramic structures efficiently due to the fixed receptive field in CNNs. This paper…

Computer Vision and Pattern Recognition · Computer Science 2022-07-13 Zhijie Shen , Chunyu Lin , Kang Liao , Lang Nie , Zishuo Zheng , Yao Zhao

Advancements in 3D rendering like Gaussian Splatting (GS) allow novel view synthesis and real-time rendering in virtual reality (VR). However, GS-created 3D environments are often difficult to edit. For scene enhancement or to incorporate…

Computer Vision and Pattern Recognition · Computer Science 2024-09-25 Hannah Schieber , Jacob Young , Tobias Langlotz , Stefanie Zollmann , Daniel Roth

In this work, we propose "tangent images," a spherical image representation that facilitates transferable and scalable $360^\circ$ computer vision. Inspired by techniques in cartography and computer graphics, we render a spherical image to…

Computer Vision and Pattern Recognition · Computer Science 2020-05-25 Marc Eder , Mykhailo Shvets , John Lim , Jan-Michael Frahm

In the task of 3D Aerial-view Scene Semantic Segmentation (3D-AVS-SS), traditional methods struggle to address semantic ambiguity caused by scale variations and structural occlusions in aerial images. This limits their segmentation accuracy…

Computer Vision and Pattern Recognition · Computer Science 2025-08-15 Xu Tang , Junan Jia , Yijing Wang , Jingjing Ma , Xiangrong Zhang