English
Related papers

Related papers: SGFormer: Spherical Geometry Transformer for 360 D…

200 papers

We introduce SurgFormer, a multiresolution gated transformer for data driven soft tissue simulation on volumetric meshes. High fidelity biomechanical solvers are often too costly for interactive use, so we train SurgFormer on solver…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Ashkan Shahbazi , Elaheh Akbari , Kyvia Pereira , Jon S. Heiselman , Annie C. Benson , Garrison L. H. Johnston , Jie Ying Wu , Nabil Simaan , Michael I. Miga , Soheil Kolouri

Aerial Image Segmentation is a top-down perspective semantic segmentation and has several challenging characteristics such as strong imbalance in the foreground-background distribution, complex background, intra-class heterogeneity,…

Computer Vision and Pattern Recognition · Computer Science 2023-10-03 Kashu Yamazaki , Taisei Hanyu , Minh Tran , Adrian de Luis , Roy McCann , Haitao Liao , Chase Rainwater , Meredith Adkins , Jackson Cothren , Ngan Le

Neural radiance fields (NeRFs) have recently emerged as a promising approach for 3D reconstruction and novel view synthesis. However, NeRF-based methods encode shape, reflectance, and illumination implicitly and this makes it challenging…

Computer Vision and Pattern Recognition · Computer Science 2023-04-10 Ruofan Liang , Jiahao Zhang , Haoda Li , Chen Yang , Yushi Guan , Nandita Vijaykumar

Diffusion models excel at 2D outpainting, but extending them to $360^\circ$ panoramic completion from unposed perspective images is challenging due to the geometric and topological mismatch between perspective projections and spherical…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Yuqin Lu , Haofeng Liu , Yang Zhou , Jun Liang , Shengfeng He , Jing Li

Recent query-based detectors have achieved remarkable progress, yet their performance remains constrained when handling objects with arbitrary orientations, especially for tiny objects capturing limited texture information. This limitation…

Computer Vision and Pattern Recognition · Computer Science 2026-02-17 Junpeng Zhang , Zewei Yang , Jie Feng , Yuhui Zheng , Ronghua Shang , Mengxuan Zhang

Feature pyramids have been widely adopted in convolutional neural networks and transformers for tasks in medical image segmentation. However, existing models generally focus on the Encoder-side Transformer for feature extraction. We further…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Hongyi Cai , Mohammad Mahdinur Rahman , Wenzhen Dong , Jingyu Wu

Forward and backward scattering provide complementary volumetric and interfacial information, yet conventional three-dimensional (3D) imaging typically accesses only one. In this Letter, we present a substrate-enhanced diffraction…

Optics · Physics 2026-03-05 Tongyu Li , Yi Shen , Dashan Dong , Danchen Jia , Jianpeng Ao , Ji-Xin Cheng , Lei Tian

The availability of affordable and portable depth sensors has made scanning objects and people simpler than ever. However, dealing with occlusions and missing parts is still a significant challenge. The problem of reconstructing a (possibly…

Computer Vision and Pattern Recognition · Computer Science 2018-04-05 Or Litany , Alex Bronstein , Michael Bronstein , Ameesh Makadia

Transformers excel when dealing with sequential data. Generalizing transformer models to geometric domains, such as manifolds, we encounter the problem of not having a well-defined global order. We propose a solution with attention heads…

Machine Learning · Computer Science 2025-07-14 M. Maurin , M. Á. Evangelista-Alvarado , P. Suárez-Serrato

Panoramic videos contain richer spatial information and have attracted tremendous amounts of attention due to their exceptional experience in some fields such as autonomous driving and virtual reality. However, existing datasets for video…

Computer Vision and Pattern Recognition · Computer Science 2024-07-30 Shilin Yan , Xiaohao Xu , Renrui Zhang , Lingyi Hong , Wenchao Chen , Wenqiang Zhang , Wei Zhang

Applying convolutional neural networks to spherical images requires particular considerations. We look to the millennia of work on cartographic map projections to provide the tools to define an optimal representation of spherical images for…

Computer Vision and Pattern Recognition · Computer Science 2020-02-24 Marc Eder , Jan-Michael Frahm

This research presents a novel depth estimation algorithm based on a Transformer-encoder architecture, tailored for the NYU and KITTI Depth Dataset. This research adopts a transformer model, initially renowned for its success in natural…

Computer Vision and Pattern Recognition · Computer Science 2024-06-25 Linhan Xia , Junbang Liu , Tong Wu

Blind face restoration is a challenging task due to the unknown and complex degradation. Although face prior-based methods and reference-based methods have recently demonstrated high-quality results, the restored images tend to contain…

Computer Vision and Pattern Recognition · Computer Science 2024-03-01 Guojing Ge , Qi Song , Guibo Zhu , Yuting Zhang , Jinglu Chen , Miao Xin , Ming Tang , Jinqiao Wang

Underwater images often exhibit poor quality, distorted color balance and low contrast due to the complex and intricate interplay of light, water, and objects. Despite the significant contributions of previous underwater enhancement…

Computer Vision and Pattern Recognition · Computer Science 2024-04-25 Weiwen Chen , Yingtie Lei , Shenghong Luo , Ziyang Zhou , Mingxian Li , Chi-Man Pun

Adverse weather conditions cause diverse and complex degradation patterns, driving the development of All-in-One (AiO) models. However, recent AiO solutions still struggle to capture diverse degradations, since global filtering methods like…

Computer Vision and Pattern Recognition · Computer Science 2025-08-01 Yuhwan Jeong , Yunseo Yang , Youngho Yoon , Kuk-Jin Yoon

Blind face restoration aims at recovering high-quality face images from those with unknown degradations. Current algorithms mainly introduce priors to complement high-quality details and achieve impressive progress. However, most of these…

Computer Vision and Pattern Recognition · Computer Science 2023-08-15 Zhouxia Wang , Jiawei Zhang , Tianshui Chen , Wenping Wang , Ping Luo

This paper presents a novel approach, Spectral-Interpretable and -Enhanced Transformer (SIEFormer), which leverages spectral analysis to reinterpret the attention mechanism within Vision Transformer (ViT) and enhance feature adaptability,…

Computer Vision and Pattern Recognition · Computer Science 2026-02-16 Chunming Li , Shidong Wang , Tong Xin , Haofeng Zhang

Pansharpening aims to generate high-resolution multi-spectral images by fusing the spatial detail of panchromatic images with the spectral richness of low-resolution MS data. However, most existing methods are evaluated under limited,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Ke Cao , Xuanhua He , Xueheng Li , Lingting Zhu , Yingying Wang , Ao Ma , Zhanjie Zhang , Man Zhou , Chengjun Xie , Jie Zhang

The generation of immersive and navigable 3D environments is increasingly prevalent with the growing adoption of virtual reality and 3D content. However, recent methods face a fundamental limitation: they cannot produce 3D worlds that…

Computer Vision and Pattern Recognition · Computer Science 2026-05-20 Antoine Schnepf , Karim Kassab , Flavian Vasile , Andrew Comport

Blind face restoration is a highly ill-posed problem that often requires auxiliary guidance to 1) improve the mapping from degraded inputs to desired outputs, or 2) complement high-quality details lost in the inputs. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2022-11-02 Shangchen Zhou , Kelvin C. K. Chan , Chongyi Li , Chen Change Loy
‹ Prev 1 4 5 6 7 8 10 Next ›