English
Related papers

Related papers: UniCal: a Single-Branch Transformer-Based Model fo…

200 papers

One of the main challenges in simultaneous localization and mapping (SLAM) is real-time processing. High-computational loads linked to data acquisition and processing complicate this task. This article presents an efficient feature…

In deep learning era, pretrained models play an important role in medical image analysis, in which ImageNet pretraining has been widely adopted as the best way. However, it is undeniable that there exists an obvious domain gap between…

Computer Vision and Pattern Recognition · Computer Science 2020-07-23 Hong-Yu Zhou , Shuang Yu , Cheng Bian , Yifan Hu , Kai Ma , Yefeng Zheng

Deep learning models frequently encounter feature uncertainty in diverse learning scenarios, significantly impacting their performance and reliability. This challenge is particularly complex in multi-modal scenarios, where models must…

Machine Learning · Computer Science 2025-06-05 Jiahao Qin , Bei Peng , Feng Liu , Guangliang Cheng , Lu Zong

We present a novel method for extrinsically calibrating a camera and a 2D Laser Rangefinder (LRF) whose beams are invisible from the camera image. We show that point-to-plane constraints from a single observation of a V-shaped calibration…

Computer Vision and Pattern Recognition · Computer Science 2018-09-05 Wenbo Dong , Volkan Isler

Millimeter-wave (mmWave) Radar--Camera fusion improves perception under adverse illumination and weather, but its performance is sensitive to Radar--Camera extrinsic calibration: residual misalignment biases Radar-to-image projection and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Yuting Wan , Liguo Sun , Jiuwu Hao , Pin LV

The use of infrastructure sensor technology for traffic detection has already been proven several times. However, extrinsic sensor calibration is still a challenge for the operator. While previous approaches are unable to calibrate the…

Computer Vision and Pattern Recognition · Computer Science 2020-08-04 Laurent Kloeker , Christian Kotulla , Lutz Eckstein

World models have demonstrated significant promise for data synthesis in autonomous driving. However, existing methods predominantly concentrate on single-modality generation, typically focusing on either multi-camera video or LiDAR…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Guosheng Zhao , Yaozeng Wang , Xiaofeng Wang , Zheng Zhu , Tingdong Yu , Guan Huang , Yongchen Zai , Ji Jiao , Changliang Xue , Xiaole Wang , Zhen Yang , Futang Zhu , Xingang Wang

In view of contemporary panoramic camera-laser scanner system, the traditional calibration method is not suitable for panoramic cameras whose imaging model is extremely nonlinear. The method based on statistical optimization has the…

Computer Vision and Pattern Recognition · Computer Science 2017-09-14 Mingwei Cao , Ming Yang , Chunxiang Wang , Yeqiang Qian , Bing Wang

Multitask learning (MTL) has become prominent for its ability to predict multiple tasks jointly, achieving better per-task performance with fewer parameters than single-task learning. Recently, decoder-focused architectures have…

Computer Vision and Pattern Recognition · Computer Science 2024-11-07 Dimitrios Sinodinos , Narges Armanfard

Vision-Transformers (ViTs) and Convolutional neural networks (CNNs) are widely used Deep Neural Networks (DNNs) for classification task. These model architectures are dependent on the number of classes in the dataset it was trained on. Any…

Computer Vision and Pattern Recognition · Computer Science 2023-05-15 Shakti N. Wadekar , Eugenio Culurciello

3D object Detection with LiDAR-camera encounters overfitting in algorithm development which is derived from the violation of some fundamental rules. We refer to the data annotation in dataset construction for theory complementing and argue…

Computer Vision and Pattern Recognition · Computer Science 2023-11-14 Junjie Huang , Yun Ye , Zhujin Liang , Yi Shan , Dalong Du

Conventional Vision Transformer simplifies visual modeling by standardizing input resolutions, often disregarding the variability of natural visual data and compromising spatial-contextual fidelity. While preliminary explorations have…

Computer Vision and Pattern Recognition · Computer Science 2025-05-30 Limeng Qiao , Yiyang Gan , Bairui Wang , Jie Qin , Shuang Xu , Siqi Yang , Lin Ma

Sparse-view Cone-Beam Computed Tomography reconstruction from limited X-ray projections remains a challenging problem in medical imaging due to the inherent undersampling of fine-grained anatomical details, which correspond to…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Cuong Tran Van , Trong-Thang Pham , Ngoc-Son Nguyen , Duy Minh Ho Nguyen , Ngan Le

Model calibration and debiasing are fundamental yet operationally expensive challenges in large-scale recommendation systems. Existing approaches treat them as separate problems requiring distinct infrastructure: post-hoc calibration…

Information Retrieval · Computer Science 2026-04-28 Hailing Cheng , Yafang Yang , Hemeng Tao , Fengyu Zhang

Unpaired image-to-image translation (UNIT) aims to map images between two visual domains without paired training data. However, given a UNIT model trained on certain domains, it is difficult for current methods to incorporate new domains…

Computer Vision and Pattern Recognition · Computer Science 2023-06-27 Siyu Huang , Jie An , Donglai Wei , Zudi Lin , Jiebo Luo , Hanspeter Pfister

LiDAR-camera extrinsic calibration (LCEC) is crucial for data fusion in intelligent vehicles. Offline, target-based approaches have long been the preferred choice in this field. However, they often demonstrate poor adaptability to…

Robotics · Computer Science 2024-06-21 Zhiwei Huang , Yikang Zhang , Qijun Chen , Rui Fan

We introduce a discriminative multimodal descriptor based on a pair of sensor readings: a point cloud from a LiDAR and an image from an RGB camera. Our descriptor, named MinkLoc++, can be used for place recognition, re-localization and loop…

Computer Vision and Pattern Recognition · Computer Science 2021-04-15 Jacek Komorowski , Monika Wysoczanska , Tomasz Trzcinski

In-context Learning enables training-free adaptation via demonstrations but remains highly sensitive to example selection and formatting. In unified multimodal models spanning understanding and generation, this sensitivity is exacerbated by…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 Yicheng Xu , Jiangning Zhang , Zhucun Xue , Teng Hu , Ran Yi , Xiaobin Hu , Yong Liu , Dacheng Tao

Multi-agent systems, e.g., automobiles and UAVs (Unmanned Ariel Vehicles), rely on the precision of onboard sensors to accurately perceive their environment, which in turn depends on the precision of onboard sensors and reliable in-field…

Signal Processing · Electrical Eng. & Systems 2026-04-23 Bichi Zhang , Holger Caesar , Raj Thilak Rajan

LiDAR place recognition is a critical capability for autonomous navigation and cross-modal localization in large-scale outdoor environments. Existing approaches predominantly depend on pre-built 3D dense maps or aerial imagery, which impose…

Computer Vision and Pattern Recognition · Computer Science 2025-08-28 Shuhao Kang , Martin Y. Liao , Yan Xia , Olaf Wysocki , Boris Jutzi , Daniel Cremers