中文
相关论文

相关论文: SOFI: Multi-Scale Deformable Transformer for Camer…

200 篇论文

Single image camera calibration is the task of estimating the camera parameters from a single input image, such as the vanishing points, focal length, and horizon line. In this work, we propose Camera calibration TRansformer with…

计算机视觉与模式识别 · 计算机科学 2021-09-07 Jinwoo Lee , Hyunsung Go , Hyunjoon Lee , Sunghyun Cho , Minhyuk Sung , Junho Kim

The fusion of LiDARs and cameras has been increasingly adopted in autonomous driving for perception tasks. The performance of such fusion-based algorithms largely depends on the accuracy of sensor calibration, which is challenging due to…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Yuxuan Xiao , Yao Li , Chengzhen Meng , Xingchen Li , Jianmin Ji , Yanyong Zhang

Reliable multi-modal calibration requires identifying which observations truly constrain the extrinsic parameters and which ones mainly add noise or ambiguity. In this paper, we propose a support-map-driven approach to multi-modal…

计算机视觉与模式识别 · 计算机科学 2026-05-25 Rajitha de Silva , Grzegorz Cielniak

This paper presents a novel semantic-based online extrinsic calibration approach, SOIC (so, I see), for Light Detection and Ranging (LiDAR) and camera sensors. Previous online calibration methods usually need prior knowledge of rough…

计算机视觉与模式识别 · 计算机科学 2020-03-26 Weimin Wang , Shohei Nobuhara , Ryosuke Nakamura , Ken Sakurada

In this paper, we uncover the hidden potential of Diffusion Transformers (DiTs) to significantly enhance generative tasks. Through an in-depth analysis of the denoising process, we demonstrate that introducing a single learned scaling…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Danil Tokhchukov , Aysel Mirzoeva , Andrey Kuznetsov , Konstantin Sobolev

Creating accurate and efficient 3D models poses significant challenges, particularly in addressing large viewpoint variations, computational complexity, and alignment discrepancies. Efficient camera path generation can help resolve these…

计算机视觉与模式识别 · 计算机科学 2025-05-26 Usha Kumari , Shuvendu Rana

For autonomous vehicles, an accurate calibration for LiDAR and camera is a prerequisite for multi-sensor perception systems. However, existing calibration techniques require either a complicated setting with various calibration targets, or…

计算机视觉与模式识别 · 计算机科学 2021-03-09 Tao Ma , Zhizheng Liu , Guohang Yan , Yikang Li

Camera calibration is an essential prerequisite for event-based vision applications. Current event camera calibration methods typically involve using flashing patterns, reconstructing intensity images, and utilizing the features extracted…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Zibin Liu , Banglei Guan , Yang Shang , Zhenbao Yu , Yifei Bian , Qifeng Yu

Accurate multi-sensor calibration is essential for deploying robust perception systems in applications such as autonomous driving and intelligent transportation. Existing LiDAR-camera calibration methods often rely on manually placed…

计算机视觉与模式识别 · 计算机科学 2025-09-03 Lei Cheng , Lihao Guo , Tianya Zhang , Tam Bang , Austin Harris , Mustafa Hajij , Mina Sartipi , Siyang Cao

Recently, while significant progress has been made in remote sensing image change captioning, existing methods fail to filter out areas unrelated to actual changes, making models susceptible to irrelevant features. In this article, we…

计算机视觉与模式识别 · 计算机科学 2024-09-20 Cong Yang , Zuchao Li , Hongzan Jiao , Zhi Gao , Lefei Zhang

Recently, deep-learning-based approaches have been widely studied for deformable image registration task. However, most efforts directly map the composite image representation to spatial transformation through the convolutional neural…

图像与视频处理 · 电气工程与系统科学 2022-07-08 Jiashun Chen , Donghuan Lu , Yu Zhang , Dong Wei , Munan Ning , Xinyu Shi , Zhe Xu , Yefeng Zheng

With information from multiple input modalities, sensor fusion-based algorithms usually out-perform their single-modality counterparts in robotics. Camera and LIDAR, with complementary semantic and depth information, are the typical choices…

计算机视觉与模式识别 · 计算机科学 2022-07-11 Akio Kodaira , Yiyang Zhou , Pengwei Zang , Wei Zhan , Masayoshi Tomizuka

Due to the lack of a definitive ground truth for the image fusion problem, the loss functions are structured based on evaluation metrics, such as the structural similarity index measure (SSIM). However, in doing so, a bias is introduced…

计算机视觉与模式识别 · 计算机科学 2024-04-25 Aytekin Erdogan , Erdem Akagündüz

Camera calibration involves estimating camera parameters to infer geometric features from captured sequences, which is crucial for computer vision and robotics. However, conventional calibration is laborious and requires dedicated…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Kang Liao , Lang Nie , Shujuan Huang , Chunyu Lin , Jing Zhang , Yao Zhao , Moncef Gabbouj , Dacheng Tao

Accurate camera calibration is essential for transforming 2D images from camera sensors into 3D world coordinates, enabling precise scene geometry interpretation and supporting sports analytics tasks such as player tracking, offside…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Nikolay S. Falaleev , Ruilong Chen

Most approaches to camera calibration rely on calibration targets of well-known geometry. During data acquisition, calibration target and camera system are typically moved w.r.t. each other, to allow image coverage and perspective…

计算机视觉与模式识别 · 计算机科学 2021-10-15 Annika Hagemann , Moritz Knorr , Christoph Stiller

Change captioning tasks aim to detect changes in image pairs observed before and after a scene change and generate a natural language description of the changes. Existing change captioning studies have mainly focused on a single…

计算机视觉与模式识别 · 计算机科学 2021-09-16 Yue Qiu , Shintaro Yamamoto , Kodai Nakashima , Ryota Suzuki , Kenji Iwata , Hirokatsu Kataoka , Yutaka Satoh

We address the problem of epipolar geometry using the motion of silhouettes. Such methods match epipolar lines or frontier points across views, which are then used as the set of putative correspondences. We introduce an approach that…

计算机视觉与模式识别 · 计算机科学 2017-04-17 Gil Ben-Artzi

Mainstream image caption models are usually two-stage captioners, i.e., calculating object features by pre-trained detector, and feeding them into a language model to generate text descriptions. However, such an operation will cause a…

计算机视觉与模式识别 · 计算机科学 2022-11-07 Bo Wang , Zhao Zhang , Mingbo Zhao , Xiaojie Jin , Mingliang Xu , Meng Wang

We introduce new linear mathematical formulations to calculate the focal length of a camera in an active platform. Through mathematical derivations, we show that the focal lengths in each direction can be estimated using only one point…

计算机视觉与模式识别 · 计算机科学 2018-06-12 Mehdi Faraji , Anup Basu
‹ 上一页 1 2 3 10 下一页 ›