中文
相关论文

相关论文: SOFI: Multi-Scale Deformable Transformer for Camer…

200 篇论文

Camera calibration is integral to robotics and computer vision algorithms that seek to infer geometric properties of the scene from visual input streams. In practice, calibration is a laborious procedure requiring specialized data…

计算机视觉与模式识别 · 计算机科学 2022-03-03 Jiading Fang , Igor Vasiljevic , Vitor Guizilini , Rares Ambrus , Greg Shakhnarovich , Adrien Gaidon , Matthew R. Walter

Time-of-flight cameras provide depth information, which is complementary to the photometric appearance of the scene in ordinary images. It is desirable to merge the depth and colour information, in order to obtain a coherent scene…

计算机视觉与模式识别 · 计算机科学 2020-12-18 Miles Hansard , Georgios Evangelidis , Quentin Pelorson , Radu Horaud

Reconstructing the 3D shape of a deformable environment from the information captured by a moving depth camera is highly relevant to surgery. The underlying challenge is the fact that simultaneously estimating camera motion and tissue…

计算机视觉与模式识别 · 计算机科学 2024-08-09 Guido Caccianiga , Julian Nubert , Cesar Cadena , Marco Hutter , Katherine J. Kuchenbecker

We propose Deep Feature Interpolation (DFI), a new data-driven baseline for automatic high-resolution image transformation. As the name suggests, it relies only on simple linear interpolation of deep convolutional features from pre-trained…

计算机视觉与模式识别 · 计算机科学 2017-06-20 Paul Upchurch , Jacob Gardner , Geoff Pleiss , Robert Pless , Noah Snavely , Kavita Bala , Kilian Weinberger

Calibrating large-scale camera arrays, such as those in dome-based setups, is time-intensive and typically requires dedicated captures of known patterns. While extrinsics in such arrays are fixed due to the physical setup, intrinsics often…

计算机视觉与模式识别 · 计算机科学 2025-08-04 Jinjiang You , Hewei Wang , Yijie Li , Mingxiao Huo , Long Van Tran Ha , Mingyuan Ma , Jinfeng Xu , Jiayi Zhang , Puzhen Wu , Shubham Garg , Wei Pu

Accurate calibration of internal parameters is a crucial yet challenging prerequisite for 3D reconstruction using light field cameras. In this paper, we propose a linear fractional transformation(LFT) parameter $\alpha$ to decoupled the…

计算机视觉与模式识别 · 计算机科学 2025-11-07 Zhong Chen , Changfeng Chen

With the recent advances in autonomous driving and the decreasing cost of LiDARs, the use of multimodal sensor systems is on the rise. However, in order to make use of the information provided by a variety of complimentary sensors, it is…

计算机视觉与模式识别 · 计算机科学 2023-12-20 Quentin Herau , Nathan Piasco , Moussab Bennehar , Luis Roldão , Dzmitry Tsishkou , Cyrille Migniot , Pascal Vasseur , Cédric Demonceaux

Video frame interpolation (VFI) aims to improve the temporal resolution of a video sequence. Most of the existing deep learning based VFI methods adopt off-the-shelf optical flow algorithms to estimate the bidirectional flows and…

计算机视觉与模式识别 · 计算机科学 2022-03-21 Tao Yang , Peiran Ren , Xuansong Xie , Xiansheng Hua , Lei Zhang

Typical Structure-from-Motion (SfM) pipelines rely on finding correspondences across images, recovering the projective structure of the observed scene and upgrading it to a metric frame using camera self-calibration constraints. Solving…

计算机视觉与模式识别 · 计算机科学 2020-07-07 Rui Gong , Danda Pani Paudel , Ajad Chhatkuli , Luc Van Gool

Cameras and LiDAR are essential sensors for autonomous vehicles. The fusion of camera and LiDAR data addresses the limitations of individual sensors but relies on precise extrinsic calibration. Recently, numerous end-to-end calibration…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Ni Ou , Zhuo Chen , Xinru Zhang , Junzheng Wang

Fourier ptychography (FP), as a computational imaging method, is a powerful tool to improve imaging resolution. Camera-scanning Fourier ptychography extends the application of FP from micro to macro creatively. Due to the non-ideal scanning…

图像与视频处理 · 电气工程与系统科学 2022-06-08 Baiqi Cui , Shaohui Zhang , Yechao Wang , Yao Hu , Qun Hao

This paper presents a new deformable convolution-based video frame interpolation (VFI) method, using a coarse to fine 3D CNN to enhance the multi-flow prediction. This model first extracts spatio-temporal features at multiple scales using a…

计算机视觉与模式识别 · 计算机科学 2024-06-11 Duolikun Danier , Fan Zhang , David Bull

Learned Image Compression (LIC) has achieved dramatic progress regarding objective and subjective metrics. MSE-based models aim to improve objective metrics while generative models are leveraged to improve visual quality measured by…

图像与视频处理 · 电气工程与系统科学 2024-05-24 Jixiang Luo , Yan Wang , Hongwei Qin

Self-supervised methods have showed promising results on depth estimation task. However, previous methods estimate the target depth map and camera ego-motion simultaneously, underusing multi-frame correlation information and ignoring the…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Songchun Zhang , Chunhui Zhao

Referring image segmentation aims to segment an object referred to by natural language expression from an image. The primary challenge lies in the efficient propagation of fine-grained semantic information from textual features to visual…

计算机视觉与模式识别 · 计算机科学 2024-04-15 Yichen Yan , Xingjian He , Sihan Chen , Jing Liu

Camera-based perception systems play a central role in modern autonomous vehicles. These camera based perception algorithms require an accurate calibration to map the real world distances to image pixels. In practice, calibration is a…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Ciarán Hogan , Ganesh Sistu , Ciarán Eising

We present a new mathematical formulation to estimate the intrinsic parameters of a camera in active or robotic platforms. We show that the focal lengths can be estimated using only one point correspondence that relates images taken before…

计算机视觉与模式识别 · 计算机科学 2018-07-27 Mehdi Faraji , Anup Basu

Transformers have demonstrated remarkable performance in natural language processing and computer vision. However, existing vision Transformers struggle to learn from limited medical data and are unable to generalize on diverse medical…

图像与视频处理 · 电气工程与系统科学 2023-04-06 Yunhe Gao , Mu Zhou , Di Liu , Zhennan Yan , Shaoting Zhang , Dimitris N. Metaxas

Image identification is one of the most challenging tasks in different areas of computer vision. Scale-invariant feature transform is an algorithm to detect and describe local features in images to further use them as an image matching…

计算机视觉与模式识别 · 计算机科学 2018-03-15 Ebrahim Karami , Mohamed Shehata , Andrew Smith

Flexible Intelligent Metasurfaces (FIMs) enable wireless systems to adapt their three-dimensional geometry through morphing, thereby providing new spatial degrees of freedom. However, continuous deformation complicates the accurate…

信号处理 · 电气工程与系统科学 2026-05-29 Vinícius L. Romano , André L. F. de Almeida , Daniel C. Araújo