中文
相关论文

相关论文: FOVAL: Calibration-Free and Subject-Invariant Fixa…

200 篇论文

Perception in the real world requires robustness to diverse viewing conditions. Existing approaches often rely on specialized architectures or training with predefined data augmentations, limiting adaptability. Taking inspiration from…

计算机视觉与模式识别 · 计算机科学 2025-09-17 Utkarsh Singhal , Ryan Feng , Stella X. Yu , Atul Prakash

Real-world datasets follow an imbalanced distribution, which poses significant challenges in rare-category object detection. Recent studies tackle this problem by developing re-weighting and re-sampling methods, that utilise the class…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Konstantinos Panagiotis Alexandridis , Ismail Elezi , Jiankang Deng , Anh Nguyen , Shan Luo

We present FoundationSLAM, a learning-based monocular dense SLAM system that addresses the absence of geometric consistency in previous flow-based approaches for accurate and robust tracking and mapping. Our core idea is to bridge flow…

计算机视觉与模式识别 · 计算机科学 2026-01-05 Yuchen Wu , Jiahe Li , Fabio Tosi , Matteo Poggi , Jin Zheng , Xiao Bai

Volume data is found in many important scientific and engineering applications. Rendering this data for visualization at high quality and interactive rates for demanding applications such as virtual reality is still not easily achievable…

图形学 · 计算机科学 2022-09-22 David Bauer , Qi Wu , Kwan-Liu Ma

There are a multitude of emerging imaging technologies that could benefit robotics. However the need for bespoke models, calibration and low-level processing represents a key barrier to their adoption. In this work we present NOCaL, Neural…

机器人学 · 计算机科学 2022-10-19 Ryan Griffiths , Jack Naylor , Donald G. Dansereau

Large vision-language models (VLMs) typically process hundreds or thousands of visual tokens per image or video frame, incurring quadratic attention cost and substantial redundancy. Existing token reduction methods often ignore the textual…

计算机视觉与模式识别 · 计算机科学 2025-12-24 Kaitong Cai , Jusheng Zhang , Jing Yang , Yijia Fan , Pengtao Xie , Jian Wang , Keze Wang

While deep networks can learn complex functions such as classifiers, detectors, and trackers, many applications require models that continually adapt to changing input distributions, changing tasks, and changing environmental conditions.…

机器学习 · 计算机科学 2022-02-21 Jathushan Rajasegaran , Chelsea Finn , Sergey Levine

Most existing methods for depth estimation from a focal stack of images employ convolutional neural networks (CNNs) using 2D or 3D convolutions over a fixed set of images. However, their effectiveness is constrained by the local properties…

计算机视觉与模式识别 · 计算机科学 2024-12-05 Xueyang Kang , Fengze Han , Abdur R. Fayjie , Patrick Vandewalle , Kourosh Khoshelham , Dong Gong

Deep learning architectures are an extremely powerful tool for recognizing and classifying images. However, they require supervised learning and normally work on vectors the size of image pixels and produce the best results when trained on…

机器学习 · 计算机科学 2020-10-20 Ryan Burt , Nina N. Thigpen , Andreas Keil , Jose C. Principe

AI tasks in the car interior like identifying and localizing externally introduced objects is crucial for response quality of personal assistants. However, computational resources of on-board systems remain highly constrained, restricting…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Sebastian Schmidt , Bálint Mészáros , Ahmet Firintepe , Stephan Günnemann

The human visual system processes images with varied degrees of resolution, with the fovea, a small portion of the retina, capturing the highest acuity region, which gradually declines toward the field of view's periphery. However, the…

计算机视觉与模式识别 · 计算机科学 2023-04-13 Beatriz Paula , Plinio Moreno

Wide field-of-view (FoV) cameras efficiently capture large portions of the scene, which makes them attractive in multiple domains, such as automotive and robotics. For such applications, estimating depth from multiple images is a critical…

计算机视觉与模式识别 · 计算机科学 2024-01-26 Daniel Lichy , Hang Su , Abhishek Badki , Jan Kautz , Orazio Gallo

Much recent work has been devoted to the problem of ensuring that a neural network's confidence scores match the true probability of being correct, i.e. the calibration problem. Of note, it was found that training with focal loss leads to…

机器学习 · 计算机科学 2023-06-21 Arindam Ghosh , Thomas Schaaf , Matthew R. Gormley

Visual autoregressive models achieve remarkable generation quality through next-scale predictions across multi-scale token pyramids. However, the conventional method uses uniform scale downsampling to build these pyramids, leading to…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Xiaofan Li , Chenming Wu , Yanpeng Sun , Jiaming Zhou , Delin Qu , Yansong Qu , Weihao Bo , Haibao Yu , Dingkang Liang

Self-driving vehicles (SDVs) require accurate calibration of LiDARs and cameras to fuse sensor data accurately for autonomy. Traditional calibration methods typically leverage fiducials captured in a controlled and structured scene and…

计算机视觉与模式识别 · 计算机科学 2024-09-30 Ze Yang , George Chen , Haowei Zhang , Kevin Ta , Ioan Andrei Bârsan , Daniel Murphy , Sivabalan Manivasagam , Raquel Urtasun

LiDAR-camera extrinsic calibration (LCEC) is crucial for data fusion in intelligent vehicles. Offline, target-based approaches have long been the preferred choice in this field. However, they often demonstrate poor adaptability to…

机器人学 · 计算机科学 2024-06-21 Zhiwei Huang , Yikang Zhang , Qijun Chen , Rui Fan

While modern deep neural networks are performant perception modules, performance (accuracy) alone is insufficient, particularly for safety-critical robotic applications such as self-driving vehicles. Robot autonomy stacks also require these…

计算机视觉与模式识别 · 计算机科学 2021-09-29 Dhaivat Bhatt , Kaustubh Mani , Dishank Bansal , Krishna Murthy , Hanju Lee , Liam Paull

Deep learning approaches have been widely adopted for precipitation nowcasting in recent years. Previous studies mainly focus on proposing new model architectures to improve pixel-wise metrics. However, they frequently result in blurry…

计算机视觉与模式识别 · 计算机科学 2024-10-31 Chiu-Wai Yan , Shi Quan Foo , Van Hoan Trinh , Dit-Yan Yeung , Ka-Hing Wong , Wai-Kin Wong

Low-Light Image Enhancement (LLIE) is a key task in computational photography and imaging. The problem of enhancing images captured during night or in dark environments has been well-studied in the computer vision literature. However,…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Juan C. Benito , Daniel Feijoo , Alvaro Garcia , Marcos V. Conde

Vanilla models for object detection and instance segmentation suffer from the heavy bias toward detecting frequent objects in the long-tailed setting. Existing methods address this issue mostly during training, e.g., by re-sampling or…

计算机视觉与模式识别 · 计算机科学 2021-12-01 Tai-Yu Pan , Cheng Zhang , Yandong Li , Hexiang Hu , Dong Xuan , Soravit Changpinyo , Boqing Gong , Wei-Lun Chao
‹ 上一页 1 2 3 10 下一页 ›