English
Related papers

Related papers: CLAMP: Majorized Plug-and-Play for Coherent 3D LID…

200 papers

Compressive image recovery is a challenging problem that requires fast and accurate algorithms. Recently, neural networks have been applied to this problem with promising results. By exploiting massively parallel GPU processing…

Machine Learning · Statistics 2017-11-08 Christopher A. Metzler , Ali Mousavi , Richard G. Baraniuk

In contrast to sparse keypoints, a handful of line segments can concisely encode the high-level scene layout, as they often delineate the main structural elements. In addition to offering strong geometric cues, they are also omnipresent in…

Computer Vision and Pattern Recognition · Computer Science 2023-03-31 Shaohui Liu , Yifan Yu , Rémi Pautrat , Marc Pollefeys , Viktor Larsson

Model-Based Image Reconstruction (MBIR) methods significantly enhance the quality of computed tomographic (CT) reconstructions relative to analytical techniques, but are limited by high computational cost. In this paper, we propose a…

Image and Video Processing · Electrical Eng. & Systems 2019-11-22 Venkatesh Sridhar , Xiao Wang , Gregery T. Buzzard , Charles A. Bouman

LiDAR-camera fusion is one of the core processes for the perception system of current automated driving systems. The typical sensor fusion process includes a list of coordinate transformation operations following system calibration.…

Robotics · Computer Science 2023-11-09 Dan Shen , Zhengming Zhang , Renran Tian , Yaobin Chen , Rini Sherony

We introduce CHAMP, a novel method for learning sequence-to-sequence, multi-hypothesis 3D human poses from 2D keypoints by leveraging a conditional distribution with a diffusion model. To predict a single output 3D pose sequence, we…

Computer Vision and Pattern Recognition · Computer Science 2025-02-25 Harry Zhang , Luca Carlone

We present a novel target-based lidar-camera extrinsic calibration methodology that can be used for non-overlapping field of view (FOV) sensors. Contrary to previous work, our methodology overcomes the non-overlapping FOV challenge using a…

Robotics · Computer Science 2025-03-04 Nicholas Charron , Huaiyuan Weng , Steven L. Waslander , Sriram Narasimhan

In this paper we deal with image classification tasks using the powerful CLIP vision-language model. Our goal is to advance the classification performance using the CLIP's image encoder, by proposing a novel Large Multimodal Model (LMM)…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Maria Tzelepi , Vasileios Mezaris

Contrastive Language-Image Pre-training (CLIP) demonstrates strong potential in medical image analysis but requires substantial data and computational resources. Due to these restrictions, existing CLIP applications in medical imaging focus…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Yuexi Du , John Onofrey , Nicha C. Dvornek

3D surface reconstruction is essential across applications of virtual reality, robotics, and mobile scanning. However, RGB-based reconstruction often fails in low-texture, low-light, and low-albedo scenes. Handheld LiDARs, now common on…

Image and Video Processing · Electrical Eng. & Systems 2024-12-02 Nikhil Behari , Aaron Young , Siddharth Somasundaram , Tzofi Klinghoffer , Akshat Dave , Ramesh Raskar

CLaMP 3 is a unified framework developed to address challenges of cross-modal and cross-lingual generalization in music information retrieval. Using contrastive learning, it aligns all major music modalities--including sheet music,…

The joint optimization of the sensor trajectory and 3D map is a crucial characteristic of Simultaneous Localization and Mapping (SLAM) systems. To achieve this, the gold standard is Bundle Adjustment (BA). Modern 3D LiDARs now retain higher…

Computer Vision and Pattern Recognition · Computer Science 2023-03-30 Luca Di Giammarino , Emanuele Giacomini , Leonardo Brizi , Omar Salem , Giorgio Grisetti

CLIP (Contrastive Language-Image Pre-training) uses contrastive learning from noise image-text pairs to excel at recognizing a wide array of candidates, yet its focus on broad associations hinders the precision in distinguishing subtle…

Computer Vision and Pattern Recognition · Computer Science 2026-05-18 Ziyu Liu , Zeyi Sun , Yuhang Zang , Wei Li , Pan Zhang , Xiaoyi Dong , Yuanjun Xiong , Dahua Lin , Jiaqi Wang

Motivated by the increasing application of low-resolution LiDAR recently, we target the problem of low-resolution LiDAR-camera calibration in this work. The main challenges are two-fold: sparsity and noise in point clouds. To address the…

Computer Vision and Pattern Recognition · Computer Science 2022-11-09 Zhikang Zhang , Zifan Yu , Suya You , Raghuveer Rao , Sanjeev Agarwal , Fengbo Ren

Light detection and ranging (LiDAR) is a ubiquitous tool to provide precise spatial awareness in various perception environments. A bionic LiDAR that can mimic human-like vision by adaptively gazing at selected regions of interest within a…

Limited-Angle Computed Tomography (LACT) is a challenging inverse problem where missing angular projections lead to incomplete sinograms and severe artifacts in the reconstructed images. While recent learning-based methods have demonstrated…

Image and Video Processing · Electrical Eng. & Systems 2025-07-09 Jiaqi Guo , Santiago López-Tapia

In autonomous driving scenarios, the collected LiDAR point clouds can be challenged by occlusion and long-range sparsity, limiting the perception of autonomous driving systems. Scene completion methods can infer the missing parts of…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Andrea Matteazzi , Dietmar Tutsch

LiDAR-based SLAM algorithms are extensively studied to providing robust and accurate positioning for autonomous driving vehicles (ADV) in the past decades. Satisfactory performance can be obtained using high-grade 3D LiDAR with 64 channels,…

Computer Vision and Pattern Recognition · Computer Science 2020-08-11 Jiang Yue , Weisong Wen , Jing Han , Li-Ta Hsu

Compared with 2D MRI, 3D MRI provides superior volumetric spatial resolution and signal-to-noise ratio. However, it is more challenging to reconstruct 3D MRI images. Current methods are mainly based on convolutional neural networks (CNN)…

Image and Video Processing · Electrical Eng. & Systems 2023-06-01 Eric Z. Chen , Chi Zhang , Xiao Chen , Yikang Liu , Terrence Chen , Shanhui Sun

Small lesions in magnetic resonance imaging (MRI) images are crucial for clinical diagnosis of many kinds of diseases. However, the MRI quality can be easily degraded by various noise, which can greatly affect the accuracy of diagnosis of…

Image and Video Processing · Electrical Eng. & Systems 2022-09-29 Haibo Yang , Shengjie Zhang , Xiaoyang Han , Botao Zhao , Yan Ren , Yaru Sheng , Xiao-Yong Zhang

Contrastive Language-Image Pre-training (CLIP) exhibits strong zero-shot classification ability on various image-level tasks, leading to the research to adapt CLIP for pixel-level open-vocabulary semantic segmentation without additional…

Computer Vision and Pattern Recognition · Computer Science 2024-11-22 Lin Sun , Jiale Cao , Jin Xie , Xiaoheng Jiang , Yanwei Pang