English
Related papers

Related papers: Upsample Anything: A Simple and Hard to Beat Basel…

200 papers

We develop a novel framework for sparse multiscale kernel approximation of large scattered data problems based on a samplet representation. Samplets form a multiresolution analysis of localized discrete signed measures and enable…

Numerical Analysis · Mathematics 2026-04-03 Sara Avesani , Gaia Fumagalli , Michael Multerer , Chiara Segala

Object detection in unmanned aerial vehicle (UAV) remote sensing images poses significant challenges due to unstable image quality, small object sizes, complex backgrounds, and environmental occlusions. Small objects, in particular, occupy…

Computer Vision and Pattern Recognition · Computer Science 2025-05-23 Xudong Wang , Yaxin Peng , Chaomin Shen

Image compression is one of the most fundamental techniques and commonly used applications in the image and video processing field. Earlier methods built a well-designed pipeline, and efforts were made to improve all modules of the pipeline…

Image and Video Processing · Electrical Eng. & Systems 2021-03-29 Yueyu Hu , Wenhan Yang , Zhan Ma , Jiaying Liu

This study explores the use of photometric techniques (shape-from-shading and uncalibrated photometric stereo) for upsampling the low-resolution depth map from an RGB-D sensor to the higher resolution of the companion RGB image. A…

Computer Vision and Pattern Recognition · Computer Science 2019-06-26 Bjoern Haefner , Songyou Peng , Alok Verma , Yvain Quéau , Daniel Cremers

Specular highlights distort appearance, obscure texture, and hinder geometric reasoning in both natural and surgical imagery. We present UnReflectAnything, an RGB-only framework that removes highlights from a single image by predicting a…

Computer Vision and Pattern Recognition · Computer Science 2025-12-12 Alberto Rota , Mert Kiray , Mert Asim Karaoglu , Patrick Ruhkamp , Elena De Momi , Nassir Navab , Benjamin Busam

Powered by large-scale pre-training, vision foundation models exhibit significant potential in open-world image understanding. However, unlike large language models that excel at directly tackling various language tasks, vision foundation…

Computer Vision and Pattern Recognition · Computer Science 2024-01-22 Yang Liu , Muzhi Zhu , Hengtao Li , Hao Chen , Xinlong Wang , Chunhua Shen

3D vision foundation models have shown strong generalization in reconstructing key 3D attributes from uncalibrated images through a single feed-forward pass. However, when deployed in online settings such as driving scenarios, predictions…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Fengyi Zhang , Tianjun Zhang , Kasra Khosoussi , Zheng Zhang , Zi Huang , Yadan Luo

We introduce Adaptive Guided Upsampling (AGU), an efficient method for upscaling low-light images capable of optimizing multiple image quality characteristics at the same time, such as reducing noise and increasing sharpness. It is based on…

Computer Vision and Pattern Recognition · Computer Science 2025-11-21 Angela Vivian Dcosta , Chunbo Song , Rafael Radkowski

Feature reassembly, i.e. feature downsampling and upsampling, is a key operation in a number of modern convolutional network architectures, e.g., residual networks and feature pyramids. Its design is critical for dense prediction tasks such…

Computer Vision and Pattern Recognition · Computer Science 2020-12-10 Jiaqi Wang , Kai Chen , Rui Xu , Ziwei Liu , Chen Change Loy , Dahua Lin

The emergence of large foundation models has propelled significant advances in various domains. The Segment Anything Model (SAM), a leading model for image segmentation, exemplifies these advances, outperforming traditional methods.…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Saurabh Yadav , Avi Gupta , Koteswar Rao Jerripothula

We address the problem of image reconstruction from incomplete measurements, encompassing both upsampling and inpainting, within a learning-based framework. Conventional supervised approaches require fully sampled ground truth data, while…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Benjamin Walder , Daniel Toader , Robert Nuster , Günther Paltauf , Peter Burgholzer , Gregor Langer , Lukas Krainer , Markus Haltmeier

The performance of image segmentation models has historically been constrained by the high cost of collecting large-scale annotated data. The Segment Anything Model (SAM) alleviates this original problem through a promptable,…

Computer Vision and Pattern Recognition · Computer Science 2026-02-04 Miguel Espinosa , Chenhongyi Yang , Linus Ericsson , Steven McDonagh , Elliot J. Crowley

Feature selection in reinforcement learning (RL), i.e. choosing basis functions such that useful approximations of the unkown value function can be obtained, is one of the main challenges in scaling RL to real-world applications. Here we…

Artificial Intelligence · Computer Science 2012-02-01 Tobias Jung , Peter Stone

Segment Anything Model (SAM), known for its remarkable zero-shot segmentation capabilities, has garnered significant attention in the community. Nevertheless, its performance is challenged when dealing with what we refer to as visually…

Computer Vision and Pattern Recognition · Computer Science 2026-01-05 Guangqian Guo , Pengfei Chen , Yong Guo , Huafeng Chen , Boqiang Zhang , Shan Gao

In image retrieval, deep local features learned in a data-driven manner have been demonstrated effective to improve retrieval performance. To realize efficient retrieval on large image database, some approaches quantize deep local features…

Image and Video Processing · Electrical Eng. & Systems 2021-12-14 Hui Wu , Min Wang , Wengang Zhou , Yang Hu , Houqiang Li

Optical Coherence Tomography (OCT) is a widely used non-invasive biomedical imaging modality that can rapidly provide volumetric images of samples. Here, we present a deep learning-based image reconstruction framework that can generate…

Image and Video Processing · Electrical Eng. & Systems 2021-07-30 Yijie Zhang , Tairan Liu , Manmohan Singh , Yilin Luo , Yair Rivenson , Kirill V. Larin , Aydogan Ozcan

Recent advances in generative AI have accelerated the production of ultra-high-resolution visual content, posing significant challenges for efficient compression and real-time decoding on end-user devices. Inspired by 3D Gaussian Splatting,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-13 Linfei Li , Lin Zhang , Zhong Wang , Ying Shen

The resurgence of convolutional neural networks (CNNs) in visual recognition tasks, exemplified by ConvNeXt, has demonstrated their capability to rival transformer-based architectures through advanced training methodologies and ViT-inspired…

Computer Vision and Pattern Recognition · Computer Science 2025-07-16 Quan Bi Pay , Vishnu Monn Baskaran , Junn Yong Loo , KokSheik Wong , Simon See

Existing deep architectures cannot operate on very large signals such as megapixel images due to computational and memory constraints. To tackle this limitation, we propose a fully differentiable end-to-end trainable model that samples and…

Computer Vision and Pattern Recognition · Computer Science 2019-07-18 Angelos Katharopoulos , François Fleuret

Generalizable rendering of an animatable human avatar from sparse inputs relies on data priors and inductive biases extracted from training on large data to avoid scene-specific optimization and to enable fast reconstruction. This raises…

Computer Vision and Pattern Recognition · Computer Science 2025-02-14 Jing Wen , Alexander G. Schwing , Shenlong Wang