中文
相关论文

相关论文: Improving Feature Stability during Upsampling -- S…

200 篇论文

This paper addresses the issue of detecting point objects in a clutter background and estimating their position by image processing. We are interested in the specific context where the object signature significantly varies with its random…

应用统计 · 统计学 2009-11-09 Vincent Samson , Frédéric Champagnat , Jean-François Giovannelli

Image retargeting is the task of making images capable of being displayed on screens with different sizes. This work should be done so that high-level visual information and low-level features such as texture remain as intact as possible to…

计算机视觉与模式识别 · 计算机科学 2019-10-18 Mahdi Ahmadi , Nader Karimi , Shadrokh Samavi

Stochastic texture filtering (STF) has re-emerged as a technique that can bring down the cost of texture filtering of advanced texture compression methods, e.g., neural texture compression. However, during texture magnification, the swapped…

图形学 · 计算机科学 2025-04-09 Bartlomiej Wronski , Matt Pharr , Tomas Akenine-Möller

Neural fields have rapidly been adopted for representing 3D signals, but their application to more classical 2D image-processing has been relatively limited. In this paper, we consider one of the most important operations in image…

Using a stride in a convolutional layer inherently introduces aliasing, which has implications for numerical stability and statistical generalization. While techniques such as the parametrizations via paraunitary systems have been used to…

机器学习 · 计算机科学 2025-07-09 Daniel Haider , Vincent Lostanlen , Martin Ehler , Nicki Holighaus , Peter Balazs

Many computer vision systems require low-cost segmentation algorithms based on deep learning, either because of the enormous size of input images or limited computational budget. Common solutions uniformly downsample the input images to…

计算机视觉与模式识别 · 计算机科学 2022-08-19 Chen Jin , Ryutaro Tanno , Thomy Mertzanidou , Eleftheria Panagiotaki , Daniel C. Alexander

A typical 2D-to-3D pipeline takes multi-view images as input, where a Vision Foundation Model (VFM) extracts features that are spatially upsampled to dense representations for 3D reconstruction. If dense features across views preserve…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Ling Xiao , Yuliang Xiu , Yue Chen , Guoming Wang , Toshihiko Yamasaki

We propose a novel method of efficient upsampling of a single natural image. Current methods for image upsampling tend to produce high-resolution images with either blurry salient edges, or loss of fine textural detail, or spurious noise…

计算机视觉与模式识别 · 计算机科学 2015-03-03 Chinmay Hegde , Oncel Tuzel , Fatih Porikli

Recent semantic segmentation methods exploit encoder-decoder architectures to produce the desired pixel-wise segmentation prediction. The last layer of the decoders is typically a bilinear upsampling procedure to recover the final…

计算机视觉与模式识别 · 计算机科学 2019-04-08 Zhi Tian , Tong He , Chunhua Shen , Youliang Yan

Increasing spatial image resolution is an often required, yet challenging task in image acquisition. Recently, it has been shown that it is possible to obtain a high resolution image by covering a low resolution sensor with a non-regular…

图像与视频处理 · 电气工程与系统科学 2022-04-11 Markus Jonscher , Jürgen Seiler , Thomas Richter , André Kaup

Transformer-based audio self-supervised learning (SSL) models commonly use spectrograms, vision-style Transformers, and masked modeling objectives. However, convolutional patchification with temporal downsampling lowers the effective…

声音 · 计算机科学 2026-05-15 Kohei Yamamoto , Kosuke Okusa

Encoder-decoder networks have found widespread use in various dense prediction tasks. However, the strong reduction of spatial resolution in the encoder leads to a loss of location information as well as boundary artifacts. To address this,…

计算机视觉与模式识别 · 计算机科学 2020-04-01 Anne S. Wannenwetsch , Stefan Roth

Spatial aliasing affects spaced microphone arrays, causing directional ambiguity above certain frequencies, degrading spatial and spectral accuracy of beamformers. Given the limitations of conventional signal processing and the scarcity of…

音频与语音处理 · 电气工程与系统科学 2025-10-21 Mateusz Guzik , Giulio Cengarle , Daniel Arteaga

Recent advances in 3D Gaussian splatting have significantly improved real-time novel view synthesis, yet insufficient geometric constraints during scene optimization often result in blurred reconstructions of fine-grained details,…

计算机视觉与模式识别 · 计算机科学 2025-08-15 Zheng Zhou , Jia-Chen Zhang , Yu-Jie Xiong , Chun-Ming Xia

Anomaly Detection is a relevant problem in numerous real-world applications, especially when dealing with images. However, little attention has been paid to the issue of changes over time in the input data distribution, which may cause a…

计算机视觉与模式识别 · 计算机科学 2024-03-26 Nikola Bugarin , Jovana Bugaric , Manuel Barusco , Davide Dalle Pezze , Gian Antonio Susto

The remarkable success in face forgery techniques has received considerable attention in computer vision due to security concerns. We observe that up-sampling is a necessary step of most face forgery techniques, and cumulative up-sampling…

计算机视觉与模式识别 · 计算机科学 2021-03-11 Honggu Liu , Xiaodan Li , Wenbo Zhou , Yuefeng Chen , Yuan He , Hui Xue , Weiming Zhang , Nenghai Yu

Semantic segmentation, which refers to pixel-wise classification of an image, is a fundamental topic in computer vision owing to its growing importance in robot vision and autonomous driving industries. It provides rich information about…

计算机视觉与模式识别 · 计算机科学 2021-03-23 Khwaja Monib Sediqi , Hyo Jong Lee

A practical constraint that comes in the way of spectrum estimation of a continuous time stationary stochastic process is the minimum separation between successively observed samples of the process. When the underlying process is not…

统计理论 · 数学 2011-06-23 Radhendushka Srivastava , Debasis Sengupta

High-resolution satellite imagery has proven useful for a broad range of tasks, including measurement of global human population, local economic livelihoods, and biodiversity, among many others. Unfortunately, high-resolution imagery is…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Yutong He , Dingjie Wang , Nicholas Lai , William Zhang , Chenlin Meng , Marshall Burke , David B. Lobell , Stefano Ermon

Saliency maps are widely used in the computer vision community for interpreting neural network classifiers. However, due to the randomness of training samples and optimization algorithms, the resulting saliency maps suffer from a…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Shizhan Gong , Jingwei Zhang , Qi Dou , Farzan Farnia