English
Related papers

Related papers: Fast and Interpretable 2D Homography Decomposition…

200 papers

As a specific semantic segmentation task, aerial imagery segmentation has been widely employed in high spatial resolution (HSR) remote sensing images understanding. Besides common issues (e.g. large scale variation) faced by general…

Computer Vision and Pattern Recognition · Computer Science 2022-02-22 Lin Huang , Qiyuan Dong , Lijun Wu , Jia Zhang , Jiang Bian , Tie-Yan Liu

Recent works based on convolutional encoder-decoder architecture and 3DMM parameterization have shown great potential for canonical view reconstruction from a single input image. Conventional CNN architectures benefit from exploiting the…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Zhiqian Lin , Jiangke Lin , Lincheng Li , Yi Yuan , Zhengxia Zou

Semantic communications (SemComs) have emerged as a promising paradigm for joint data and task-oriented transmissions, combining the demands for both the bit-accurate delivery and end-to-end (E2E) distortion minimization. However, current…

Information Theory · Computer Science 2025-08-12 Dongxu Li , Kai Yuan , Jianhao Huang , Chuan Huang , Xiaoqi Qin , Shuguang Cui , Ping Zhang

To address the limitations of Transformer decoders in capturing edge details, recognizing local textures and modeling spatial continuity, this paper proposes a novel decoder framework specifically designed for medical image segmentation,…

Computer Vision and Pattern Recognition · Computer Science 2025-12-08 Fan Zhang , Zhiwei Gu , Hua Wang

The Hough transform (HT) is a fundamental tool across various domains, from classical image analysis to neural networks and tomography. Two key aspects of the algorithms for computing the HT are their computational complexity and accuracy -…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Danil Kazimirov , Dmitry Nikolaev

This paper proposes a novel Affine Subspace Representation (ASR) descriptor to deal with affine distortions induced by viewpoint changes. Unlike the traditional local descriptors such as SIFT, ASR inherently encodes local information of…

Computer Vision and Pattern Recognition · Computer Science 2014-07-21 Zhenhua Wang , Bin Fan , Fuchao Wu

Recent works have shown that depth information can be obtained from Dual-Pixel (DP) sensors. A DP arrangement provides two views in a single shot, thus resembling a stereo image pair with a tiny baseline. However, the different point spread…

Computer Vision and Pattern Recognition · Computer Science 2023-06-14 Sagi Monin , Sagi Katz , Georgios Evangelidis

In this paper, we propose a coarse-to-fine integration solution inspired by the classical ICP algorithm, to pairwise 3D point cloud registration with two improvements of hybrid metric spaces (eg, BSC feature and Euclidean geometry spaces)…

Computer Vision and Pattern Recognition · Computer Science 2018-08-14 Yue Pan , Bisheng Yang , Fuxun Liang , Zhen Dong

This work introduces a kernel-independent, multilevel, adaptive algorithm for efficiently evaluating a discrete convolution kernel with a given source distribution. The method is based on linear algebraic tools such as low rank…

Numerical Analysis · Mathematics 2025-07-11 Anna Yesypenko , Chao Chen , Per-Gunnar Martinsson

To prepare images for better segmentation, we need preprocessing applications, such as smoothing, to reduce noise. In this paper, we present an enhanced computation method for smoothing 2D object in binary case. Unlike existing approaches,…

Distributed, Parallel, and Cluster Computing · Computer Science 2016-03-31 Ramzi Mahmoudi , Mohamed Akil

We introduce an end-to-end learning framework for image-to-image composition, aiming to plausibly compose an object represented as a cropped patch from an object image into a background scene image. As our approach emphasizes more on…

Computer Vision and Pattern Recognition · Computer Science 2022-12-05 Hang Zhou , Rui Ma , Ling-Xiao Zhang , Lin Gao , Ali Mahdavi-Amiri , Hao Zhang

Most existing human pose estimation (HPE) methods exploit multi-scale information by fusing feature maps of four different spatial sizes, \ie $1/4$, $1/8$, $1/16$, and $1/32$ of the input image. There are two drawbacks of this strategy: 1)…

Computer Vision and Pattern Recognition · Computer Science 2021-07-23 Zhengxiong Luo , Zhicheng Wang , Yan Huang , Liang Wang , Tieniu Tan , Erjin Zhou

For every Gaussian kernel density estimator $f(x)=\sum_i a_i \exp(-\lVert x-x_i\rVert^2/2h^2)$ associated to a point cloud $\mathcal{D}=\{x_1,...,x_N\}\subset \mathbb{R}^d$, we define a nested family of closed subspaces…

Algebraic Topology · Mathematics 2024-05-02 Erik Carlsson , John Carlsson

Connected component analysis (CCA) has been heavily used to label binary images and classify segments. However, it has not been well-exploited to segment multi-valued natural images. This work proposes a novel multi-value segmentation…

Computer Vision and Pattern Recognition · Computer Science 2014-02-12 Dibyendu Mukherjee

Feature matching between image pairs is a fundamental problem in computer vision that drives many applications, such as SLAM. Recently, semi-dense matching approaches have achieved substantial performance enhancements and established a…

Computer Vision and Pattern Recognition · Computer Science 2024-11-12 Xiaolong Wang , Lei Yu , Yingying Zhang , Jiangwei Lao , Lixiang Ru , Liheng Zhong , Jingdong Chen , Yu Zhang , Ming Yang

Video processing systems such as HEVC requiring low energy consumption needed for the multimedia market has lead to extensive development in fast algorithms for the efficient approximation of 2-D DCT transforms. The DCT is employed in a…

Multimedia · Computer Science 2015-01-14 U. S. Potluri , A. Madanayake , R. J. Cintra , F. M. Bayer , S. Kulasekera , A. Edirisuriya

Purpose: The expanded encoding model incorporates spatially- and time-varying field perturbations for correction during reconstruction. So far, these reconstructions have used the conjugate gradient method with early stopping used as…

Hyperdimensional (HD) computing offers an attractive alternative to deep networks for edge learning due to its simplicity, fast prototype-based inference, and compatibility with online updates. However, standard pixel-based HD encoders are…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Arpan Kusari

The ventral, dorsal, and lateral streams in high-level human visual cortex are implicated in distinct functional processes. Yet, deep neural networks (DNNs) trained on a single task model the entire visual system surprisingly well, hinting…

Machine Learning · Computer Science 2025-10-13 Ammar I Marvi , Nancy G Kanwisher , Meenakshi Khosla

This paper describes the application of a real coded genetic algorithm (GA) to align two or more 2-D images by means of image registration. The proposed search strategy is a transformation parameters-based approach involving the affine…

Neural and Evolutionary Computing · Computer Science 2012-04-11 Mosab Bazargani , António dos Anjos , Fernando G. Lobo , Ali Mollahosseini , Hamid Reza Shahbazkia