English
Related papers

Related papers: On Ullman's theorem in computer vision

200 papers

Human motion capture either requires multi-camera systems or is unreliable when using single-view input due to depth ambiguities. Meanwhile, mirrors are readily available in urban environments and form an affordable alternative by recording…

Computer Vision and Pattern Recognition · Computer Science 2024-05-17 Daniel Ajisafe , James Tang , Shih-Yang Su , Bastian Wandt , Helge Rhodin

We introduce the simple idea of adaptive view planning to multi-view synthesis, aiming to improve both occlusion revelation and 3D consistency for single-view 3D reconstruction. Instead of producing an unordered set of views independently…

Graphics · Computer Science 2025-11-11 Yizhi Wang , Mingrui Zhao , Hao Zhang

3D localization in Multimodal Large Language Models (MLLMs), including 3D object detection and 3D visual grounding, is fundamentally limited by camera intrinsic ambiguity: the same image admits different 3D scenes under different cameras.…

Computer Vision and Pattern Recognition · Computer Science 2026-05-20 Xueying Jiang , Wenhao Li , Quanhao Qian , Deli Zhao , Shijian Lu , Gongjie Zhang , Ran Xu

The recent years have given rise to a large number of techniques for "looking around corners", i.e., for reconstructing occluded objects from time-resolved measurements of indirect light reflections off a wall. While the direct view of…

Image and Video Processing · Electrical Eng. & Systems 2023-07-19 Jonathan Klein , Martin Laurenzis , Matthias B. Hullin , Julian Iseringhausen

While object reconstruction has made great strides in recent years, current methods typically require densely captured images and/or known camera poses, and generalize poorly to novel object categories. To step toward object reconstruction…

Computer Vision and Pattern Recognition · Computer Science 2024-01-29 Hanwen Jiang , Zhenyu Jiang , Kristen Grauman , Yuke Zhu

A non-iterative auto-calibration algorithm is presented. It deals with a minimal set of six scene points in three views taken by a camera with fixed but unknown intrinsic parameters. Calibration is based on the image correspondences only.…

Computer Vision and Pattern Recognition · Computer Science 2014-11-12 Evgeniy Martyushev

We propose a self-supervised capsule architecture for 3D point clouds. We compute capsule decompositions of objects through permutation-equivariant attention, and self-supervise the process by training with pairs of randomly rotated…

Computer Vision and Pattern Recognition · Computer Science 2021-11-29 Weiwei Sun , Andrea Tagliasacchi , Boyang Deng , Sara Sabour , Soroosh Yazdani , Geoffrey Hinton , Kwang Moo Yi

Reconstructing the geometry and appearance of objects from photographs taken in different environments is difficult as the illumination and therefore the object appearance vary across captured images. This is particularly challenging for…

Computer Vision and Pattern Recognition · Computer Science 2024-12-20 Hadi Alzayer , Philipp Henzler , Jonathan T. Barron , Jia-Bin Huang , Pratul P. Srinivasan , Dor Verbin

We present a new framework for multi-view geometry in computer vision. A camera is a mapping between $\mathbb{P}^3$ and a line congruence. This model, which ignores image planes and measurements, is a natural abstraction of traditional…

Algebraic Geometry · Mathematics 2016-12-28 Jean Ponce , Bernd Sturmfels , Matthew Trager

The conical Radon transform, which assigns to a given function $f$ on $\mathbb R^3$ its integrals over conical surfaces, arises in several imaging techniques, e.g. in astronomy and homeland security, especially when the so-called Compton…

Mathematical Physics · Physics 2018-03-28 Sunghwan Moon

Estimating the pose of an object from a monocular image is an inverse problem fundamental in computer vision. The ill-posed nature of this problem requires incorporating deformation priors to solve it. In practice, many materials do not…

Computer Vision and Pattern Recognition · Computer Science 2023-03-20 Oriol Barbany , Adrià Colomé , Carme Torras

Human re-rendering from a single image is a starkly under-constrained problem, and state-of-the-art algorithms often exhibit undesired artefacts, such as over-smoothing, unrealistic distortions of the body parts and garments, or implausible…

Computer Vision and Pattern Recognition · Computer Science 2021-01-12 Kripasindhu Sarkar , Dushyant Mehta , Weipeng Xu , Vladislav Golyanik , Christian Theobalt

We study algebraic varieties associated with the camera resectioning problem. We characterize these resectioning varieties' multigraded vanishing ideals using Gr\"obner basis techniques. As an application, we derive and re-interpret…

Algebraic Geometry · Mathematics 2023-09-11 Erin Connelly , Timothy Duff , Jessie Loucks-Tavitas

Single-image piece-wise planar 3D reconstruction aims to simultaneously segment plane instances and recover 3D plane parameters from an image. Most recent approaches leverage convolutional neural networks (CNNs) and achieve promising…

Computer Vision and Pattern Recognition · Computer Science 2019-04-25 Zehao Yu , Jia Zheng , Dongze Lian , Zihan Zhou , Shenghua Gao

We propose a modular framework for single-view indoor scene 3D reconstruction, where several core modules are powered by diffusion techniques. Traditional approaches for this task often struggle with the complex instance shapes and…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Yuxiao Li

In this paper we prove theorems characterizing the decomposition of equivariant feature spaces, filters and a structural preservation theorem for invariant subspace chains in group equivariant convolutional neural networks(G-CNN).…

Representation Theory · Mathematics 2025-07-14 Bich Van Nguyen , Nguyen Cao Manh Thang

3D object detection from monocular images has proven to be an enormously challenging task, with the performance of leading systems not yet achieving even 10\% of that of LiDAR-based counterparts. One explanation for this performance gap is…

Computer Vision and Pattern Recognition · Computer Science 2018-11-21 Thomas Roddick , Alex Kendall , Roberto Cipolla

Despite the growing use of transformer models in computer vision, a mechanistic understanding of these networks is still needed. This work introduces a method to reverse-engineer Vision Transformers trained to solve image classification…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Martina G. Vilas , Timothy Schaumlöffel , Gemma Roig

We give an explicit plane-by-plane filtered back-projection reconstruction algorithm for the transverse ray transform of symmetric second rank tensor fields on Euclidean 3-space, using data from rotation about three orthogonal axes. We show…

Analysis of PDEs · Mathematics 2016-11-03 Naeem M. Desai , William R. B. Lionheart

In this paper, we propose a pipeline to generate 3D point cloud of an object from a single-view RGB image. Most previous work predict the 3D point coordinates from single RGB images directly. We decompose this problem into depth estimation…

Computer Vision and Pattern Recognition · Computer Science 2020-10-27 Wei Zeng , Sezer Karaoglu , Theo Gevers