中文
相关论文

相关论文: On Ullman's theorem in computer vision

200 篇论文

Human motion capture either requires multi-camera systems or is unreliable when using single-view input due to depth ambiguities. Meanwhile, mirrors are readily available in urban environments and form an affordable alternative by recording…

计算机视觉与模式识别 · 计算机科学 2024-05-17 Daniel Ajisafe , James Tang , Shih-Yang Su , Bastian Wandt , Helge Rhodin

We introduce the simple idea of adaptive view planning to multi-view synthesis, aiming to improve both occlusion revelation and 3D consistency for single-view 3D reconstruction. Instead of producing an unordered set of views independently…

图形学 · 计算机科学 2025-11-11 Yizhi Wang , Mingrui Zhao , Hao Zhang

3D localization in Multimodal Large Language Models (MLLMs), including 3D object detection and 3D visual grounding, is fundamentally limited by camera intrinsic ambiguity: the same image admits different 3D scenes under different cameras.…

计算机视觉与模式识别 · 计算机科学 2026-05-20 Xueying Jiang , Wenhao Li , Quanhao Qian , Deli Zhao , Shijian Lu , Gongjie Zhang , Ran Xu

The recent years have given rise to a large number of techniques for "looking around corners", i.e., for reconstructing occluded objects from time-resolved measurements of indirect light reflections off a wall. While the direct view of…

图像与视频处理 · 电气工程与系统科学 2023-07-19 Jonathan Klein , Martin Laurenzis , Matthias B. Hullin , Julian Iseringhausen

While object reconstruction has made great strides in recent years, current methods typically require densely captured images and/or known camera poses, and generalize poorly to novel object categories. To step toward object reconstruction…

计算机视觉与模式识别 · 计算机科学 2024-01-29 Hanwen Jiang , Zhenyu Jiang , Kristen Grauman , Yuke Zhu

A non-iterative auto-calibration algorithm is presented. It deals with a minimal set of six scene points in three views taken by a camera with fixed but unknown intrinsic parameters. Calibration is based on the image correspondences only.…

计算机视觉与模式识别 · 计算机科学 2014-11-12 Evgeniy Martyushev

We propose a self-supervised capsule architecture for 3D point clouds. We compute capsule decompositions of objects through permutation-equivariant attention, and self-supervise the process by training with pairs of randomly rotated…

计算机视觉与模式识别 · 计算机科学 2021-11-29 Weiwei Sun , Andrea Tagliasacchi , Boyang Deng , Sara Sabour , Soroosh Yazdani , Geoffrey Hinton , Kwang Moo Yi

Reconstructing the geometry and appearance of objects from photographs taken in different environments is difficult as the illumination and therefore the object appearance vary across captured images. This is particularly challenging for…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Hadi Alzayer , Philipp Henzler , Jonathan T. Barron , Jia-Bin Huang , Pratul P. Srinivasan , Dor Verbin

We present a new framework for multi-view geometry in computer vision. A camera is a mapping between $\mathbb{P}^3$ and a line congruence. This model, which ignores image planes and measurements, is a natural abstraction of traditional…

代数几何 · 数学 2016-12-28 Jean Ponce , Bernd Sturmfels , Matthew Trager

The conical Radon transform, which assigns to a given function $f$ on $\mathbb R^3$ its integrals over conical surfaces, arises in several imaging techniques, e.g. in astronomy and homeland security, especially when the so-called Compton…

数学物理 · 物理学 2018-03-28 Sunghwan Moon

Estimating the pose of an object from a monocular image is an inverse problem fundamental in computer vision. The ill-posed nature of this problem requires incorporating deformation priors to solve it. In practice, many materials do not…

计算机视觉与模式识别 · 计算机科学 2023-03-20 Oriol Barbany , Adrià Colomé , Carme Torras

Human re-rendering from a single image is a starkly under-constrained problem, and state-of-the-art algorithms often exhibit undesired artefacts, such as over-smoothing, unrealistic distortions of the body parts and garments, or implausible…

计算机视觉与模式识别 · 计算机科学 2021-01-12 Kripasindhu Sarkar , Dushyant Mehta , Weipeng Xu , Vladislav Golyanik , Christian Theobalt

We study algebraic varieties associated with the camera resectioning problem. We characterize these resectioning varieties' multigraded vanishing ideals using Gr\"obner basis techniques. As an application, we derive and re-interpret…

代数几何 · 数学 2023-09-11 Erin Connelly , Timothy Duff , Jessie Loucks-Tavitas

Single-image piece-wise planar 3D reconstruction aims to simultaneously segment plane instances and recover 3D plane parameters from an image. Most recent approaches leverage convolutional neural networks (CNNs) and achieve promising…

计算机视觉与模式识别 · 计算机科学 2019-04-25 Zehao Yu , Jia Zheng , Dongze Lian , Zihan Zhou , Shenghua Gao

We propose a modular framework for single-view indoor scene 3D reconstruction, where several core modules are powered by diffusion techniques. Traditional approaches for this task often struggle with the complex instance shapes and…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Yuxiao Li

In this paper we prove theorems characterizing the decomposition of equivariant feature spaces, filters and a structural preservation theorem for invariant subspace chains in group equivariant convolutional neural networks(G-CNN).…

表示论 · 数学 2025-07-14 Bich Van Nguyen , Nguyen Cao Manh Thang

3D object detection from monocular images has proven to be an enormously challenging task, with the performance of leading systems not yet achieving even 10\% of that of LiDAR-based counterparts. One explanation for this performance gap is…

计算机视觉与模式识别 · 计算机科学 2018-11-21 Thomas Roddick , Alex Kendall , Roberto Cipolla

Despite the growing use of transformer models in computer vision, a mechanistic understanding of these networks is still needed. This work introduces a method to reverse-engineer Vision Transformers trained to solve image classification…

计算机视觉与模式识别 · 计算机科学 2023-10-31 Martina G. Vilas , Timothy Schaumlöffel , Gemma Roig

We give an explicit plane-by-plane filtered back-projection reconstruction algorithm for the transverse ray transform of symmetric second rank tensor fields on Euclidean 3-space, using data from rotation about three orthogonal axes. We show…

偏微分方程分析 · 数学 2016-11-03 Naeem M. Desai , William R. B. Lionheart

In this paper, we propose a pipeline to generate 3D point cloud of an object from a single-view RGB image. Most previous work predict the 3D point coordinates from single RGB images directly. We decompose this problem into depth estimation…

计算机视觉与模式识别 · 计算机科学 2020-10-27 Wei Zeng , Sezer Karaoglu , Theo Gevers