English
Related papers

Related papers: Light Cones For Vision: Simple Causal Priors For V…

200 papers

Variants of accuracy and precision are the gold-standard by which the computer vision community measures progress of perception algorithms. One reason for the ubiquity of these metrics is that they are largely task-agnostic; we in general…

Computer Vision and Pattern Recognition · Computer Science 2020-04-21 Jonah Philion , Amlan Kar , Sanja Fidler

Understanding the 3-dimensional structure of the world is a core challenge in computer vision and robotics. Neural rendering approaches learn an implicit 3D model by predicting what a camera would see from an arbitrary viewpoint. We extend…

Computer Vision and Pattern Recognition · Computer Science 2019-11-13 Josh Tobin , OpenAI Robotics , Pieter Abbeel

We introduce a canonical, compact topology, which we call weakly causal, naturally generated by the causal site of J. D. Christensen and L. Crane, a pointless algebraic structure motivated by certain problems of quantum gravity. We show…

Mathematical Physics · Physics 2013-11-14 Martin Kovár , Alena Chernikava

Spatial knowledge is a fundamental building block for the development of advanced perceptive and cognitive abilities. Traditionally, in robotics, the Euclidean (x,y,z) coordinate system and the agent's forward model are defined a priori. We…

Machine Learning · Computer Science 2020-10-30 Alban Laflaquière

This paper aims to classify and locate objects accurately and efficiently, without using bounding box annotations. It is challenging as objects in the wild could appear at arbitrary locations and in different scales. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2016-04-14 Chen Sun , Manohar Paluri , Ronan Collobert , Ram Nevatia , Lubomir Bourdev

Effective light cones, characterized by Lieb-Robinson bounds, emerge in nonrelativistic local quantum systems. Here, we present several analytical results derived from logarithmic light cones (LLCs). Possible origins of LLCs include the…

Quantum Physics · Physics 2025-09-17 Yu Zeng , Alioscia Hamma , Yu-Ran Zhang , Qiang Liu , Rengang Li , Heng Fan , Wu-Ming Liu

In General Relativity the metric can be recovered from the structure of the lightcones and a measure giving the volume element. Since the causal structure seems to be simpler than the Lorentzian manifold structure, this suggests that it is…

General Relativity and Quantum Cosmology · Physics 2015-04-28 Ovidiu Cristinel Stoica

Evolution of visual object recognition architectures based on Convolutional Neural Networks & Convolutional Deep Belief Networks paradigms has revolutionized artificial Vision Science. These architectures extract & learn the real world…

Computer Vision and Pattern Recognition · Computer Science 2015-09-08 Atul Laxman Katole , Krishna Prasad Yellapragada , Amish Kumar Bedi , Sehaj Singh Kalra , Mynepalli Siva Chaitanya

This paper describes an optimized single-stage deep convolutional neural network to detect objects in urban environments, using nothing more than point cloud data. This feature enables our method to work regardless the time of the day and…

Computer Vision and Pattern Recognition · Computer Science 2018-05-21 Kazuki Minemura , Hengfui Liau , Abraham Monrroy , Shinpei Kato

Current multimodal LLMs encode images as static visual prefixes and rely on text-based reasoning, lacking goal-driven and adaptive visual access. Inspired by human visual perception-where attention is selectively and sequentially shifted…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Guangfu Guo , Xiaoqian Lu , Yue Feng , Mingming Sun

Object-centric representations using slots have shown the advances towards efficient, flexible and interpretable abstraction from low-level perceptual features in a compositional scene. Current approaches randomize the initial state of…

Computer Vision and Pattern Recognition · Computer Science 2023-08-23 Ning Gao , Bernard Hohmann , Gerhard Neumann

The accelerating advancement of generative models has introduced new challenges for detecting AI-generated images, especially in real-world scenarios where novel generation techniques emerge rapidly. Existing learning paradigms are likely…

Computer Vision and Pattern Recognition · Computer Science 2026-03-30 Qinghui He , Haifeng Zhang , Xiuli Bi , Bo Liu , Chi-Man Pun , Bin Xiao

Slot Attention (SA) with pretrained diffusion models has recently shown promise for object-centric learning (OCL), but suffers from slot entanglement and weak alignment between object slots and image content. We propose Contrastive…

Computer Vision and Pattern Recognition · Computer Science 2026-02-20 Bac Nguyen , Yuhta Takida , Naoki Murata , Chieh-Hsin Lai , Toshimitsu Uesaka , Stefano Ermon , Yuki Mitsufuji

Accurate 3D scene description is fundamental to robotic navigation and augmented reality, yet current dense captioning methods face significant limitations in processing sparse point cloud data. % Existing approaches that apply Euclidean…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Ziyao He , Yingjie Liu , ZhangYangRui , Mingsong Chen , Xuan Tang , Xian Wei

An important, if relatively less well known aspect of the singularity theorems in Lorentzian Geometry is to understand how their conclusions fare upon weakening or suppression of one or more of their hypotheses. Then, theorems with modified…

General Relativity and Quantum Cosmology · Physics 2014-08-20 I. P. Costa e Silva , J. L. Flores

Transformer architectures are now central to sequence modeling tasks. At its heart is the attention mechanism, which enables effective modeling of long-term dependencies in a sequence. Recently, transformers have been successfully applied…

Computer Vision and Pattern Recognition · Computer Science 2022-06-16 Lin Zheng , Huijie Pan , Lingpeng Kong

I introduce a family of closeness functions between causal Lorentzian geometries of finite volume and arbitrary underlying topology. When points are randomly scattered in a Lorentzian manifold, with uniform density according to the volume…

General Relativity and Quantum Cosmology · Physics 2015-06-25 Luca Bombelli

Spatial reasoning from monocular images is essential for autonomous driving, yet current Vision-Language Models (VLMs) still struggle with fine-grained geometric perception, particularly under large scale variation and ambiguous object…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Yanchun Cheng , Rundong Wang , Xulei Yang , Alok Prakash , Daniela Rus , Marcelo H Ang , ShiJie Li

State-of-the-art two-stage object detectors apply a classifier to a sparse set of object proposals, relying on region-wise features extracted by RoIPool or RoIAlign as inputs. The region-wise features, in spite of aligning well with the…

Computer Vision and Pattern Recognition · Computer Science 2021-09-01 Zhao-Min Chen , Xin Jin , Borui Zhao , Xiu-Shen Wei , Yanwen Guo

Recent learning-based visual localization methods use global descriptors to disambiguate visually similar places, but existing approaches often derive these descriptors from geometric cues alone (e.g., covisibility graphs), limiting their…

Computer Vision and Pattern Recognition · Computer Science 2026-01-09 Son Tung Nguyen , Alejandro Fontan , Michael Milford , Tobias Fischer