中文
相关论文

相关论文: Linton Stereo Illusion: Response on Johnston (1991…

200 篇论文

We present a new illusion that challenges our understanding of stereo vision. The illusion consists of a larger circle at 50cm, and smaller circle in front of it at 40cm, with constant angular sizes throughout. We move the larger circle…

神经元与认知 · 定量生物学 2026-02-20 Paul Linton

Video see-through (VST) technology aims to seamlessly blend virtual and physical worlds by reconstructing reality through cameras. While manufacturers promise perceptual fidelity, it remains unclear how close these systems are to…

人机交互 · 计算机科学 2026-01-07 Jialin Wang , Songming Ping , Kemu Xu , Yue Li , Hai-Ning Liang

Stereo matching is one of the longest-standing problems in computer vision with close to 40 years of studies and research. Throughout the years the paradigm has shifted from local, pixel-level decision to various forms of discrete and…

计算机视觉与模式识别 · 计算机科学 2021-04-01 Matteo Poggi , Fabio Tosi , Konstantinos Batsos , Philippos Mordohai , Stefano Mattoccia

Visual translation tolerance refers to our capacity to recognize objects over a wide range of different retinal locations. Although translation is perhaps the simplest spatial transform that the visual system needs to cope with, the extent…

神经元与认知 · 定量生物学 2020-12-09 Ryan Blything , Valerio Biscione , Ivan I. Vankov , Casimir J. H. Ludwig , Jeffrey S. Bowers

The Expanding Hole Illusion is a compelling visual phenomenon in which a static, concentric pattern evokes a strong perception of continuous forward motion. Despite its simplicity, this illusion challenges our understanding of how the brain…

神经元与认知 · 定量生物学 2025-01-16 Nasim Nematzadeh , David M. W. Powers

We revisit the problem of visual depth estimation in the context of autonomous vehicles. Despite the progress on monocular depth estimation in recent years, we show that the gap between monocular and stereo depth accuracy remains large$-$a…

计算机视觉与模式识别 · 计算机科学 2020-07-09 Nikolai Smolyanskiy , Alexey Kamenev , Stan Birchfield

Novel view synthesis is an important problem in computer vision and graphics. Over the years a large number of solutions have been put forward to solve the problem. However, the large-baseline novel view synthesis problem is far from being…

计算机视觉与模式识别 · 计算机科学 2018-05-08 Tewodros Habtegebrial , Kiran Varanasi , Christian Bailer , Didier Stricker

While GPT-4V(ision) impressively models both visual and textual information simultaneously, it's hallucination behavior has not been systematically assessed. To bridge this gap, we introduce a new benchmark, namely, the Bias and…

机器学习 · 计算机科学 2023-11-08 Chenhang Cui , Yiyang Zhou , Xinyu Yang , Shirley Wu , Linjun Zhang , James Zou , Huaxiu Yao

By comparing biological and artificial perception through the lens of illusions, we highlight critical differences in how each system constructs visual reality. Understanding these divergences can inform the development of more robust,…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Jianyi Yang , Junyi Ye , Ankan Dash , Guiling Wang

Multimodal Large Reasoning Models (MLRMs) have achieved remarkable strides in visual reasoning through test time compute scaling, yet long chain reasoning remains prone to hallucinations. We identify a concerning phenomenon termed the…

人工智能 · 计算机科学 2026-05-29 Zhe Qian , Yanbiao Ma , Zhuohan Ouyang , Zhonghua Wang , Zhongxing Xu , Fei Luo , Xinyu Liu , Zongyuan Ge , Yike Guo , Jungong Han

Illusions are entertaining, but they are also a useful diagnostic tool in cognitive science, philosophy, and neuroscience. A typical illusion shows a gap between how something "really is" and how something "appears to be", and this gap…

神经元与认知 · 定量生物学 2024-12-30 Tomer Ullman

In parallel with the success of CNNs to solve vision problems, there is a growing interest in developing methodologies to understand and visualize the internal representations of these networks. How the responses of a trained CNN encode the…

计算机视觉与模式识别 · 计算机科学 2015-11-18 Ivet Rafegas , Maria Vanrell

Vision-and-Language Navigation (VLN) tasks agents with locating specific objects in unseen environments using natural language instructions and visual cues. Many existing VLN approaches typically follow an 'observe-and-reason' schema, that…

机器人学 · 计算机科学 2026-02-04 Yanjia Huang , Mingyang Wu , Renjie Li , Zhengzhong Tu

Neural networks have greatly boosted performance in computer vision by learning powerful representations of input data. The drawback of end-to-end training for maximal overall performance are black-box models whose hidden representations…

计算机视觉与模式识别 · 计算机科学 2020-04-29 Patrick Esser , Robin Rombach , Björn Ommer

The visual distortion effects visible to an observer traveling around and descending to the surface of an extremely compact star are described. Specifically, trips to a ``normal" neutron star, a black hole, and an ultracompact neutron star…

天体物理学 · 物理学 2008-11-26 Robert J. Nemiroff

Reconstructing the 3D shape of an object using several images under different light sources is a very challenging task, especially when realistic assumptions such as light propagation and attenuation, perspective viewing geometry and…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Fotios Logothetis , Roberto Mecca , Ignas Budvytis , Roberto Cipolla

Vision-and-Language Navigation (VLN) is a cornerstone of embodied intelligence. However, current agents often suffer from significant performance degradation when transitioning from simulation to real-world deployment, primarily due to…

Hallucination has been widely recognized to be a significant drawback for large language models (LLMs). There have been many works that attempt to reduce the extent of hallucination. These efforts have mostly been empirical so far, which…

计算与语言 · 计算机科学 2025-02-14 Ziwei Xu , Sanjay Jain , Mohan Kankanhalli

The Large Synoptic Survey Telescope (LSST) is conceived as an 8.4-m telescope with CCD or CMOS focal plane covering most of a field 0.6 m in diameter, the latter exceeding the size of the largest photographic plates ever used in astronomy.…

天体物理学 · 物理学 2015-06-10 Alistair R. Walker

Stereo vision generally involves the computation of pixel correspondences and estimation of disparities between rectified image pairs. In many applications, including simultaneous localization and mapping (SLAM) and 3D object detection, the…

计算机视觉与模式识别 · 计算机科学 2020-11-11 WeiQin Chuah , Ruwan Tennakoon , Reza Hoseinnezhad , Alireza Bab-Hadiashar , David Suter
‹ 上一页 1 2 3 10 下一页 ›