中文
相关论文

相关论文: Modeling geometric-optical illusions: A variationa…

200 篇论文

Visual-Language Models (VLMs) have demonstrated exceptional cross-modal understanding across various tasks, including zero-shot classification, image captioning, and visual question answering. However, their robustness to physically…

计算机视觉与模式识别 · 计算机科学 2026-04-20 Chengyin Hu , Xuemeng Sun , Jiaju Han , Qike Zhang , Xiang Chen , Xin Wang , Yiwei Wei , Jiahua Long

Human-object interaction (HOI) detection aims to detect interactions between humans and objects in images. While recent advances have improved performance on existing benchmarks, their evaluations mainly focus on overall prediction accuracy…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Lemeng Wang , Qinqian Lei , Vidhi Bakshi , Daniel Yi , Yifan Liu , Jiacheng Hou , Asher Seng Hao , Zheda Mai , Wei-Lun Chao , Robby T. Tan , Bo Wang

Rendering realistic human-object interactions (HOIs) from sparse-view inputs is a challenging yet crucial task for various real-world applications. Existing methods often struggle to simultaneously achieve high rendering quality, physical…

图形学 · 计算机科学 2026-04-10 Weiquan Wang , Jun Xiao , Yi Yang , Yueting Zhuang , Long Chen

Recently, multiple formulations of vision problems as probabilistic inversions of generative models based on computer graphics have been proposed. However, applications to 3D perception from natural images have focused on low-dimensional…

计算机视觉与模式识别 · 计算机科学 2014-07-08 Tejas D. Kulkarni , Vikash K. Mansinghka , Pushmeet Kohli , Joshua B. Tenenbaum

Reconstructing 3D geometry and appearance from a sparse set of fixed cameras is a foundational task with broad applications, yet it remains fundamentally constrained by the limited viewpoints. We show that this bound can be broken by…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Ryosuke Hirai , Kohei Yamashita , Antoine Guédon , Ryo Kawahara , Vincent Lepetit , Ko Nishino

Adaptive optics normally concerns the feedback correction of phase aberrations. Such correction has been of benefit in various optical systems, with applications ranging in scale from astronomical telescopes to super-resolution microscopes.…

光学 · 物理学 2021-10-07 Chao He , Jacopo Antonello , Martin J. Booth

Vision, as an inexpensive yet information rich sensor, is commonly used for perception on autonomous mobile robots. Unfortunately, accurate vision-based perception requires a number of assumptions about the environment to hold -- some…

机器人学 · 计算机科学 2019-08-01 Sadegh Rabiee , Joydeep Biswas

Visual Inertial Odometry (VIO) is one of the most established state estimation methods for mobile platforms. However, when visual tracking fails, VIO algorithms quickly diverge due to rapid error accumulation during inertial data…

机器人学 · 计算机科学 2023-06-13 Russell Buchanan , Varun Agrawal , Marco Camurri , Frank Dellaert , Maurice Fallon

We study the effect of adversarial perturbations on the task of monocular depth prediction. Specifically, we explore the ability of small, imperceptible additive perturbations to selectively alter the perceived geometry of the scene. We…

计算机视觉与模式识别 · 计算机科学 2020-12-10 Alex Wong , Safa Cicek , Stefano Soatto

Gauss's Lemma is revised by showing that the point set association of the double tangential space with the tangential space of a Riemannian manifold is not the identity. The latter point set association is called a metrical distortion, an…

微分几何 · 数学 2026-03-09 Stephan Voellinger

The Geometry of Interaction purpose is to give a semantic of proofs or programs accounting for their dynamics. The initial presentation, translated as an algebraic weighting of paths in proofnets, led to a better characterization of the…

计算机科学中的逻辑 · 计算机科学 2008-04-10 Marc de Falco

Even when neglecting diffraction effects, the well-known equations of geometrical optics (GO) are not entirely accurate. Traditional GO treats wave rays as classical particles, which are completely described by their coordinates and…

等离子体物理 · 物理学 2017-04-05 D. E. Ruiz , I. Y. Dodin

We propose a novel diffusion-based framework for reconstructing 3D geometry of hand-held objects from monocular RGB images by leveraging hand-object interaction as geometric guidance. Our method conditions a latent diffusion model on an…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Ayce Idil Aytekin , Helge Rhodin , Rishabh Dabral , Christian Theobalt

A layout to image (L2I) generation model aims to generate a complicated image containing multiple objects (things) against natural background (stuff), conditioned on a given layout. Built upon the recent advances in generative adversarial…

计算机视觉与模式识别 · 计算机科学 2021-03-23 Sen He , Wentong Liao , Michael Ying Yang , Yongxin Yang , Yi-Zhe Song , Bodo Rosenhahn , Tao Xiang

Vision-Language Models (VLMs) occasionally generate outputs that contradict input images, constraining their reliability in real-world applications. While visual prompting is reported to suppress hallucinations by augmenting prompts with…

计算机视觉与模式识别 · 计算机科学 2025-02-24 Masayo Tomita , Katsuhiko Hayashi , Tomoyuki Kaneko

The reconstruction of gap-free signals from observation data is a critical challenge for numerous application domains, such as geoscience and space-based earth observation, when the available sensors or the data collection processes lead to…

图像与视频处理 · 电气工程与系统科学 2022-11-15 Maxime Beauchamp , Joseph Thompson , Hugo Georgenthum , Quentin Febvre , Ronan Fablet

The geometrical diffraction theory, in the sense of Keller,is here reconsidered as an obstacle problem in the Riemannian geometry. The first result is the proof of the existence and the analysis of the main properties of the diffracted…

数学物理 · 物理学 2007-05-23 Enrico De Micheli , Giacomo Monti Bragadin , Giovanni Alberto Viano

Reliable perception is fundamental for safety critical decision making in autonomous driving. Yet, vision based object detector neural networks remain vulnerable to uncertainty arising from issues such as data bias and distributional…

计算机视觉与模式识别 · 计算机科学 2025-10-21 Nishad Sahu , Shounak Sural , Aditya Satish Patil , Ragunathan , Rajkumar

Accurate geometric surface reconstruction, providing essential environmental information for navigation and manipulation tasks, is critical for enabling robotic self-exploration and interaction. Recently, 3D Gaussian Splatting (3DGS) has…

计算机视觉与模式识别 · 计算机科学 2025-07-17 Tengfei Wang , Xin Wang , Yongmao Hou , Zhaoning Zhang , Yiwei Xu , Zongqian Zhan

Despite the significant success of Large Vision-Language models(LVLMs), these models still suffer hallucinations when describing images, generating answers that include non-existent objects. It is reported that these models tend to…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Bin Li , Dehong Gao , Yeyuan Wang , Linbo Jin , Shanqing Yu , Xiaoyan Cai , Libin Yang