中文
相关论文

相关论文: Scanpath Prediction in Panoramic Videos via Expect…

200 篇论文

Probabilistic Manifold Decomposition (PMD)\cite{doi:10.1137/25M1738863}, developed in our earlier work, provides a nonlinear model reduction by embedding high-dimensional dynamics onto low-dimensional probabilistic manifolds. The PMD has…

数值分析 · 数学 2026-01-13 Jiaming Guo , Dunhui Xiao

Predicting human gaze behavior within computer vision is integral for developing interactive systems that can anticipate user attention, address fundamental questions in cognitive science, and hold implications for fields like…

图像与视频处理 · 电气工程与系统科学 2024-07-02 Akash Awasthi , Ngan Le , Zhigang Deng , Rishi Agrawal , Carol C. Wu , Hien Van Nguyen

This paper presents a method for generating probabilistic descent trajectories in simulations of real-world airspace. A dataset of 116,066 trajectories harvested from Mode S radar returns in UK airspace was used to train and test the model.…

系统与控制 · 电气工程与系统科学 2025-10-09 Amy Hodgkin , Nick Pepper , Marc Thomas

In limited-view computed tomography reconstruction, iterative image reconstruction with sparsity-exploiting methods, such as total variation (TV) minimization, inspired by compressive sensing, potentially claims large reductions in sampling…

医学物理 · 物理学 2016-01-26 Bin Yan , Wenkun Zhang , Lei Li , Hanming Zhang , Linyuan Wang

Our aim is to estimate the perspective-effected geometric distortion of a scene from a video feed. In contrast to all previous work we wish to achieve this using from low-level, spatio-temporally local motion features used in commercial…

计算机视觉与模式识别 · 计算机科学 2015-04-22 Ognjen Arandjelovic , Duc-Son Pham , Svetha Venkatesh

Social intelligence is an important requirement for enabling robots to collaborate with people. In particular, human path prediction is an essential capability for robots in that it prevents potential collision with a human and allows the…

机器人学 · 计算机科学 2020-06-30 Hee-Seung Moon , Jiwon Seo

This paper proposes a simple self-supervised approach for learning a representation for visual correspondence from raw video. We cast correspondence as prediction of links in a space-time graph constructed from video. In this graph, the…

计算机视觉与模式识别 · 计算机科学 2020-12-04 Allan Jabri , Andrew Owens , Alexei A. Efros

Human movement prediction is difficult as humans naturally exhibit complex behaviors that can change drastically from one environment to the next. In order to alleviate this issue, we propose a prediction framework that decouples short-term…

机器人学 · 计算机科学 2020-03-19 Philipp Kratzer , Marc Toussaint , Jim Mainprice

Few-shot, fine-grained classification in computer vision poses significant challenges due to the need to differentiate subtle class distinctions with limited data. This paper presents a novel method that enhances the Contrastive…

计算机视觉与模式识别 · 计算机科学 2025-04-24 Eric Brouwer , Jan Erik van Woerden , Gertjan Burghouts , Matias Valdenegro-Toro , Marco Zullich

Constrained motion planning is a common but challenging problem in robotic manipulation. In recent years, data-driven constrained motion planning algorithms have shown impressive planning speed and success rate. Among them, the latent…

机器人学 · 计算机科学 2026-01-01 Jiawei Zhang , Chengchao Bai , Wei Pan , Tianhang Liu , Jifeng Guo

The performance of computer vision models in certain real-world applications (e.g., rare wildlife observation) is limited by the small number of available images. Expanding datasets using pre-trained generative models is an effective way to…

计算机视觉与模式识别 · 计算机科学 2024-12-25 Changjian Chen , Fei Lv , Yalong Guan , Pengcheng Wang , Shengjie Yu , Yifan Zhang , Zhuo Tang

Standard lossy image compression algorithms aim to preserve an image's appearance, while minimizing the number of bits needed to transmit it. However, the amount of information actually needed by a user for downstream tasks -- e.g.,…

计算机视觉与模式识别 · 计算机科学 2021-08-10 Siddharth Reddy , Anca D. Dragan , Sergey Levine

Conformal prediction provides rigorous, distribution-free uncertainty guarantees, but often yields prohibitively large prediction sets in structured domains such as routing, planning, or sequential recommendation. We introduce "graph-based…

机器学习 · 计算机科学 2026-03-31 Sreenivas Gollapudi , Kostas Kollias , Kamesh Munagala , Aravindan Vijayaraghavan

Accurate video prediction by deep neural networks, especially for dynamic regions, is a challenging task in computer vision for critical applications such as autonomous driving, remote working, and telemedicine. Due to inherent…

计算机视觉与模式识别 · 计算机科学 2024-12-05 Kazuki Kotoyori , Shota Hirose , Heming Sun , Jiro Katto

We propose to learn a probabilistic motion model from a sequence of images. Besides spatio-temporal registration, our method offers to predict motion from a limited number of frames, useful for temporal super-resolution. The model is based…

计算机视觉与模式识别 · 计算机科学 2019-09-24 Julian Krebs , Tommaso Mansi , Nicholas Ayache , Hervé Delingette

Lossy image coding standards such as JPEG and MPEG have successfully achieved high compression rates for human consumption of multimedia data. However, with the increasing prevalence of IoT devices, drones, and self-driving cars, machines…

计算机视觉与模式识别 · 计算机科学 2023-10-03 Chen-Hsiu Huang , Ja-Ling Wu

The majority of the approaches to the automatic recovery of a panoramic image from a set of partial views are suboptimal in the sense that the input images are aligned, or registered, pair by pair, e.g., consecutive frames of a video clip.…

计算机视觉与模式识别 · 计算机科学 2010-10-20 Bernardo Esteves Pires , Pedro M. Q. Aguiar

The last years have seen a surge in models predicting the scanpaths of fixations made by humans when viewing images. However, the field is lacking a principled comparison of those models with respect to their predictive power. In the past,…

计算机视觉与模式识别 · 计算机科学 2021-10-05 Matthias Kümmerer , Matthias Bethge

In this paper, we focus on the task of optimizing the parameters in Parametrized Quantum Circuits (PQCs). While popular algorithms, such as Simultaneous Perturbation Stochastic Approximation (SPSA), limit the number of circuit-execution to…

量子物理 · 物理学 2025-11-18 Sayantan Pramanik , M Girish Chandra

We address the problem of novel view video prediction; given a set of input video clips from a single/multiple views, our network is able to predict the video from a novel view. The proposed approach does not require any priors and is able…

计算机视觉与模式识别 · 计算机科学 2021-06-09 Sarah Shiraz , Krishna Regmi , Shruti Vyas , Yogesh S. Rawat , Mubarak Shah