中文
相关论文

相关论文: Look Twice: A Generalist Computational Model Predi…

200 篇论文

Scene viewing is used to study attentional selection in complex but still controlled environments. One of the main observations on eye movements during scene viewing is the inhomogeneous distribution of fixation locations: While some parts…

神经元与认知 · 定量生物学 2018-11-14 Hans A. Trukenbrod , Simon Barthelmé , Felix A. Wichmann , Ralf Engbert

Humans read texts at a varying pace, while machine learning models treat each token in the same way in terms of a computational process. Therefore, we ask, does it help to make models act more like humans? In this paper, we convert this…

计算与语言 · 计算机科学 2023-11-02 Xinting Huang , Jiajing Wan , Ioannis Kritikos , Nora Hollenstein

We study the ability of human observer to control fixation duration during execution of visual search tasks. We conducted the eye-tracking experiments with natural and synthetic images and found the dependency of fixation duration on…

神经元与认知 · 定量生物学 2020-04-24 Alexander Vasilyev

We present Sapiens, a family of models for four fundamental human-centric vision tasks -- 2D pose estimation, body-part segmentation, depth estimation, and surface normal prediction. Our models natively support 1K high-resolution inference…

计算机视觉与模式识别 · 计算机科学 2024-08-28 Rawal Khirodkar , Timur Bagautdinov , Julieta Martinez , Su Zhaoen , Austin James , Peter Selednik , Stuart Anderson , Shunsuke Saito

Predicting where people look in natural scenes has attracted a lot of interest in computer vision and computational neuroscience over the past two decades. Two seemingly contrasting categories of cues have been proposed to influence where…

计算机视觉与模式识别 · 计算机科学 2015-04-01 Ali Borji , James Tanner

Visual repetition is ubiquitous in our world. It appears in human activity (sports, cooking), animal behavior (a bee's waggle dance), natural phenomena (leaves in the wind) and in urban environments (flashing lights). Estimating visual…

计算机视觉与模式识别 · 计算机科学 2018-06-20 Tom F. H. Runia , Cees G. M. Snoek , Arnold W. M. Smeulders

The Human visual perception of the world is of a large fixed image that is highly detailed and sharp. However, receptor density in the retina is not uniform: a small central region called the fovea is very dense and exhibits high…

计算机视觉与模式识别 · 计算机科学 2017-01-26 Alon Hazan , Yuval Harel , Ron Meir

The importance of an element in a visual stimulus is commonly associated with the fixations during a free-viewing task. We argue that fixations are not always correlated with attention or awareness of visual objects. We suggest to filter…

神经元与认知 · 定量生物学 2017-12-07 Xi Wang , Marc Alexa

A major goal of computational neuroscience has been to explain how the primate ventral visual stream (VVS) transforms visual input into temporally evolving neural representations that support robust visual perception. Historically, most…

神经元与认知 · 定量生物学 2026-01-21 Matteo Dunnhofer , Maren Wehrheim , Hamidreza Ramezanpour , Sabine Muzellec , Kohitij Kar

In an experiment involving semantic search, the visual movements of sample populations subjected to visual and aural input were tracked in a taskless paradigm. The probability distributions of saccades and fixations were obtained and…

统计力学 · 物理学 2015-05-27 D. P. Shinde , Anita Mehta , R. K. Mishra

The human visual system processes images with varied degrees of resolution, with the fovea, a small portion of the retina, capturing the highest acuity region, which gradually declines toward the field of view's periphery. However, the…

计算机视觉与模式识别 · 计算机科学 2023-04-13 Beatriz Paula , Plinio Moreno

Animals (especially humans) have an amazing ability to learn new tasks quickly, and switch between them flexibly. How brains support this ability is largely unknown, both neuroscientifically and algorithmically. One reasonable supposition…

机器学习 · 计算机科学 2017-06-23 Kevin T. Feigelis , Daniel L. K. Yamins

Attention is fundamental to both biological and artificial intelligence, yet research on animal attention and AI self attention remains largely disconnected. We propose a Recurrent Vision Transformer (Recurrent ViT) that integrates…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Jonathan Morgan , Badr Albanna , James P. Herman

Humans actively observe the visual surroundings by focusing on salient objects and ignoring trivial details. However, computer vision models based on convolutional neural networks (CNN) often analyze visual input all at once through a…

计算机视觉与模式识别 · 计算机科学 2024-09-30 Minkyu Choi , Yizhen Zhang , Kuan Han , Xiaokai Wang , Zhongming Liu

The word-based account of saccades drawn by a central gravity of the PVL is supported by two pillars of evidences. The first is the finding of the initial fixation location on a word resembled a normal distribution (Rayner, 1979). The other…

神经元与认知 · 定量生物学 2015-03-19 Yanping Liu , Huan Wei

Active perception and foveal vision are the foundations of the human visual system. While foveal vision reduces the amount of information to process during a gaze fixation, active perception will change the gaze direction to the most…

计算机视觉与模式识别 · 计算机科学 2022-08-25 Alexandre M. F. Dias , Luís Simões , Plinio Moreno , Alexandre Bernardino

We aim to ask and answer an essential question "how quickly do we react after observing a displayed visual target?" To this end, we present psychophysical studies that characterize the remarkable disconnect between human saccadic behaviors…

人机交互 · 计算机科学 2022-05-06 Budmonde Duinkharjav , Praneeth Chakravarthula , Rachel Brown , Anjul Patney , Qi Sun

Recent progress in diverse intelligence has shown simple learning capacities below the organism level - single cells and even molecular networks. However, there are still many knowledge gaps around learning capacity above the organism…

种群与进化 · 定量生物学 2026-05-29 Adrita Samanta , Hananel Hazan , Michael Levin

In real-world scene perception human observers generate sequences of fixations to move image patches into the high-acuity center of the visual field. Models of visual attention developed over the last 25 years aim to predict two-dimensional…

神经元与认知 · 定量生物学 2022-08-15 Lisa Schwetlick , Daniel Backhaus , Ralf Engbert

Systems based on bag-of-words models from image features collected at maxima of sparse interest point operators have been used successfully for both computer visual object and action recognition tasks. While the sparse, interest-point based…

计算机视觉与模式识别 · 计算机科学 2013-12-31 Stefan Mathe , Cristian Sminchisescu