中文
相关论文

相关论文: Separating Components of Attention and Surprise

200 篇论文

We consider the problem of precision matrix estimation where, due to extraneous confounding of the underlying precision matrix, the data are independent but not identically distributed. While such confounding occurs in many scientific…

机器学习 · 统计学 2019-07-01 Sinong Geng , Mladen Kolar , Oluwasanmi Koyejo

The attention mechanism is a fundamental component of the Transformer model, contributing to interactions among distinct tokens, in contrast to earlier feed-forward neural networks. In general, the attention scores are determined simply by…

计算与语言 · 计算机科学 2024-10-11 Chuanyang Zheng , Yihang Gao , Han Shi , Jing Xiong , Jiankai Sun , Jingyao Li , Minbin Huang , Xiaozhe Ren , Michael Ng , Xin Jiang , Zhenguo Li , Yu Li

Models based on the Transformer architecture have achieved better accuracy than the ones based on competing architectures for a large set of tasks. A unique feature of the Transformer is its universal application of a self-attention…

机器学习 · 计算机科学 2020-10-01 Nan Ding , Xinjie Fan , Zhenzhong Lan , Dale Schuurmans , Radu Soricut

Attention is a powerful and ubiquitous mechanism for allowing neural models to focus on particular salient pieces of information by taking their weighted average when making predictions. In particular, multi-headed attention is a driving…

计算与语言 · 计算机科学 2019-11-05 Paul Michel , Omer Levy , Graham Neubig

Attention mechanisms have recently boosted performance on a range of NLP tasks. Because attention layers explicitly weight input components' representations, it is also often assumed that attention can be used to identify information that…

计算与语言 · 计算机科学 2019-06-11 Sofia Serrano , Noah A. Smith

Understanding how high-level concepts are represented within artificial neural networks is a fundamental challenge in the field of artificial intelligence. While existing literature in explainable AI emphasizes the importance of labeling…

机器学习 · 计算机科学 2024-05-17 Abhilekha Dalal , Rushrukh Rayan , Pascal Hitzler

Attention mechanisms, especially self-attention, have played an increasingly important role in deep feature representation for visual tasks. Self-attention updates the feature at each position by computing a weighted sum of features using…

计算机视觉与模式识别 · 计算机科学 2021-06-01 Meng-Hao Guo , Zheng-Ning Liu , Tai-Jiang Mu , Shi-Min Hu

In order to remain adaptable to a dynamic environment, neural activity must be simultaneously both sensitive and stable. To solve this problem, the brain has been hypothesised to sit near a critical boundary. Yet, precisely how criticality…

神经元与认知 · 定量生物学 2023-04-07 Brandon R. Munn , Eli J. Müller , James M. Shine

The appearance of an object triggers an orienting gaze movement toward its location. The movement consists of a rapid rotation of the eyes, the saccade, which is accompanied by a head rotation if the target eccentricity exceeds the…

神经元与认知 · 定量生物学 2024-09-17 Laurent Goffart

Abrupt shifts in ecosystems, brains, markets, and climate are often diagnosed as signs of approaching a tipping point, i.e. a critical bifurcation where stability is lost. Here we reveal a broader and more deceptive mechanism:…

混沌动力学 · 物理学 2025-10-06 Virgile Troude , Sandro Claudio Lera , Ke Wu , Didier Sornette

In simple perceptual decisions the brain has to identify a stimulus based on noisy sensory samples from the stimulus. Basic statistical considerations state that the reliability of the stimulus information, i.e., the amount of noise in the…

神经元与认知 · 定量生物学 2015-09-08 Sebastian Bitzer , Stefan J. Kiebel

To date, it is still impossible to sample the entire mammalian brain with single-neuron precision. This forces one to either use spikes (focusing on few neurons) or to use coarse-sampled activity (averaging over many neurons, e.g. LFP).…

神经元与认知 · 定量生物学 2022-12-12 Joao Pinheiro Neto , Franz Paul Spitzner , Viola Priesemann

Background. From information theory, surprisal is a measurement of how unexpected an event is. Statistical language models provide a probabilistic approximation of natural languages, and because surprisal is constructed with the probability…

计算与语言 · 计算机科学 2022-04-18 James Caddy , Markus Wagner , Christoph Treude , Earl T. Barr , Miltiadis Allamanis

Adaptive streaming of 360-degree video relies on viewport prediction to allocate bandwidth efficiently. Current approaches predominantly use visual saliency or historical gaze patterns, neglecting the role of spatial audio in guiding user…

多媒体 · 计算机科学 2026-01-07 Arman Nik Khah , Ravi Prakash

In neural-decoding studies, recordings of participants' responses to stimuli are used to train models. In recent years, there has been an explosion of publications detailing applications of innovations from deep-learning research to…

神经元与认知 · 定量生物学 2025-08-04 Jack A. Kilgallen , Barak A. Pearlmutter , Jeffrey Mark Siskind

Our ability to track multiple objects in a dynamic environment enables us to perform everyday tasks such as driving, playing team sports, and walking in a crowded mall. Despite more than three decades of literature on multiple object…

神经元与认知 · 定量生物学 2022-08-01 Yannick Roy , Jocelyn Faubert

Attention mechanisms have seen wide adoption in neural NLP models. In addition to improving predictive performance, these are often touted as affording transparency: models equipped with attention provide a distribution over attended-to…

计算与语言 · 计算机科学 2019-05-10 Sarthak Jain , Byron C. Wallace

Selective attention enables humans to efficiently process visual stimuli by enhancing important elements and filtering out irrelevant information. Locating visual attention is fundamental in neuroscience with potential applications in…

信号处理 · 电气工程与系统科学 2025-09-19 Yuanyuan Yao , Wout De Swaef , Simon Geirnaert , Alexander Bertrand

In safety-critical applications like medical diagnosis, certainty associated with a model's prediction is just as important as its accuracy. Consequently, uncertainty estimation and reduction play a crucial role. Uncertainty in predictions…

图像与视频处理 · 电气工程与系统科学 2023-09-12 Abhishek Singh Sambyal , Narayanan C. Krishnan , Deepti R. Bathula

This paper proposes a neuronal circuitry layout and synaptic plasticity principles that allow the (pyramidal) neuron to act as a "combinatorial switch". Namely, the neuron learns to be more prone to generate spikes given those combinations…

生物物理 · 物理学 2017-05-09 Marat M. Rvachev