中文
相关论文

相关论文: A sub-Riemannian model of the visual cortex with f…

200 篇论文

Recent advances in multimodal large language models (MLLMs) have enabled impressive progress in vision-language understanding, yet their high computational cost limits deployment in resource-constrained scenarios such as robotic…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Quoc-Huy Trinh

We show how a Hopfield network with modifiable recurrent connections undergoing slow Hebbian learning can extract the underlying geometry of an input space. First, we use a slow/fast analysis to derive an averaged system whose dynamics…

神经元与认知 · 定量生物学 2011-02-02 Mathieu N. Galtier , Olivier D. Faugeras , Paul C. Bressloff

Human vision possesses a special type of visual processing systems called peripheral vision. Partitioning the entire visual field into multiple contour regions based on the distance to the center of our gaze, the peripheral vision provides…

计算机视觉与模式识别 · 计算机科学 2022-10-14 Juhong Min , Yucheng Zhao , Chong Luo , Minsu Cho

We consider the evolution model proposed in [9, 6] to describe illusory contrast perception phenomena induced by surrounding orientations. Firstly, we highlight its analogies and differences with the widely used Wilson-Cowan equations [48],…

计算机视觉与模式识别 · 计算机科学 2020-07-20 Marcelo Bertalmío , Luca Calatroni , Valentina Franceschi , Benedetta Franceschiello , Dario Prandi

The idea of using the recurrent neural network for visual attention has gained popularity in computer vision community. Although the recurrent attention model (RAM) leverages the glimpses with more large patch size to increasing its scope,…

计算机视觉与模式识别 · 计算机科学 2021-11-16 Gang Chen

Immersive video offers the freedom to navigate inside virtualized environment. Instead of streaming the bulky immersive videos entirely, a viewport (also referred to as field of view, FoV) adaptive streaming is preferred. We often stream…

多媒体 · 计算机科学 2018-02-19 Shaowei Xie , Qiu Shen , Yiling Xu , Qiaojian Qian , Shaowei Wang , Zhan Ma , Wenjun Zhang

Human vision is foveated, with variable resolution peaking at the center of a large field of view; this reflects an efficient trade-off for active sensing, allowing eye-movements to bring different parts of the world into focus with other…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Nicholas M. Blauch , George A. Alvarez , Talia Konkle

The transformer models have shown promising effectiveness in dealing with various vision tasks. However, compared with training Convolutional Neural Network (CNN) models, training Vision Transformer (ViT) models is more difficult and relies…

计算机视觉与模式识别 · 计算机科学 2022-07-28 Jiawang Bai , Li Yuan , Shu-Tao Xia , Shuicheng Yan , Zhifeng Li , Wei Liu

We have put forwards a unified quantitative framework of vision and audition, based on existing data and theories. According to this model, the retina is a feedforward network self-adaptive to inputs in a specific period. After fully grown,…

计算机视觉与模式识别 · 计算机科学 2014-06-24 Peilei Liu , Ting Wang

Receptive profiles of V1 cortical cells are very heterogeneous and act by differentiating the stimulus image as operators changing from point to point. A lightness and color constancy image can be reconstructed as the solution of the…

偏微分方程分析 · 数学 2023-08-01 Alessandro Sarti , Mattia Galeotti , Giovanna Citti

In this paper, we will study the following pattern recognition problem: Every pattern is a 3-dimensional graph, its surface can be split up into some regions, every region is composed of the pixels with the approximately same colour value…

神经元与认知 · 定量生物学 2017-03-07 YongHong Chen

Recently, we put forwarded a redox molecular hypothesis involving the natural biophysical substrate of visual perception and imagery. Here, we explicitly propose that the feedback and feedforward iterative operation processes can be…

神经元与认知 · 定量生物学 2011-03-31 I. Bokkon , V. Salari , J. Tuszynski

Low-light image enhancement is a promising solution to tackle the problem of insufficient sensitivity of human vision system (HVS) to perceive information in low light environments. Previous Retinex-based works always accomplish enhancement…

图像与视频处理 · 电气工程与系统科学 2020-05-18 Xiaoxiao Li , Xiaopeng Guo , Liye Mei , Mingyu Shang , Jie Gao , Maojing Shu , Xiang Wang

We report on simultaneous recordings from cells in all layers of visual cortex and models developed to capture the higher order structure of population spiking activity. Specifically, we use Ising, Restricted Boltzmann Machine (RBM) and…

神经元与认知 · 定量生物学 2013-01-03 Urs Köster , Jascha Sohl-Dickstein , Charles M. Gray , Bruno A. Olshausen

Decoding visual stimuli from neural responses recorded by functional Magnetic Resonance Imaging (fMRI) presents an intriguing intersection between cognitive neuroscience and machine learning, promising advancements in understanding human…

计算机视觉与模式识别 · 计算机科学 2023-12-29 Jingyuan Sun , Mingxiao Li , Zijiao Chen , Yunhao Zhang , Shaonan Wang , Marie-Francine Moens

Neural activity in the visual cortex of blind humans persists in the absence of visual stimuli. However, little is known about the preservation of visual representation capacity in these cortical regions, which could have significant…

Some biological mechanisms of early vision are comparatively well understood, but they have yet to be evaluated for their ability to accurately predict and explain human judgments of image similarity. From well-studied simple connectivity…

计算机视觉与模式识别 · 计算机科学 2020-08-18 Elijah Bowen , Antonio Rodriguez , Damian Sowinski , Richard Granger

Equipping the rototranslation group $SE(2)$ with a sub-Riemannian structure inspired by the visual cortex V1, we propose algorithms for image inpainting and enhancement based on hypoelliptic diffusion. We innovate on previous…

计算机视觉与模式识别 · 计算机科学 2024-11-28 Francesco Ballerin , Erlend Grong

Convolutional neural networks (CNNs) are remarkably successful in many computer vision tasks. However, the high cost of inference is problematic for embedded and real-time systems, so there are many studies on compressing the networks. On…

计算机视觉与模式识别 · 计算机科学 2021-12-15 Akihiro Imamura , Nana Arizumi

While Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in general visual understanding, they frequently falter in fine-grained perception tasks that require identifying tiny objects or discerning subtle…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Jilong Zhu , Yang Feng