中文
相关论文

相关论文: Development of spatial coarse-to-fine processing i…

200 篇论文

Convolutional Neural Networks (CNNs) have dominated the majority of computer vision tasks. However, CNNs' vulnerability to adversarial attacks has raised concerns about deploying these models to safety-critical applications. In contrast,…

计算机视觉与模式识别 · 计算机科学 2024-05-13 Keng-Hsin Liao , Chin-Yuan Yeh , Hsi-Wen Chen , Ming-Syan Chen

Spiking neural networks (SNNs) offer a biologically inspired computing paradigm with significant potential for energy-efficient neural processing. Among neural coding schemes of SNNs, Time-To-First-Spike (TTFS) coding, which encodes…

神经与进化计算 · 计算机科学 2026-03-25 Yi Lu , Jianhao Ding , Zhaofei Yu

Category-selectivity in the brain describes the observation that certain spatially localized areas of the cerebral cortex tend to respond robustly and selectively to stimuli from specific limited categories. One of the most well known…

神经元与认知 · 定量生物学 2021-12-21 T. Anderson Keller , Qinghe Gao , Max Welling

Sensory neuroscience seeks to understand how the brain encodes natural environments. However, neural coding has largely been studied using simplified stimuli. In order to assess whether the brain's coding strategy depend on the stimulus…

Learning features invariant to arbitrary transformations in the data is a requirement for any recognition system, biological or artificial. It is now widely accepted that simple cells in the primary visual cortex respond to features while…

神经与进化计算 · 计算机科学 2020-12-14 Jayanta K. Dutta , Bonny Banerjee

A primary challenge in developing synthetic spatial hearing systems, particularly underwater, is accurately modeling sound scattering. Biological organisms achieve 3D spatial hearing by exploiting sound scattering off their bodies to…

声音 · 计算机科学 2026-03-03 Siminfar Samakoush Galougah , Pranav Pulijala , Ramani Duraiswami

This study investigates the spatial reasoning capabilities of vision-language models (VLMs) through Chain-of-Thought (CoT) prompting and reinforcement learning. We begin by evaluating the impact of different prompting strategies and find…

计算机视觉与模式识别 · 计算机科学 2025-07-21 Binbin Ji , Siddharth Agrawal , Qiance Tang , Yvonne Wu

A wide range of evidence points toward the existence of a common algorithm underlying the processing of information throughout the cerebral cortex. Several hypothesized features of this cortical algorithm are reviewed, including sparse…

神经元与认知 · 定量生物学 2014-11-19 Michael R. Ferrier

Spatial transcriptomics (ST) provides spatially resolved measurements of gene expression, enabling characterization of the molecular landscape of human tissue beyond histological assessment as well as localized readouts that can be aligned…

While LLM-based TTS models exhibit zero-shot emotion and speaker cloning, their cloning fidelity and pronunciation clarity degrade on unseen domains. Fine-tuning is essential for adaptation, yet uniform approaches overlook specific…

音频与语音处理 · 电气工程与系统科学 2026-03-09 Tianrui Wang , Meng Ge , Cheng Gong , Chunyu Qiang , Haoyu Wang , Zikang Huang , Yu Jiang , Ye Ni , Yuheng Lu , Xiaobao Wang , Engsiong Chng , Xie Chen , Longbiao Wang , Jianwu Dang

Current multimodal LLMs encode images as static visual prefixes and rely on text-based reasoning, lacking goal-driven and adaptive visual access. Inspired by human visual perception-where attention is selectively and sequentially shifted…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Guangfu Guo , Xiaoqian Lu , Yue Feng , Mingming Sun

Targeted electrical stimulation of the brain perturbs neural networks and modulates their rhythmic activity both at the site of stimulation and at remote brain regions. Understanding, or even predicting, this neuromodulatory effect is…

神经元与认知 · 定量生物学 2021-06-29 Christoforos Papasavvas , Peter Neal Taylor , Yujiang Wang

Lane detection (LD) plays a crucial role in enhancing the L2+ capabilities of autonomous driving, capturing widespread attention. The Post-Processing Quantization (PTQ) could facilitate the practical application of LD models, enabling fast…

计算机视觉与模式识别 · 计算机科学 2024-05-14 Yunqian Fan , Xiuying Wei , Ruihao Gong , Yuqing Ma , Xiangguo Zhang , Qi Zhang , Xianglong Liu

Understanding the information processing roles of cortical circuits is an outstanding problem in neuroscience and artificial intelligence. The theoretical setting of Bayesian inference has been suggested as a framework for understanding…

神经元与认知 · 定量生物学 2018-08-06 Dileep George , Alexander Lavin , J. Swaroop Guntupalli , David Mely , Nick Hay , Miguel Lazaro-Gredilla

Spiking neural networks (SNN) are able to learn spatiotemporal features while using less energy, especially on neuromorphic hardware. The most widely used spiking neuron in deep learning is the Leaky Integrate and Fire (LIF) neuron. LIF…

神经与进化计算 · 计算机科学 2023-08-08 Sidi Yaya Arnaud Yarga , Sean U. N. Wood

Slow waves (SWs) are spatio-temporal patterns of cortical activity that occur both during natural sleep and anesthesia and are preserved across species. Even though electrophysiological recordings have been largely used to characterize…

Spatial transcriptomics (ST) provides high-resolution pathological images and whole-transcriptomic expression profiles at individual spots across whole-slide scales. This setting makes it an ideal data source to develop multimodal…

计算机视觉与模式识别 · 计算机科学 2024-11-27 Yuxiang Lin , Ling Luo , Ying Chen , Xushi Zhang , Zihui Wang , Wenxian Yang , Mengsha Tong , Rongshan Yu

In this paper we propose a deep learning approach for segmenting sub-cortical structures of the human brain in Magnetic Resonance (MR) image data. We draw inspiration from a state-of-the-art Fully-Convolutional Neural Network (F-CNN)…

计算机视觉与模式识别 · 计算机科学 2016-02-08 Mahsa Shakeri , Stavros Tsogkas , Enzo Ferrante , Sarah Lippe , Samuel Kadoury , Nikos Paragios , Iasonas Kokkinos

Low-light remote sensing images generally feature high resolution and high spatial complexity, with continuously distributed surface features in space. This continuity in scenes leads to extensive long-range correlations in spatial domains…

计算机视觉与模式识别 · 计算机科学 2024-09-09 Zishu Yao , Guodong Fan , Jinfu Fan , Min Gan , C. L. Philip Chen

Contemporary Vision-Language Models (VLMs) achieve strong performance on a wide range of tasks by pairing a vision encoder with a pre-trained language model, fine-tuned for visual-text inputs. Yet despite these gains, it remains unclear how…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Lachin Naghashyar , Hunar Batra , Ashkan Khakzar , Philip Torr , Ronald Clark , Christian Schroeder de Witt , Constantin Venhoff