中文
相关论文

相关论文: A bio-inspired image coder with temporal scalabili…

200 篇论文

Screenshot-to-code generation aims to translate user interface screenshots into executable frontend code that faithfully reproduces the target layout and style. Existing multimodal large language models perform this mapping directly from…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Jie Deng , Kaichun Yao , Libo Zhang

In the past five years, the use of generative and foundational AI systems has greatly improved the decoding of brain activity. Visual perception, in particular, can now be decoded from functional Magnetic Resonance Imaging (fMRI) with…

图像与视频处理 · 电气工程与系统科学 2024-03-15 Yohann Benchetrit , Hubert Banville , Jean-Rémi King

The timing of individual neuronal spikes is essential for biological brains to make fast responses to sensory stimuli. However, conventional artificial neural networks lack the intrinsic temporal coding ability present in biological…

神经与进化计算 · 计算机科学 2020-11-18 Iulia M. Comsa , Krzysztof Potempa , Luca Versari , Thomas Fischbacher , Andrea Gesmundo , Jyrki Alakuijala

The visual system is hierarchically organized to process visual information in successive stages. Neural representations vary drastically across the first stages of visual processing: at the output of the retina, ganglion cell receptive…

神经元与认知 · 定量生物学 2019-01-07 Jack Lindsey , Samuel A. Ocko , Surya Ganguli , Stephane Deny

Understanding how neural activity gives rise to perception is a central challenge in neuroscience. We address the problem of decoding visual information from high-density intracortical recordings in primates, using the THINGS Ventral Stream…

神经元与认知 · 定量生物学 2026-01-19 Matteo Ciferri , Matteo Ferrante , Nicola Toschi

Thanks to the latest advancements in wavefront shaping, optical methods have proven crucial to achieve imaging and control light in multiply scattering media, like biological tissues. However, the stability times of living biological…

光学 · 物理学 2023-09-21 Lorenzo Valzania , Sylvain Gigan

Vision transformer has achieved impressive performance for many vision tasks. However, it may suffer from high redundancy in capturing local features for shallow layers. Local self-attention or early-stage convolutions are thus utilized,…

计算机视觉与模式识别 · 计算机科学 2024-01-26 Huaibo Huang , Xiaoqiang Zhou , Jie Cao , Ran He , Tieniu Tan

We consider the problem of ultra-low bit rate visual communication for remote vision analysis, human interactions and control in challenging scenarios with very low communication bandwidth, such as deep space exploration, battlefield…

计算机视觉与模式识别 · 计算机科学 2025-11-03 Weiming Chen , Yijia Wang , Zhihan Zhu , Zhihai He

As a bio-inspired vision sensor, the spike camera emulates the operational principles of the fovea, a compact retinal region, by employing spike discharges to encode the accumulation of per-pixel luminance intensity. Leveraging its high…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Lin Zhu , Xianzhang Chen , Xiao Wang , Hua Huang

Creativity, a process that generates novel and meaningful ideas, involves increased association between task-positive (control) and task-negative (default) networks in the human brain. Inspired by this seminal finding, in this study we…

人工智能 · 计算机科学 2020-04-24 Payel Das , Brian Quanz , Pin-Yu Chen , Jae-wook Ahn , Dhruv Shah

The random walker method for image segmentation is a popular tool for semi-automatic image segmentation, especially in the biomedical field. However, its linear asymptotic run time and memory requirements make application to 3D datasets of…

计算机视觉与模式识别 · 计算机科学 2022-08-24 Dominik Drees , Florian Eilers , Xiaoyi Jiang

Cell imaging and analysis are fundamental to biomedical research because cells are the basic functional units of life. Among different cell-related analysis, cell counting and detection are widely used. In this paper, we focus on one common…

计算机视觉与模式识别 · 计算机科学 2019-04-19 Haoyi Liang , Aijaz Naik , Cedric L. Williams , Jaideep Kapur , Daniel S. Weller

Few-step image generation has seen rapid progress, with consistency and meanflow-based methods significantly reducing the number of sampling steps. Despite their low inference cost, these approaches often suffer from training instability…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Tung Do , Thuan Hoang Nguyen , Hao Li

We present a complete system for real-time rendering of scenes with complex appearance previously reserved for offline use. This is achieved with a combination of algorithmic and system level innovations. Our appearance model utilizes…

Recently, the power of unconditional image synthesis has significantly advanced through the use of Generative Adversarial Networks (GANs). The task of inverting an image into its corresponding latent code of the trained GAN is of utmost…

计算机视觉与模式识别 · 计算机科学 2021-08-25 Yuval Alaluf , Or Patashnik , Daniel Cohen-Or

Active vision enables dynamic visual perception, offering an alternative to static feedforward architectures in computer vision, which rely on large datasets and high computational resources. Biological selective attention mechanisms allow…

计算机视觉与模式识别 · 计算机科学 2026-02-11 Giulia D'Angelo , Victoria Clerico , Chiara Bartolozzi , Matej Hoffmann , P. Michael Furlong , Alexander Hadjiivanov

Although temporal coding through spike-time patterns has long been of interest in neuroscience, the specific structures that could be useful for spike-time codes remain highly unclear. Here, we introduce a new analytical approach, using…

神经元与认知 · 定量生物学 2022-11-15 Federico W. Pasini , Alexandra N. Busch , Ján Mináč , Krishnan Padmanabhan , Lyle Muller

Decoding images from brain activity has been a challenge. Owing to the development of deep learning, there are available tools to solve this problem. The decoded image, which aims to map neural spike trains to low-level visual features and…

计算机视觉与模式识别 · 计算机科学 2022-07-19 Wenyi Li , Shengjie Zheng , Yufan Liao , Rongqi Hong , Weiliang Chen , Chenggnag He , Xiaojian Li

Visual reconstruction algorithms are an interpretive tool that map brain activity to pixels. Past reconstruction algorithms employed brute-force search through a massive library to select candidate images that, when passed through an…

神经元与认知 · 定量生物学 2023-05-03 Reese Kneeland , Jordyn Ojeda , Ghislain St-Yves , Thomas Naselaris

Transformers, the de-facto standard for language modeling, have been recently applied for vision tasks. This paper introduces sparse queries for vision transformers to exploit the intrinsic spatial redundancy of natural images and save…

计算机视觉与模式识别 · 计算机科学 2023-01-11 Lin Song , Songyang Zhang , Songtao Liu , Zeming Li , Xuming He , Hongbin Sun , Jian Sun , Nanning Zheng