English
Related papers

Related papers: Development of spatial coarse-to-fine processing i…

200 papers

Foundation models (FMs) have emerged as a powerful paradigm, enabling a diverse range of data analytics and knowledge discovery tasks across scientific fields. Inspired by the success of FMs, particularly large language models, researchers…

Machine Learning · Computer Science 2025-11-27 Sean Bin Yang , Ying Sun , Yunyao Cheng , Yan Lin , Kristian Torp , Jilin Hu

In this paper, we propose spatial propagation networks for learning the affinity matrix for vision tasks. We show that by constructing a row/column linear propagation model, the spatially varying transformation matrix exactly constitutes an…

Computer Vision and Pattern Recognition · Computer Science 2017-10-04 Sifei Liu , Shalini De Mello , Jinwei Gu , Guangyu Zhong , Ming-Hsuan Yang , Jan Kautz

In recent years, there has been increasing interest in developing models and tools to address the complex patterns of connectivity found in brain tissue. Specifically, this is due to a need to understand how emergent properties emerge from…

Neurons and Cognition · Quantitative Biology 2022-04-15 Sean Knight , Navjot Gadda

Deep generative approaches have obtained great success in image inpainting recently. However, most generative inpainting networks suffer from either over-smooth results or aliasing artifacts. The former lacks high-frequency details, while…

Computer Vision and Pattern Recognition · Computer Science 2023-07-18 Ze Lu , Yalei Lv , Wenqi Wang , Pengfei Xiong

This paper introduces an innovative end-to-end model-based deep learning approach for efficient electromagnetic analysis of high-dimensional frequency selective surfaces (FSS). Unlike traditional data-driven methods that require large…

Machine Learning · Computer Science 2024-10-23 Cheima Hammami , Lucas Polo-López , Luc Le Magoarou

In studying primate vision, a large body of work focuses on the first feedforward sweep. During this initial time window, information is thought to pass through ventral stream regions in a stage-like fashion in an effort to extract…

Neurons and Cognition · Quantitative Biology 2026-04-15 Daniel Anthes , Sushrut Thorat , Anna Mitola , Paolo Papale , Peter König , Tim C Kietzmann

Temporal human action detection aims to identify and localize action segments within untrimmed videos, serving as a pivotal task in video understanding. Despite the progress achieved by prior architectures like CNN and Transformer models,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-13 Yicheng Qiu , Keiji Yanai

Today the gold standard for in vivo imaging through scattering tissue is the point-scanning two-photon microscope (PSTPM). Especially in neuroscience, PSTPM is widely used for deep-tissue imaging in the brain. However, due to sequential…

Image and Video Processing · Electrical Eng. & Systems 2020-01-03 Zhun Wei , Josiah R. Boivin , Yi Xue , Xudong Chen , Peter T. C. So , Elly Nedivi , Dushan N. Wadduwage

Remote sensing image captioning aims to generate semantically accurate descriptions that are closely linked to the visual features of remote sensing images. Existing approaches typically emphasize fine-grained extraction of visual features…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Maofu Liu , Jiahui Liu , Xiaokang Zhang

Spatio-temporal feature encoding is essential for encoding facial expression dynamics in video sequences. At test time, most spatio-temporal encoding methods assume that a temporally segmented sequence is fed to a learned model, which could…

Computer Vision and Pattern Recognition · Computer Science 2017-11-30 Wissam J. Baddar , Yong Man Ro

Developing channel-adaptive deep joint source-channel coding (JSCC) systems is a critical challenge in wireless image transmission. While recent advancements have been made, most existing approaches are designed for static channel…

Image and Video Processing · Electrical Eng. & Systems 2024-12-12 Hanlei Li , Guangyi Zhang , Kequan Zhou , Yunlong Cai , Guanding Yu

Spatial referring is a fundamental capability of embodied robots to interact with the 3D physical world. However, even with the powerful pretrained vision language models (VLMs), recent approaches are still not qualified to accurately…

Several studies with brain signals suggested that bottom-up and top-down influences are exerted through distinct frequency bands among visual cortical areas. It has been recently shown that theta and gamma rhythms subserve feedforward,…

Neurons and Cognition · Quantitative Biology 2021-11-24 Leonardo Dalla Porta , Daniel M. Castro , Mauro Copelli , Pedro V. Carelli , Fernanda S. Matias

A linear neural network is proposed for mamalian vision system in which backward connections from the primary visual cortex (V1) to the lateral geniculate nucleus play a key role. The backward connections control the flow of information…

Neurons and Cognition · Quantitative Biology 2007-05-23 Ted Hesselroth , Klaus Schulten

Small Earth data are geoscience observations with limited short-term monitoring variability, providing sparse but meaningful measurements, typically exhibiting spatiotemporal correlations. Spatiotemporal forecasting on such data is crucial…

Machine Learning · Computer Science 2025-10-13 Yuting Yang , Gang Mei , Zhengjing Ma , Nengxiong Xu , Jianbing Peng

Supervised fine-tuning (SFT) is a critical step in aligning large language models (LLMs) with human instructions and values, yet many aspects of SFT remain poorly understood. We trained a wide range of base models on a variety of datasets…

Computation and Language · Computer Science 2025-10-31 Yuto Harada , Yusuke Yamauchi , Yusuke Oda , Yohei Oseki , Yusuke Miyao , Yu Takagi

Multimodal reasoning in vision-language models (VLMs) typically relies on a two-stage process: supervised fine-tuning (SFT) and reinforcement learning (RL). In standard SFT, all tokens contribute equally to the loss, even though reasoning…

Artificial Intelligence · Computer Science 2026-03-20 Shaked Perek , Ben Wiesel , Avihu Dekel , Nimrod Shabtay , Eli Schwartz

The primary visual cortex processes a large amount of visual information, however, due to its large receptive fields, when multiple stimuli fall within one receptive field, there are computational problems. To solve this problem, the visual…

Neurons and Cognition · Quantitative Biology 2019-04-18 Linda Wang

Deep neural networks trained on Functional Connectivity (FC) networks extracted from functional Magnetic Resonance Imaging (fMRI) data have gained popularity due to the increasing availability of data and advances in model architectures,…

Machine Learning · Computer Science 2023-12-05 Jungwon Choi , Seongho Keum , EungGu Yun , Byung-Hoon Kim , Juho Lee

The human visual system contains a hierarchical sequence of modules that take part in visual perception at different levels of abstraction, i.e., superordinate, basic, and subordinate levels. One important question is to identify the…

Neurons and Cognition · Quantitative Biology 2018-03-12 Matin N. Ashtiani , Saeed Reza Kheradpisheh , Timothée Masquelier , Mohammad Ganjtabesh
‹ Prev 1 8 9 10 Next ›