中文
相关论文

相关论文: Cognitive-mapping and contextual pyramid based Dig…

200 篇论文

We present Spatial Lifting (SL), a novel methodology for dense prediction tasks. SL operates by lifting standard inputs, such as 2D images, into a higher-dimensional space and subsequently processing them using networks designed for that…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Mingzhi Xu , Yizhe Zhang

In cognitive science and AI, a longstanding question is whether machines learn representations that align with those of the human mind. While current models show promise, it remains an open question whether this alignment is superficial or…

神经元与认知 · 定量生物学 2025-10-27 Craig Sanders , Billy Dickson , Sahaj Singh Maini , Robert Nosofsky , Zoran Tiganj

Recently, deep-learning-based approaches have been widely studied for deformable image registration task. However, most efforts directly map the composite image representation to spatial transformation through the convolutional neural…

图像与视频处理 · 电气工程与系统科学 2022-07-08 Jiashun Chen , Donghuan Lu , Yu Zhang , Dong Wei , Munan Ning , Xinyu Shi , Zhe Xu , Yefeng Zheng

Cephalometric tracing method is usually used in orthodontic diagnosis and treatment planning. In this paper, we propose a deep learning based framework to automatically detect anatomical landmarks in cephalometric X-ray images. We train the…

图像与视频处理 · 电气工程与系统科学 2020-09-30 Zhusi Zhong , Jie Li , Zhenxi Zhang , Zhicheng Jiao , Xinbo Gao

We provide a dataset for enabling Deep Generative Models (DGMs) in engineering design and propose methods to automate data labeling by utilizing large-scale foundation models. GeoBiked is curated to contain 4 355 bicycle images, annotated…

计算机视觉与模式识别 · 计算机科学 2025-05-23 Phillip Mueller , Sebastian Mueller , Lars Mikelsons

We propose a novel landmarks-assisted collaborative end-to-end deep framework for automatic 4D FER. Using 4D face scan data, we calculate its various geometrical images, and afterwards use rank pooling to generate their dynamic images…

计算机视觉与模式识别 · 计算机科学 2020-02-10 Muzammil Behzad , Nhat Vo , Xiaobai Li , Guoying Zhao

Radiological images such as computed tomography (CT) and X-rays render anatomy with intrinsic structures. Being able to reliably locate the same anatomical structure across varying images is a fundamental task in medical image analysis. In…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Ke Yan , Jinzheng Cai , Dakai Jin , Shun Miao , Dazhou Guo , Adam P. Harrison , Youbao Tang , Jing Xiao , Jingjing Lu , Le Lu

In deformable object manipulation, we often want to interact with specific segments of an object that are only defined in non-deformed models of the object. We thus require a system that can recognize and locate these segments in sensor…

计算机视觉与模式识别 · 计算机科学 2023-11-14 Pit Henrich , Balázs Gyenes , Paul Maria Scheikl , Gerhard Neumann , Franziska Mathis-Ullrich

Use of the electroencephalogram (EEG) and machine learning approaches to recognize emotions can facilitate affective human computer interactions. However, the type of EEG data constitutes an obstacle for cross-individual EEG feature…

机器学习 · 计算机科学 2021-05-26 Xiaolong Zhong , Zhong Yin

The domain of computer vision has experienced significant advancements in facial-landmark detection, becoming increasingly essential across various applications such as augmented reality, facial recognition, and emotion analysis. Unlike…

计算机视觉与模式识别 · 计算机科学 2024-04-10 Zong-Wei Hong , Yu-Chen Lin

Motivated by recent findings from cognitive neural science, we advocate the use of a dual-level model for concept representations: the embodied level consists of concept-oriented feature representations, and the symbolic level consists of…

机器学习 · 计算机科学 2022-03-02 Daniel T. Chang

Reconstructing and understanding 3D structures from a limited number of images is a well-established problem in computer vision. Traditional methods usually break this task into multiple subtasks, each requiring complex transformations…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Zhiwen Fan , Jian Zhang , Wenyan Cong , Peihao Wang , Renjie Li , Kairun Wen , Shijie Zhou , Achuta Kadambi , Zhangyang Wang , Danfei Xu , Boris Ivanovic , Marco Pavone , Yue Wang

In this paper, we introduce a contextual grounding approach that captures the context in corresponding text entities and image regions to improve the grounding accuracy. Specifically, the proposed architecture accepts pre-trained text token…

计算机视觉与模式识别 · 计算机科学 2019-11-07 Farley Lai , Ning Xie , Derek Doran , Asim Kadav

A cognitive map is an internal model which encodes the abstract relationships among entities in the world, giving humans and animals the flexibility to adapt to new situations, with a strong out-of-distribution (OOD) generalization that…

机器学习 · 计算机科学 2026-05-12 Victor Rambaud , Salvador Mascarenhas , Yair Lakretz

When approaching the semantic segmentation of overhead imagery in the decimeter spatial resolution range, successful strategies usually combine powerful methods to learn the visual appearance of the semantic classes (e.g. convolutional…

计算机视觉与模式识别 · 计算机科学 2018-08-24 Michele Volpi , Devis Tuia

Precise volumetric delineation of hippocampal structures is essential for quantifying neurodevelopmental trajectories in pre-term and term infants, where subtle morphological variations may carry prognostic significance. While foundation…

图像与视频处理 · 电气工程与系统科学 2026-03-02 Annayah Usman , Behraj Khan , Tahir Qasim Syed

Effective human-robot interaction, such as in robot learning from human demonstration, requires the learning agent to be able to ground abstract concepts (such as those contained within instructions) in a corresponding high-dimensional…

计算机视觉与模式识别 · 计算机科学 2018-10-03 Yordan Hristov , Alex Lascarides , Subramanian Ramamoorthy

Contextual information has been shown to be powerful for semantic segmentation. This work proposes a novel Context-based Tandem Network (CTNet) by interactively exploring the spatial contextual information and the channel contextual…

计算机视觉与模式识别 · 计算机科学 2021-04-21 Zechao Li , Yanpeng Sun , Jinhui Tang

We propose a modular and scalable framework for dense coregistration and cosegmentation with two key characteristics: first, we substitute ground truth data with the semantic map output of a classifier; second, we combine this output with…

计算机视觉与模式识别 · 计算机科学 2016-07-25 Mahsa Shakeri , Enzo Ferrante , Stavros Tsogkas , Sarah Lippe , Samuel Kadoury , Iasonas Kokkinos , Nikos Paragios

Semantic representations can be framed as a structured, dynamic knowledge space through which humans navigate to retrieve and manipulate meaning. To investigate how humans traverse this geometry, we introduce a framework that represents…

计算与语言 · 计算机科学 2026-04-15 Felipe D. Toro-Hernández , Jesuino Vieira Filho , Rodrigo M. Cabral-Carvalho