English
Related papers

Related papers: Disentangling Hippocampal Shape Variations: A Stud…

200 papers

We propose a sequential variational autoencoder to learn disentangled representations of sequential data (e.g., videos and audios) under self-supervision. Specifically, we exploit the benefits of some readily accessible supervisory signals…

Computer Vision and Pattern Recognition · Computer Science 2020-05-26 Yizhe Zhu , Martin Renqiang Min , Asim Kadav , Hans Peter Graf

Representation learning constitutes a pivotal cornerstone in contemporary deep learning paradigms, offering a conduit to elucidate distinctive features within the latent space and interpret the deep models. Nevertheless, the inherent…

Computer Vision and Pattern Recognition · Computer Science 2024-02-07 Siyuan Dai , Kai Ye , Kun Zhao , Ge Cui , Haoteng Tang , Liang Zhan

The scarcity of training data and the large speaker variation in dysarthric speech lead to poor accuracy and poor speaker generalization of spoken language understanding systems for dysarthric speech. Through work on the speech features, we…

Audio and Speech Processing · Electrical Eng. & Systems 2022-10-25 Jinzi Qi , Hugo Van hamme

Susceptibility tensor imaging (STI) is an emerging magnetic resonance imaging technique that characterizes the anisotropic tissue magnetic susceptibility with a second-order tensor model. STI has the potential to provide information for…

Image and Video Processing · Electrical Eng. & Systems 2022-09-13 Zhenghan Fang , Kuo-Wei Lai , Peter van Zijl , Xu Li , Jeremias Sulam

Diffusion kurtosis imaging (DKI), is an imaging modality that yields novel disease biomarkers and in combination with nervous tissue modeling, provides access to microstructural parameters. Recently, DKI and subsequent estimation of…

Quantitative Methods · Quantitative Biology 2019-04-09 Andrey Chuhutin , Brian Hansen , Agnieszka Wlodarczyk , Trevor Owens , Noam Shemesh , Sune Nørhøj Jespersen

Generative modeling of 3D brain MRIs presents difficulties in achieving high visual fidelity while ensuring sufficient coverage of the data distribution. In this work, we propose to address this challenge with composable, multiscale…

Image and Video Processing · Electrical Eng. & Systems 2023-01-12 Jaivardhan Kapoor , Jakob H. Macke , Christian F. Baumgartner

Leveraging the fact that speaker identity and content vary on different time scales, \acrlong{fhvae} (\acrshort{fhvae}) uses different latent variables to symbolize these two attributes. Disentanglement of these attributes is carried out by…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-16 Yuying Xie , Thomas Arildsen , Zheng-Hua Tan

We present a new approach for representing and reconstructing multidimensional magnetic resonance imaging (MRI) data. Our method builds on a novel, learned feature-based image representation that disentangles different types of features,…

Image and Video Processing · Electrical Eng. & Systems 2026-01-01 Ruiyang Zhao , Fan Lam

Variational autoencoders (VAEs) learn representations of data by jointly training a probabilistic encoder and decoder network. Typically these models encode all features of the data into a single variable. Here we are interested in learning…

Purpose. Brain Magnetic Resonance Images (MRIs) are essential for the diagnosis of neurological diseases. Recently, deep learning methods for unsupervised anomaly detection (UAD) have been proposed for the analysis of brain MRI. These…

Image and Video Processing · Electrical Eng. & Systems 2021-09-15 Marcel Bengs , Finn Behrendt , Julia Krüger , Roland Opfer , Alexander Schlaefer

Sensory input from multiple sources is crucial for robust and coherent human perception. Different sources contribute complementary explanatory factors. Similarly, research studies often collect multimodal imaging data, each of which can…

Contrastive Analysis VAE (CA-VAEs) is a family of Variational auto-encoders (VAEs) that aims at separating the common factors of variation between a background dataset (BG) (i.e., healthy subjects) and a target dataset (TG) (i.e., patients)…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Robin Louiset , Edouard Duchesnay , Antoine Grigis , Benoit Dufumier , Pietro Gori

Physical imaging is a foundational characterization method in areas from condensed matter physics and chemistry to astronomy and spans length scales from atomic to universe. Images encapsulate crucial data regarding atomic bonding,…

We present a self-supervised method to disentangle factors of variation in high-dimensional data that does not rely on prior knowledge of the underlying variation profile (e.g., no assumptions on the number or distribution of the individual…

Machine Learning · Computer Science 2022-09-23 Eric Yeats , Frank Liu , David Womble , Hai Li

In this article, we analyze the morphometry of hippocampus in subjects with very mild dementia of Alzheimer's type (DAT) and nondemented controls and how it changes over a two-year period. Morphometric differences with respect to a template…

Objective speech disorder classification for speakers with communication difficulty is desirable for diagnosis and administering therapy. With the current state of speech technology, it is evident to propose neural networks for this…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-15 Jinzi Qi , Hugo Van hamme

Learning precise representations of users and items to fit observed interaction data is the fundamental task of collaborative filtering. Existing studies usually infer entangled representations to fit such interaction data, neglecting to…

Information Retrieval · Computer Science 2024-01-11 Zhiqiang Guo , Guohui Li , Jianjun Li , Chaoyang Wang , Si Shi

An organ shape atlas, which represents the shape and position of the organs and skeleton of a living body using a small number of parameters, is expected to have a wide range of clinical applications, including intraoperative guidance and…

Image and Video Processing · Electrical Eng. & Systems 2025-06-19 Zijie Wang , Ryuichi Umehara , Mitsuhiro Nakamura , Megumi Nakao

Existing zero-shot skeleton-based action recognition methods utilize projection networks to learn a shared latent space of skeleton features and semantic embeddings. The inherent imbalance in action recognition datasets, characterized by…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Sheng-Wei Li , Zi-Xiang Wei , Wei-Jie Chen , Yi-Hsin Yu , Chih-Yuan Yang , Jane Yung-jen Hsu

Functional connectivity (FC) derived from resting-state fMRI is widely used to characterize large-scale brain network alterations in neurological and psychiatric disorders. However, FC construction critically depends on the choice of brain…

Neurons and Cognition · Quantitative Biology 2026-05-11 Minheng Chen , Chao Cao , Jing Zhang , Tianming Liu , Dajiang Zhu