English
Related papers

Related papers: Two Sample Test for Eigendecompositions of Functio…

200 papers

Hypothesis testing is a statistical inference approach used to determine whether data supports a specific hypothesis. An important type is the two-sample test, which evaluates whether two sets of data points are from identical…

Machine Learning · Computer Science 2025-01-08 Weizhi Li , Visar Berisha , Gautam Dasarathy

Functional differentiation in the brain emerges as distinct regions specialize and is key to understanding brain function as a complex system. Previous research has modeled this process using artificial neural networks with specific…

Neurons and Cognition · Quantitative Biology 2025-11-17 Yuki Tomoda , Ichiro Tsuda , Yutaka Yamaguti

Functional data analysis, which handles data arising from curves, surfaces, volumes, manifolds and beyond in a variety of scientific fields, is a rapidly developing area in modern statistics and data science in the recent decades. The…

Methodology · Statistics 2020-08-21 Xiaoke Zhang , Wu Xue , Qiyue Wang

Many fMRI analyses examine functional connectivity, or statistical dependencies among remote brain regions. Yet popular methods for studying whole-brain functional connectivity often yield results that are difficult to interpret. Factor…

Methodology · Statistics 2024-09-24 Kyle Stanley , Nicole Lazar , Matthew Reimherr

Tests of independence are an important tool in applications, specifically in connection with the detection of a relationship between variables; they also have initiated many developments in statistical theory. In the present paper we build…

Statistics Theory · Mathematics 2026-05-13 L. Baringhaus , R. Grübel

The problem of testing equality of the entire second order structure of two independent functional linear processes is considered. A fully functional $L^2$-type test is developed which evaluates, over all frequencies, the Hilbert-Schmidt…

Methodology · Statistics 2020-04-15 Anne Leucht , Efstathios Paparoditis , Theofanis Sapatinas

Neural Network-based active learning (NAL) is a cost-effective data selection technique that utilizes neural networks to select and train on a small subset of samples. While existing work successfully develops various effective or…

Machine Learning · Computer Science 2024-06-07 Dake Bu , Wei Huang , Taiji Suzuki , Ji Cheng , Qingfu Zhang , Zhiqiang Xu , Hau-San Wong

In modern data analysis, nonparametric measures of discrepancies between random variables are particularly important. The subject is well-studied in the frequentist literature, while the development in the Bayesian setting is limited where…

Methodology · Statistics 2022-01-25 Qinyi Zhang , Veit Wild , Sarah Filippi , Seth Flaxman , Dino Sejdinovic

Background. Large Language Models (LLMs) hold promise for improving genetic variant literature review in clinical testing. We assessed Generative Pretrained Transformer 4's (GPT-4) performance, nondeterminism, and drift to inform its…

We consider two-sample tests for high-dimensional data under two disjoint models: the strongly spiked eigenvalue (SSE) model and the non-SSE (NSSE) model. We provide a general test statistic as a function of a positive-semidefinite matrix.…

Statistics Theory · Mathematics 2016-11-28 Makoto Aoshima , Kazuyoshi Yata

In this article we present a multivariate model for determining the different syntactic, semantic, and form (surface-structure) processes underlying the comprehension of simple phrases. This model is applied to EEG signals recorded during a…

Computation and Language · Computer Science 2018-04-17 Sabine Ploux , Viviane Déprez

A key property of neural networks is their capacity of adapting to data during training. Yet, our current mathematical understanding of feature learning and its relationship to generalization remain limited. In this work, we provide a…

Machine Learning · Statistics 2024-10-25 Yatin Dandi , Luca Pesce , Hugo Cui , Florent Krzakala , Yue M. Lu , Bruno Loureiro

Multi-task learning and self-training are two common ways to improve a machine learning model's performance in settings with limited training data. Drawing heavily on ideas from those two approaches, we suggest transductive auxiliary task…

Computation and Language · Computer Science 2019-09-24 Johannes Bjerva , Katharina Kann , Isabelle Augenstein

Global brain activity self-organizes into discrete patterns characterized by distinct behavioral observables and modes of information processing. The human thalamocortical system is a densely connected network where local neural activation…

We consider the problem of two-sample testing in a semi-supervised setting with abundant unlabeled covariate data. Standard two-sample tests neglect covariate information, which has the potential to significantly boost performance. However,…

Machine Learning · Statistics 2026-05-05 Gyumin Lee , Shubhanshu Shekhar , Ilmun Kim

Modelling the dynamics of interactions in a neuronal ensemble is an important problem in functional connectivity research. One popular framework is latent factor models (LFMs), which have achieved notable success in decoding neuronal…

Methodology · Statistics 2023-05-18 Meixi Chen , Martin Lysy , David Moorman , Reza Ramezan

Hypothesis testing for the slope function in functional linear regression is of both practical and theoretical interest. We develop a novel test for the nullity of the slope function, where testing the slope function is transformed into…

Methodology · Statistics 2024-04-02 Yinan Lin , Zhenhua Lin

In this paper, we have proposed a brain signal classification method, which uses eigenvalues of the covariance matrix as features to classify images (topomaps) created from the brain signals. The signals are recorded during the answering of…

Machine Learning · Computer Science 2019-05-01 Saeed Bamatraf , Muhammad Hussain , Emad-ul-Haq Qazi , Hatim Aboalsamh

We study the influence of different activation functions in the output layer of deep neural network models for soft and hard label prediction in the learning with disagreement task. In this task, the goal is to quantify the amount of…

Computation and Language · Computer Science 2024-01-05 Peyman Hosseini , Mehran Hosseini , Sana Sabah Al-Azzawi , Marcus Liwicki , Ignacio Castro , Matthew Purver

We present a general framework for hypothesis testing on distributions of sets of individual examples. Sets may represent many common data sources such as groups of observations in time series, collections of words in text or a batch of…

Methodology · Statistics 2021-02-03 Alexis Bellot , Mihaela van der Schaar