English
Related papers

Related papers: Differential gene expression analysis via two-comp…

200 papers

The estimation of covariance matrices of gene expressions has many applications in cancer systems biology. Many gene expression studies, however, are hampered by low sample size and it has therefore become popular to increase sample size by…

Sine-skewed circular distributions are identifiable and have easily-computable trigonometric moments and a simple random number generation algorithm, whereas they are known to have relatively low levels of asymmetry. This study proposes a…

Methodology · Statistics 2024-02-16 Yoichi Miyata , Takayuki Shiohama , Toshihiro Abe

Estimating parameters of mixture model has wide applications ranging from classification problems to estimating of complex distributions. Most of the current literature on estimating the parameters of the mixture densities are based on…

Machine Learning · Statistics 2020-06-23 Yuantong Li , Qi Ma , Sujit K. Ghosh

The linking genotype to phenotype is the fundamental aim of modern genetics. We focus on study of links between gene expression data and phenotype data through integrative analysis. We propose three approaches. 1) The inherent complexity of…

Quantitative Methods · Quantitative Biology 2015-06-30 Min Xu

This paper provides a generic framework of component analysis (CA) methods introducing a new expression for scatter matrices and Gram matrices, called Generalized Pairwise Expression (GPE). This expression is quite compact but highly…

Computer Vision and Pattern Recognition · Computer Science 2012-10-09 Akisato Kimura , Masashi Sugiyama , Sakano Hitoshi , Hirokazu Kameoka

The goal of anomaly detection is to identify anomalous samples from normal ones. In this paper, a small number of anomalies are assumed to be available at the training stage, but they are assumed to be collected only from several anomaly…

Machine Learning · Computer Science 2022-05-03 Bowen Tian , Qinliang Su , Jian Yin

In unsupervised classification, Hidden Markov Models (HMM) are used to account for a neighborhood structure between observations. The emission distributions are often supposed to belong to some parametric family. In this paper, a…

Machine Learning · Statistics 2012-06-25 Stevenn Volant , Caroline Bérard , Marie-Laure Martin-Magniette , Stéphane Robin

Because of the high cost of commercial genotyping chip technologies, many investigations have used a two-stage design for genome-wide association studies, using part of the sample for an initial discovery of ``promising'' SNPs at a less…

Mixtures of factor analyzers (MFA) provide a powerful tool for modelling high-dimensional datasets. In recent years, several generalizations of MFA have been developed where the normality assumption of the factors and/or of the errors was…

Methodology · Statistics 2018-10-29 Sharon X. Lee , Tsung-I Lin , Geoffrey J. McLachlan

Variational methods are widely used for approximate posterior inference. However, their use is typically limited to families of distributions that enjoy particular conjugacy properties. To circumvent this limitation, we propose a family of…

Machine Learning · Computer Science 2012-06-22 Samuel Gershman , Matt Hoffman , David Blei

The skew-normal and related families are flexible and asymmetric parametric models suitable for modelling a diverse range of systems. We show that the multivariate maximum of a high-dimensional extended skew-normal random sample has…

Methodology · Statistics 2018-10-02 Boris Beranger , Simone A. Padoan , Yangfan Xu , Scott A. Sisson

Recently, Generative Adversarial Networks (GANs) have emerged as a popular alternative for modeling complex high dimensional distributions. Most of the existing works implicitly assume that the clean samples from the target distribution are…

Machine Learning · Statistics 2019-02-14 Mohammadreza Soltani , Swayambhoo Jain , Abhinav Sambasivan

There is an extensive literature on methods for meta-analysis of diagnostic studies, but it mainly focuses on a single test. However, the better understanding of a particular disease has led to the development of multiple tests. A…

Methodology · Statistics 2020-10-19 Aristidis K. Nikoloulopoulos

We study semiparametric factor models in high-dimensional panels where the factor loadings consist of a nonparametric component explained by observed covariates and an idiosyncratic component capturing unobserved heterogeneity. A key…

Methodology · Statistics 2025-12-09 Sijie Zheng

Finite mixtures of matrix normal distributions are a powerful tool for classifying three-way data in unsupervised problems. The distribution of each component is assumed to be a matrix variate normal density. The mixture model can be…

Methodology · Statistics 2013-03-07 Cinzia Viroli

In this study, we propose a robust mixture regression procedure based on the skew t distribution to model heavy-tailed and/or skewed errors in a mixture regression setting. Using the scale mixture representation of the skew t distribution,…

Statistics Theory · Mathematics 2017-06-12 Fatma Zehra Doğru , Olcay Arslan

Gene expression analysis aims at identifying the genes able to accurately predict biological parameters like, for example, disease subtyping or progression. While accurate prediction can be achieved by means of many different techniques,…

Methodology · Statistics 2008-09-11 Christine De Mol , Sofia Mosci , Magali Traskine , Alessandro Verri

Signal-processing molecules inside cells are often present at low copy number, which necessitates probabilistic models to account for intrinsic noise. Probability distributions have traditionally been found using simulation-based approaches…

Molecular Networks · Quantitative Biology 2009-11-09 Andrew Mugler , Aleksandra M. Walczak , Chris H. Wiggins

Biological structure and function depend on complex regulatory interactions between many genes. A wealth of gene expression data is available from high-throughput genome-wide measurement technologies, but effective gene regulatory network…

Molecular Networks · Quantitative Biology 2016-03-28 Arwen Vanice Bradley , Ye Henry Li , Bokyung Choi , Wing Hung Wong

Mixtures of skew-t distributions offer a flexible choice for model-based clustering. A mixture model of this sort can be implemented using a variety of formulations of the skew-t distribution. Herein we develop a mixture of skew-t factor…

Methodology · Statistics 2017-10-09 Paula M. Murray , Ryan P. Browne , Paul D. McNicholas