English
Related papers

Related papers: Sparse Nested Markov models with Log-linear Parame…

200 papers

The data torrent unleashed by current and upcoming astronomical surveys demands scalable analysis methods. Many machine learning approaches scale well, but separating the instrument measurement from the physical effects of interest, dealing…

Computation · Statistics 2023-04-19 Johannes Buchner

Enumerating the directed acyclic graphs (DAGs) of a Markov equivalence class (MEC) is an important primitive in causal analysis. The central resource from the perspective of computational complexity is the delay, that is, the time an…

Artificial Intelligence · Computer Science 2023-12-19 Marcel Wienöbst , Malte Luttermann , Max Bannach , Maciej Liśkiewicz

A penalized maximum likelihood estimation approach is proposed for discrete-time hidden Markov models where covariates affect the observed responses and serial dependence is considered. The proposed penalized maximum likelihood method…

Methodology · Statistics 2025-07-04 Luca Brusa , Fulvia Pennoni , Francesco Bartolucci , Romina Peruilh Bagolini

Nonnegative matrix factorization is a powerful technique to realize dimension reduction and pattern recognition through single-layer data representation learning. Deep learning, however, with its carefully designed hierarchical structure,…

Computer Vision and Pattern Recognition · Computer Science 2017-07-31 Zhenxing Guo , Shihua Zhang

"Mixed Data" comprising a large number of heterogeneous variables (e.g. count, binary, continuous, skewed continuous, among other data types) are prevalent in varied areas such as genomics and proteomics, imaging genetics, national…

Statistics Theory · Mathematics 2014-11-04 Eunho Yang , Pradeep Ravikumar , Genevera I. Allen , Yulia Baker , Ying-Wooi Wan , Zhandong Liu

We consider (nonparametric) sparse (generalized) additive models (SpAM) for classification. The design of a SpAM classifier is based on minimizing the logistic loss with a sparse group Lasso/Slope-type penalties on the coefficients of…

Statistics Theory · Mathematics 2024-05-16 Felix Abramovich

This paper introduces a novel class of models for binary data, which we call log-mean linear models. The characterizing feature of these models is that they are specified by linear constraints on the log-mean linear parameter, defined as a…

Methodology · Statistics 2013-01-14 Alberto Roverato , Monia Lupparelli , Luca La Rocca

Markov combination is an operation that takes two statistical models and produces a third whose marginal distributions include those of the original models. Building upon and extending existing work in the Gaussian case, we develop Markov…

Statistics Theory · Mathematics 2025-09-24 Orlando Marigliano , Eva Riccomagno

Ising models describe the joint probability distribution of a vector of binary feature variables. Typically, not all the variables interact with each other and one is interested in learning the presumably sparse network structure of the…

Machine Learning · Computer Science 2019-07-09 Frank Nussbaum , Joachim Giesen

We characterize the effectiveness of a classical algorithm for recovering the Markov graph of a general discrete pairwise graphical model from i.i.d. samples. The algorithm is (appropriately regularized) maximum conditional log-likelihood,…

Machine Learning · Computer Science 2019-06-20 Shanshan Wu , Sujay Sanghavi , Alexandros G. Dimakis

Many important datasets contain samples that are missing one or more feature values. Maintaining the interpretability of machine learning models in the presence of such missing data is challenging. Singly or multiply imputing missing values…

Machine Learning · Computer Science 2026-02-10 Hayden McTavish , Jon Donnelly , Margo Seltzer , Cynthia Rudin

This paper introduces sparse dynamic chain graph models for network inference in high dimensional non-Gaussian time series data. The proposed method parametrized by a precision matrix that encodes the intra time-slice conditional…

Methodology · Statistics 2018-05-28 Pariya Behrouzi , Fentaw Abegaz , Ernst C. Wit

While discrete latent variable models have had great success in self-supervised learning, most models assume that frames are independent. Due to the segmental nature of phonemes in speech perception, modeling dependencies among latent…

Computation and Language · Computer Science 2022-11-01 Sung-Lin Yeh , Hao Tang

In this paper, we study the problem of sparse mixed linear regression on an unlabeled dataset that is generated from linear measurements from two different regression parameter vectors. Since the data is unlabeled, our task is not only to…

Machine Learning · Computer Science 2022-09-12 Adarsh Barik , Jean Honorio

Graphical models encode conditional independence statements of a multivariate distribution via a graph. Traditionally, the marginal distributions in a graphical model are assumed to be Gaussian. In this paper, we propose a three-level…

Methodology · Statistics 2025-05-01 Luis E. Nieto-Barajas , Simón Lunagómez

A statistical language model assigns probability to strings of arbitrary length. Unfortunately, it is not possible to gather reliable statistics on strings of arbitrary length from a finite corpus. Therefore, a statistical language model…

cmp-lg · Computer Science 2008-02-03 Eric Sven Ristad , Robert G. Thomas

We introduce a new embarrassingly parallel parameter learning algorithm for Markov random fields with untied parameters which is efficient for a large class of practical models. Our algorithm parallelizes naturally over cliques and, for…

Machine Learning · Statistics 2014-02-06 Yariv Dror Mizrahi , Misha Denil , Nando de Freitas

Denoising diffusion probabilistic models have recently received much research attention since they outperform alternative approaches, such as GANs, and currently provide state-of-the-art generative performance. The superior performance of…

Computer Vision and Pattern Recognition · Computer Science 2022-03-17 Dmitry Baranchuk , Ivan Rubachev , Andrey Voynov , Valentin Khrulkov , Artem Babenko

Directed acyclic graphs (DAGs) are a popular framework to express multivariate probability distributions. Acyclic directed mixed graphs (ADMGs) are generalizations of DAGs that can succinctly capture much richer sets of conditional…

Machine Learning · Statistics 2010-09-01 Ricardo Silva , Charles Blundell , Yee Whye Teh

Motivated by problems from neuroimaging in which existing approaches make use of "mass univariate" analysis which neglects spatial structure entirely, but the full joint modelling of all quantities of interest is computationally infeasible,…

Methodology · Statistics 2022-04-19 Denishrouf Thesingarajah , Adam M. Johansen