English
Related papers

Related papers: Learning Energy-Based Models by Self-normalising t…

200 papers

Energy-based modeling is a promising approach to unsupervised learning, which yields many downstream applications from a single model. The main difficulty in learning energy-based models with the "contrastive approaches" is the generation…

Machine Learning · Computer Science 2021-11-30 Kirill Neklyudov , Priyank Jaini , Max Welling

Energy-based models (EBMs) are generative models that are usually trained via maximum likelihood estimation. This approach becomes challenging in generic situations where the trained energy is non-convex, due to the need to sample the Gibbs…

Machine Learning · Computer Science 2022-02-16 Carles Domingo-Enrich , Alberto Bietti , Marylou Gabrié , Joan Bruna , Eric Vanden-Eijnden

Consider a setting with $N$ independent individuals, each with an unknown parameter, $p_i \in [0, 1]$ drawn from some unknown distribution $P^\star$. After observing the outcomes of $t$ independent Bernoulli trials, i.e., $X_i \sim…

Statistics Theory · Mathematics 2019-02-13 Ramya Korlakai Vinayak , Weihao Kong , Gregory Valiant , Sham M. Kakade

Plasticity Loss is an increasingly important phenomenon that refers to the empirical observation that as a neural network is continually trained on a sequence of changing tasks, its ability to adapt to a new task diminishes over time. We…

Machine Learning · Computer Science 2025-09-30 Vivek F. Farias , Adam D. Jozefiak

Despite their advantages, normalizing flows generally suffer from several shortcomings including their tendency to generate unrealistic data (e.g., images) and their failing to detect out-of-distribution data. One reason for these…

Machine Learning · Statistics 2022-07-13 Florentin Coeurdoux , Nicolas Dobigeon , Pierre Chainais

In pseudo-labeling (PL), which is a type of semi-supervised learning, pseudo-labels are assigned based on the confidence scores provided by the classifier; therefore, accurate confidence is important for successful PL. In this study, we…

Computer Vision and Pattern Recognition · Computer Science 2024-04-16 Masahito Toba , Seiichi Uchida , Hideaki Hayashi

Stochastic Differential Equations (SDEs) are used as statistical models in many disciplines. However, intractable likelihood functions for SDEs make inference challenging, and we need to resort to simulation-based techniques to estimate and…

Methodology · Statistics 2014-08-12 Grant Schneider , Peter F. Craigmile , Radu Herbei

Auto-regressive sequence generative models trained by Maximum Likelihood Estimation suffer the exposure bias problem in practical finite sample scenarios. The crux is that the number of training samples for Maximum Likelihood Estimation is…

Machine Learning · Statistics 2020-07-14 Yuxuan Song , Ning Miao , Hao Zhou , Lantao Yu , Mingxuan Wang , Lei Li

In many statistical learning problems, the target functions to be optimized are highly non-convex in various model spaces and thus are difficult to analyze. In this paper, we compute \emph{Energy Landscape Maps} (ELMs) which characterize…

Machine Learning · Statistics 2014-10-03 Maria Pavlovskaia , Kewei Tu , Song-Chun Zhu

We explore past and recent developments in rare-event probability estimation with a particular focus on a novel Monte Carlo technique Empirical Likelihood Maximization (ELM). This is a versatile method that involves sampling from a sequence…

Computation · Statistics 2013-12-12 A. Huang , Z. I. Botev

The nearest neighbor spacing distribution (NNSD) is one of common methods in statistical analysis of nuclear energy levels. In this paper, we have proposed Maximum Likelihood Estimation (MLE) method to evaluate parameter of (NNSD)'s which…

Nuclear Theory · Physics 2015-05-20 M. A. Jafarizadeh , N. Fouladi , H. Sabri , B. Rashidian Maleki

Robust autonomous driving requires agents to accurately identify unexpected areas (anomalies) in urban scenes. To this end, some critical issues remain open: how to design advisable metric to measure anomalies, and how to properly generate…

Computer Vision and Pattern Recognition · Computer Science 2025-01-03 Yuanpeng Tu , Yuxi Li , Boshen Zhang , Liang Liu , Jiangning Zhang , Yabiao Wang , Cai Rong Zhao

Biomolecular Neural Networks (BNNs), artificial neural networks with biologically synthesizable architectures, achieve universal function approximation capabilities beyond simple biological circuits. However, training BNNs remains…

Machine Learning · Computer Science 2025-09-09 Eric Palanques-Tost , Hanna Krasowski , Murat Arcak , Ron Weiss , Calin Belta

We introduce two new particle-based algorithms for learning latent variable models via marginal maximum likelihood estimation, including one which is entirely tuning-free. Our methods are based on the perspective of marginal maximum…

Machine Learning · Statistics 2024-03-04 Louis Sharrock , Daniel Dodd , Christopher Nemeth

Schr\"odinger Bridge (SB) is an entropy-regularized optimal transport problem that has received increasing attention in deep generative modeling for its mathematical flexibility compared to the Scored-based Generative Model (SGM). However,…

Machine Learning · Statistics 2023-04-04 Tianrong Chen , Guan-Horng Liu , Evangelos A. Theodorou

We build on auto-encoding sequential Monte Carlo (AESMC): a method for model and proposal learning based on maximizing the lower bound to the log marginal likelihood in a broad family of structured probabilistic models. Our approach relies…

Machine Learning · Statistics 2018-04-06 Tuan Anh Le , Maximilian Igl , Tom Rainforth , Tom Jin , Frank Wood

Maximum likelihood (ML) estimation using Newton's method in nonlinear state space models (SSMs) is a challenging problem due to the analytical intractability of the log-likelihood and its gradient and Hessian. We estimate the gradient and…

Computation · Statistics 2016-03-11 Manon Kok , Johan Dahlin , Thomas B. Schön , Adrian Wills

We study three fundamental statistical-learning problems: distribution estimation, property estimation, and property testing. We establish the profile maximum likelihood (PML) estimator as the first unified sample-optimal approach to a wide…

Machine Learning · Statistics 2019-07-12 Yi Hao , Alon Orlitsky

Supervised fine-tuning (SFT) is the standard approach for post-training large language models (LLMs), yet it often shows limited generalization. We trace this limitation to its default training objective: negative log likelihood (NLL).…

Computation and Language · Computer Science 2026-05-25 Gaotang Li , Ruizhong Qiu , Xiusi Chen , Heng Ji , Hanghang Tong

Bayesian Neural Networks (BNNs) offer a principled and natural framework for proper uncertainty quantification in the context of deep learning. They address the typical challenges associated with conventional deep learning methods, such as…

Computation · Statistics 2024-11-13 Zahra Moslemi , Yang Meng , Shiwei Lan , Babak Shahbaba
‹ Prev 1 3 4 5 6 7 10 Next ›