English
Related papers

Related papers: Unimodal probability distributions for deep ordina…

200 papers

In this paper, a new natural discrete version of the one parameter polynomial exponential family of distributions have been proposed and studied. The distribution is named as Natural Discrete One Parameter Polynomial Exponential (NDOPPE)…

Statistics Theory · Mathematics 2020-07-08 Sudhansu S. Maiti , Molay Kumar Ruidas , Sumanta Adhya

Deterministic neural nets have been shown to learn effective predictors on a wide range of machine learning problems. However, as the standard approach is to train the network to minimize a prediction loss, the resultant model remains…

Machine Learning · Computer Science 2018-11-02 Murat Sensoy , Lance Kaplan , Melih Kandemir

In this work, we present a method to generate probability distributions and classes of probability distributions, which broadens a process of probability distribution construction. In this method, distribution classes are built from…

Statistics Theory · Mathematics 2021-08-16 Cícero Carlos Ramos de Brito , Leandro Chaves Rêgo , Wilson Rosa de Oliveira

Distributional reinforcement learning algorithms have attempted to utilize estimated uncertainty for exploration, such as optimism in the face of uncertainty. However, using the estimated variance for optimistic exploration may cause biased…

Machine Learning · Computer Science 2023-12-06 Taehyun Cho , Seungyub Han , Heesoo Lee , Kyungjae Lee , Jungwoo Lee

Probabilistic models with discrete latent variables naturally capture datasets composed of discrete classes. However, they are difficult to train efficiently, since backpropagation through discrete variables is generally not possible. We…

Machine Learning · Statistics 2017-04-25 Jason Tyler Rolfe

A probability distribution over the Boolean cube is monotone if flipping the value of a coordinate from zero to one can only increase the probability of an element. Given samples of an unknown monotone distribution over the Boolean cube, we…

Data Structures and Algorithms · Computer Science 2020-02-11 Ronitt Rubinfeld , Arsen Vasilyan

We propose a method for the accurate estimation of rare event or failure probabilities for expensive-to-evaluate numerical models in high dimensions. The proposed approach combines ideas from large deviation theory and adaptive importance…

Computation · Statistics 2023-03-28 Shanyin Tong , Georg Stadler

The binomial, the negative binomial, the Poisson, the compound Poisson and the Erlang distribution do all admit integral representations with respect to its (continuous) parameter. We use the Margulis-Russo type formulas for Bernoulli and…

Probability · Mathematics 2026-02-05 Guenter Last , Sergei Zuyev

Existing integer-valued autoregressive (INAR) models for count random fields suffer from difficulties in characterizing the stationary marginal distribution and in computing conditional probabilities (as required for likelihood inference).…

Methodology · Statistics 2026-05-15 Christian H. Weiß , Angelika Silbernagel

A wide variety of complex physical systems described by unitary matrices have been shown numerically to satisfy level statistics predicted by Dyson's circular ensemble. We argue that the impact of localization in such systems is to provide…

Condensed Matter · Physics 2009-10-28 K. A. Muttalib , M. E. H. Ismail

The idea behind Poisson approximation to the binomial distribution was used in [J. de la Cal, F. Luquin, J. Approx. Theory, 68(3), 1992, 322-329] and subsequent papers in order to establish the convergence of suitable sequences of positive…

Probability · Mathematics 2022-08-18 Ana-Maria Acu , Margareta Heilmann , Ioan Rasa , Andra Seserman

A general piecewise (including pointwise) probability distribution with space-saving notation and its hierarchical particular cases are considered. The explicit closed-form normalization, expectation, and variance formulas along with the…

Probability · Mathematics 2022-02-01 Lev Gelimson

The Poisson multinomial distribution (PMD) describes the distribution of the sum of $n$ independent but non-identically distributed random vectors, in which each random vector is of length $m$ with 0/1 valued elements and only one of its…

Computation · Statistics 2022-01-13 Zhengzhi Lin , Yueyao Wang , Yili Hong

Accurate estimation of predictive uncertainty in modern neural networks is critical to achieve well calibrated predictions and detect out-of-distribution (OOD) inputs. The most promising approaches have been predominantly focused on…

Machine Learning · Computer Science 2020-07-13 Shreyas Padhy , Zachary Nado , Jie Ren , Jeremiah Liu , Jasper Snoek , Balaji Lakshminarayanan

Cross-entropy loss and focal loss are the most common choices when training deep neural networks for classification problems. Generally speaking, however, a good loss function can take on much more flexible forms, and should be tailored for…

Computer Vision and Pattern Recognition · Computer Science 2022-05-12 Zhaoqi Leng , Mingxing Tan , Chenxi Liu , Ekin Dogus Cubuk , Xiaojie Shi , Shuyang Cheng , Dragomir Anguelov

In this paper, we adopt a probability distribution estimation perspective to explore the optimization mechanisms of supervised classification using deep neural networks. We demonstrate that, when employing the Fenchel-Young loss, despite…

Machine Learning · Computer Science 2025-04-01 Binchuan Qi , Wei Gong , Li Li

In this paper, we explore ordinal classification (in the context of deep neural networks) through a simple modification of the squared error loss which not only allows it to not only be sensitive to class ordering, but also allows the…

Machine Learning · Statistics 2017-01-10 Christopher Beckham , Christopher Pal

This paper introduces a convenient strategy for coding and predicting sequences of independent, identically distributed random variables generated from a large alphabet of size $m$. In particular, the size of the sample is allowed to be…

Information Theory · Computer Science 2014-01-17 Xiao Yang , Andrew R. Barron

We propose data thinning, an approach for splitting an observation into two or more independent parts that sum to the original observation, and that follow the same distribution as the original observation, up to a (known) scaling of a…

Methodology · Statistics 2023-11-22 Anna Neufeld , Ameer Dharamshi , Lucy L. Gao , Daniela Witten

Exponential random graph models are an important tool in the statistical analysis of data. However, Bayesian parameter estimation for these models is extremely challenging, since evaluation of the posterior distribution typically involves…

Computation · Statistics 2017-05-05 Lampros Bouranis , Nial Friel , Florian Maire