English
Related papers

Related papers: Better estimates from binned income data: Interpol…

200 papers

Continuous normalizing flows (CNFs) can model data distributions with expressive infinite-length architectures. But this modeling involves computationally expensive process of solving an ordinary differential equation (ODE) during maximum…

Machine Learning · Computer Science 2024-10-15 Denis Gudovskiy , Tomoyuki Okuno , Yohei Nakata

In astrophysics a common goal is to infer the flux distribution of populations of scientifically interesting objects such as pulsars or supernovae. In practice, inference for the flux distribution is often conducted using the cumulative…

Applications · Statistics 2015-03-03 Raymond K. W. Wong , Paul Baines , Alexander Aue , Thomas C. M. Lee , Vinay L. Kashyap

This article proposes a new method to estimate an existing mutual information based dependence measure using histogram density estimates. Finding a suitable bin length for histogram is an open problem. We propose a new way of computing the…

Information Theory · Computer Science 2015-09-15 Namita Jain , C. A. Murthy

We study three notions of uncertainty quantification -- calibration, confidence intervals and prediction sets -- for binary classification in the distribution-free setting, that is without making any distributional assumptions on the data.…

Machine Learning · Statistics 2022-02-17 Chirag Gupta , Aleksandr Podkopaev , Aaditya Ramdas

Irregular errors such as heteroscedasticity and nonnormality remain major challenges in linear modeling. These issues often lead to biased inference and unreliable measures of uncertainty. Classical remedies, such as robust standard errors…

Methodology · Statistics 2026-03-05 Elsayed Elamir

Histopolation, or interpolation on segments, is a mathematical technique used to approximate a function $f$ over a given interval $I=[a,b]$ by exploiting integral information over a set of subintervals of $I$. Unlike classical polynomial…

Numerical Analysis · Mathematics 2025-08-12 Francesco Dell'Accio , Francesco Larosa , Federico Nudo , Najoua Siar

Post-selection inference (PoSI) is a statistical technique for obtaining valid confidence intervals and p-values when hypothesis generation and testing use the same source of data. PoSI can be used on a range of popular algorithms including…

Methodology · Statistics 2023-05-23 Erik Drysdale

Accurate conditional prediction in the regression setting plays an important role in many real-world problems. Typically, a point prediction often falls short since no attempt is made to quantify the prediction accuracy. Classically, under…

Methodology · Statistics 2025-09-04 Kejin Wu , Dimitris N. Politis

Motivated by non-linear, non-Gaussian, distributed multi-sensor/agent navigation and tracking applications, we propose a multi-rate consensus/fusion based framework for distributed implementation of the particle filter (CF/DPF). The CF/DPF…

Distributed, Parallel, and Cluster Computing · Computer Science 2012-09-06 Arash Mohammadi , Amir Asif

We propose lookahead diffusion probabilistic models (LA-DPMs) to exploit the correlation in the outputs of the deep neural networks (DNNs) over subsequent timesteps in diffusion probabilistic models (DPMs) to refine the mean estimation of…

Artificial Intelligence · Computer Science 2023-04-25 Guoqiang Zhang , Niwa Kenta , W. Bastiaan Kleijn

Effective decision making requires understanding the uncertainty inherent in a prediction. In regression, this uncertainty can be estimated by a variety of methods; however, many of these methods are laborious to tune, generate…

Machine Learning · Statistics 2021-12-02 Tianhui Zhou , Yitong Li , Yuan Wu , David Carlson

As predictive algorithms grow in popularity, using the same dataset to both train and test a new model has become routine across research, policy, and industry. Sample-splitting attains valid inference on model properties by using separate…

Econometrics · Economics 2025-11-27 Bruno Fava

We study the sample complexity of learning a uniform approximation of an $n$-dimensional cumulative distribution function (CDF) within an error $\epsilon > 0$, when observations are restricted to a minimal one-bit feedback. This serves as a…

Machine Learning · Computer Science 2026-05-12 Matteo Castiglioni , Anna Lunghi , Alberto Marchesi

Score matching is an approach to learning probability distributions parametrized up to a constant of proportionality (e.g. Energy-Based Models). The idea is to fit the score of the distribution, rather than the likelihood, thus avoiding the…

Machine Learning · Computer Science 2024-01-31 Yilong Qin , Andrej Risteski

We give a new explicitly invertible approximation of the normal cumulative distribution function: $\Phi(x) \simeq 1/2 + 1/2 \sqrt{1-{e}^{-x^2\frac{17+{x}^{2}}{26.694+2x^2}}}$, $\forall x \ge 0$, with absolute error $<4.00\cdot 10^{-5}$,…

Statistics Theory · Mathematics 2012-11-28 Alessandro Soranzo , Emanuela Epure

We introduce a two-parameter family of discrepancy measures, termed \emph{$(G,f)$-divergences}, obtained by applying a non-decreasing function $G$ to an $f$-divergence $D_f$. Building on Csisz\'ar's formulation of mutual $f$-information, we…

Information Theory · Computer Science 2026-01-23 Hamidreza Abin , Mahdi Zinati , Amin Gohari , Mohammad Hossein Yassaee , Mohammad Mahdi Mojahedian

To explain individual differences in development, behavior, and cognition, most previous studies focused on projecting resting-state functional MRI (fMRI) based functional connectivity (FC) data into a low-dimensional space via linear…

Neurons and Cognition · Quantitative Biology 2018-11-01 Li Xiao , Julia M. Stephen , Tony W. Wilson , Vince D. Calhoun , Yu-Ping Wang

Polynomial distribution can be applied to dynamical systems in certain situations. Macroeconomic systems characterized by economic variables such as income and wealth can be modelled similarly using polynomials. We extend our previous work…

General Finance · Quantitative Finance 2016-03-29 Elvis Oltean

Key to effective generic, or "black-box", variational inference is the selection of an approximation to the target density that balances accuracy and speed. Copula models are promising options, but calibration of the approximation can be…

Methodology · Statistics 2022-07-01 Michael Stanley Smith , Rubén Loaiza-Maya

Integrated IPD-AD analysis, which combines individual participant data (IPD) with aggregate data (AD), is increasingly recognized as an effective strategy for generating more reliable and generalizable inferences from heterogeneous studies.…

Methodology · Statistics 2026-03-03 Ming-Yueh Huang , Jing Qin , Chiung-Yu Huang