English
Related papers

Related papers: Stochastic Ordering under Weaker Likelihood-Ratio …

200 papers

Diversification is usually viewed as a reliable way to reduce risk, yet it can dramatically fail for heavy-tailed losses with infinite mean: pooling independent losses of this type may increase tail risk at every threshold. We study this…

Risk Management · Quantitative Finance 2026-03-11 Léonard Vincent

In typical applications of Bayesian optimization, minimal assumptions are made about the objective function being optimized. This is true even when researchers have prior information about the shape of the function with respect to one or…

Machine Learning · Statistics 2016-12-30 Michael Jauch , Víctor Peña

A convergence theorem for the continuous weak approximation of the solution of stochastic differential equations by general one step methods is proved, which is an extension of a theorem due to Milstein. As an application, uniform second…

Numerical Analysis · Mathematics 2013-03-19 Kristian Debrabant , Andreas Rößler

Many common loss functions such as mean-squared-error, cross-entropy, and reconstruction loss are unnecessarily rigid. Under a probabilistic interpretation, these common losses correspond to distributions with fixed shapes and scales. We…

Machine Learning · Computer Science 2020-10-05 Mark Hamilton , Evan Shelhamer , William T. Freeman

``Localization'' has proven to be a valuable tool in the Statistical Learning literature as it allows sharp risk bounds in terms of the problem geometry. Localized bounds seem to be much less exploited in the Stochastic Optimization…

Optimization and Control · Mathematics 2023-03-30 Roberto I. Oliveira , Philip Thompson

In the present article, we show the existence of a coupled fixed point for an order preserving mapping in a preordered left K-complete quasi-pseudometric space using a preorder induced by an appropriate function. We also define the concept…

General Mathematics · Mathematics 2014-11-14 Yaé Ulrich Gaba

Inverse classification, the process of making meaningful perturbations to a test point such that it is more likely to have a desired classification, has previously been addressed using data from a single static point in time. Such an…

Machine Learning · Computer Science 2016-11-15 Michael T. Lash , W. Nick Street

We generalize stochastic subgradient descent methods to situations in which we do not receive independent samples from the distribution over which we optimize, but instead receive samples that are coupled over time. We show that as long as…

Optimization and Control · Mathematics 2012-08-02 John C. Duchi , Alekh Agarwal , Mikael Johansson , Michael I. Jordan

We provide a novel computer-assisted technique for systematically analyzing first-order methods for optimization. In contrast with previous works, the approach is particularly suited for handling sublinear convergence rates and stochastic…

Optimization and Control · Mathematics 2021-12-22 Adrien Taylor , Francis Bach

Two methods are proposed for high-dimensional shape-constrained regression and classification. These methods reshape pre-trained prediction rules to satisfy shape constraints like monotonicity and convexity. The first method can be applied…

Machine Learning · Statistics 2018-05-17 Matt Bonakdarpour , Sabyasachi Chatterjee , Rina Foygel Barber , John Lafferty

We consider non-convex stochastic optimization using first-order algorithms for which the gradient estimates may have heavy tails. We show that a combination of gradient clipping, momentum, and normalized gradient descent yields convergence…

Machine Learning · Computer Science 2021-11-10 Ashok Cutkosky , Harsh Mehta

We provide elementary proofs of several results concerning the possible outcomes arising from a fixed profile within the class of positional voting systems. Our arguments enable a simple and explicit construction of paradoxical profiles,…

Combinatorics · Mathematics 2020-08-17 Jacqueline Anderson , Brian Camara , John Pike

Adaptive optimization methods (such as Adam) play a major role in LLM pretraining, significantly outperforming Gradient Descent (GD). Recent studies have proposed new smoothness assumptions on the loss function to explain the advantages of…

Machine Learning · Computer Science 2025-12-02 Robin Yadav , Shuo Xie , Tianhao Wang , Zhiyuan Li

Systems may depend on parameters which one may control, or which serve to optimise the system, or are imposed externally, or they could be uncertain. This last case is taken as the ``Leitmotiv'' for the following. A reduced order model is…

Machine Learning · Computer Science 2025-02-17 Hermann G. Matthies

We extend the empirical likelihood of Owen [Ann. Statist. 18 (1990) 90-120] by partitioning its domain into the collection of its contours and mapping the contours through a continuous sequence of similarity transformations onto the full…

Statistics Theory · Mathematics 2013-11-11 Min Tsao , Fan Wu

Let $X_1, X_2,\ldots, X_n$ (resp. $Y_1, Y_2,\ldots, Y_n$) be independent random variables such that $X_i$ (resp. $Y_i$) follows generalized exponential distribution with shape parameter $\theta_i$ and scale parameter $\lambda_i$ (resp.…

Applications · Statistics 2016-01-18 Amarjit Kundu , Shovan Chowdhury , Asok K. Nanda , Nil Kamal Hazra

Stochastic ordering among distributions has been considered in a variety of scenarios. Economic studies often involve research about the ordering of investment strategies or social welfare. However, as noted in the literature, stochastic…

Lower semi-continuity (\texttt{LSC}) is a critical assumption in many foundational optimisation theory results; however, in many cases, \texttt{LSC} is stronger than necessary. This has led to the introduction of numerous weaker continuity…

Optimization and Control · Mathematics 2025-04-11 Jacob Westerhout , Xin Guo , Hien Duy Nguyen

In finite mixtures of location-scale distributions, if there is no constraint or penalty on the parameters, then the maximum likelihood estimator does not exist because the likelihood is unbounded. To avoid this problem, we consider a…

Statistics Theory · Mathematics 2011-03-04 Kentaro Tanaka

The Whittle likelihood is a widely used and computationally efficient pseudo-likelihood. However, it is known to produce biased parameter estimates for large classes of models. We propose a method for de-biasing Whittle estimates for…