中文
相关论文

相关论文: General oracle inequalities for a penalized log-li…

200 篇论文

We consider the problem of simultaneous variable selection and estimation in additive, partially linear models for longitudinal/clustered data. We propose an estimation procedure via polynomial splines to estimate the nonparametric…

统计理论 · 数学 2013-02-04 Shujie Ma , Qiongxia Song , Li Wang

For data with high-dimensional covariates but small to moderate sample sizes, the analysis of single datasets often generates unsatisfactory results. The integrative analysis of multiple independent datasets provides an effective way of…

统计方法学 · 统计学 2015-01-19 Yuan Huang , Qingzhao Zhang , Sanguo Zhang , Jian Huang , Shuangge Ma

We study the law of the iterated logarithm (LIL) for the maximum likelihood estimation of the parameters (as a convex optimization problem) in the generalized linear models with independent or weakly dependent ($\rho$-mixing, $m$-dependent)…

统计理论 · 数学 2020-04-28 Xiaowei Yang , Shuang Song , Huiming Zhang

The paper deals with conditional linear information inequalities valid for entropy functions induced by discrete random variables. Specifically, the so-called conditional Ingleton inequalities are in the center of interest: these are valid…

信息论 · 计算机科学 2022-03-16 Milan Studeny

This paper studies oracle properties of $\ell_1$-penalized least squares in nonparametric regression setting with random design. We show that the penalized least squares estimator satisfies sparsity oracle inequalities, i.e., bounds in…

统计理论 · 数学 2007-08-03 Florentina Bunea , Alexandre Tsybakov , Marten Wegkamp

Empirical likelihood is a very important nonparametric approach which is of wide application. However, it is hard and even infeasible to calculate the empirical log-likelihood ratio statistic with massive data. The main challenge is the…

统计方法学 · 统计学 2024-01-24 Qihua Wang , Jinye Du , Ying Sheng

In this paper, we establish some general forms of the law of the iterated logarithm for independent random variables in a sub-linear expectation space, where the random variables are not necessarily identically distributed. Exponential…

概率论 · 数学 2021-06-16 Li-Xin Zhang

We consider a finite mixture of Gaussian regression model for high- dimensional data, where the number of covariates may be much larger than the sample size. We propose to estimate the unknown conditional mixture density by a maximum…

统计理论 · 数学 2014-09-05 Emilie Devijver

We present a construction of the basic operators of stochastic analysis (gradient and divergence) for a class of discrete-time normal martingales called obtuse random walks. The approach is based on the chaos representation property and…

概率论 · 数学 2015-02-18 Uwe Franz , Tarek Hamdi

In this study, we consider unsupervised clustering of categorical vectors that can be of different size using mixture. We use likelihood maximization to estimate the parameters of the underlying mixture model and a penalization technique to…

统计理论 · 数学 2017-09-08 Esther Derman , Erwan Le Pennec

There is some disagreement on whether Likert scale data should be treated as ordinal or continuous. This paper treats Likert data as ordinal, uses non-parametric hypothesis testing, and clustering to validate those variables that have…

计算机与社会 · 计算机科学 2020-11-20 Matthew Norris

This article investigates unsupervised classification techniques for categorical multivariate data. The study employs multivariate multinomial mixture modeling, which is a type of model particularly applicable to multilocus genotypic data.…

统计理论 · 数学 2014-03-11 Dominique Bontemps , Wilson Toussile

The log-likelihood of a generative model often involves both positive and negative terms. For a temporal multivariate point process, the negative term sums over all the possible event types at each time and also integrates over all the…

机器学习 · 计算机科学 2020-11-03 Hongyuan Mei , Tom Wan , Jason Eisner

We consider linear structural equation models that are associated with mixed graphs. The structural equations in these models only involve observed variables, but their idiosyncratic error terms are allowed to be correlated and…

统计计算 · 统计学 2017-10-10 Y. Samuel Wang , Mathias Drton

For obtaining causal inferences that are objective, and therefore have the best chance of revealing scientific truths, carefully designed and executed randomized experiments are generally considered to be the gold standard. Observational…

应用统计 · 统计学 2008-11-12 Donald B. Rubin

Typical Bayesian methods for models with latent variables (or random effects) involve directly sampling the latent variables along with the model parameters. In high-level software code for model definitions (using, e.g., BUGS, JAGS, Stan),…

统计计算 · 统计学 2022-12-12 E. C. Merkle , D. Furr , S. Rabe-Hesketh

We study the problem of offline learning in automated decision systems under the contextual bandits model. We are given logged historical data consisting of contexts, (randomized) actions, and (nonnegative) rewards. A common goal is to…

机器学习 · 计算机科学 2019-01-16 Yifei Ma , Yu-Xiang Wang , Balakrishnan , Narayanaswamy

We extend the correspondence between two-stage coding procedures in data compression and penalized likelihood procedures in statistical estimation. Traditionally, this had required restriction to countable parameter spaces. We show how to…

统计理论 · 数学 2015-05-08 Sabyasachi Chatterjee , Andrew Barron

We introduce a class of Markov chains, that contains the model of stochastic approximation by averaging and non-averaging. Using martingale approximation method, we establish various deviation inequalities for separately Lipschitz functions…

概率论 · 数学 2022-09-16 Xiequan Fan , Pierre Alquier , Paul Doukhan

The quality of consequences in a decision making problem under (severe) uncertainty must often be compared among different targets (goals, objectives) simultaneously. In addition, the evaluations of a consequence's performance under the…

人工智能 · 计算机科学 2022-12-15 Christoph Jansen , Georg Schollmeyer , Thomas Augustin