中文
相关论文

相关论文: The sign of the logistic regression coefficient

200 篇论文

We study the residual bootstrap (RB) method in the context of high-dimensional linear regression. Specifically, we analyze the distributional approximation of linear contrasts $c^{\top} (\hat{\beta}_{\rho}-\beta)$, where…

统计理论 · 数学 2016-07-05 Miles E. Lopes

We study linear regressions in a context where the outcome of interest and some of the covariates are observed in two different datasets that cannot be matched. Traditional approaches obtain point identification by relying, often…

计量经济学 · 经济学 2025-11-18 Xavier D'Haultfoeuille , Christophe Gaillac , Arnaud Maurel

This short note is to point the reader to notice that the proof of high dimensional asymptotic normality of MLE estimator for logistic regression under the regime $p_n=o(n)$ given in paper: "Maximum likelihood estimation in logistic…

统计理论 · 数学 2018-01-29 Huiming Zhang

We consider linear random coefficient regression models, where the regressors are allowed to have a finite support. First, we investigate identifiability, and show that the means and the variances and covariances of the random coefficients…

统计理论 · 数学 2023-06-16 Philipp Hermann , Hajo Holzmann

In the critical beta-splitting model of a random $n$-leaf binary tree, leaf-sets are recursively split into subsets, and a set of $m$ leaves is split into subsets containing $i$ and $m-i$ leaves with probabilities proportional to…

概率论 · 数学 2024-09-09 David Aldous , Boris Pittel

The Statistical Learning Theory (SLT) provides the theoretical guarantees for supervised machine learning based on the Empirical Risk Minimization Principle (ERMP). Such principle defines an upper bound to ensure the uniform convergence of…

In Change point detection task Likelihood Ratio Test (LRT) is sequentially applied in a sliding window procedure. Its high values indicate changes of parametric distribution in the data sequence. Correspondingly LRT values require…

统计理论 · 数学 2017-10-23 Nazar Buzun , Valeriy Avanesov

It is well known that if a random vector satisfies a log-Sobolev inequality, all of its marginals have subgaussian tails. In the spirit of the KLS conjecture, we investigate whether this implication can be reversed under a log-concavity…

泛函分析 · 数学 2026-02-17 Pierre Bizeul

We derive upper bounds for random design linear regression with dependent ($\beta$-mixing) data absent any realizability assumptions. In contrast to the strictly realizable martingale noise regime, no sharp instance-optimal non-asymptotics…

机器学习 · 计算机科学 2023-10-30 Ingvar Ziemann , Stephen Tu , George J. Pappas , Nikolai Matni

The method of 1-bit ("sign-sign") random projections has been a popular tool for efficient search and machine learning on large datasets. Given two $D$-dim data vectors $u$, $v\in\mathbb{R}^D$, one can generate $x = \sum_{i=1}^D u_i r_i$,…

统计方法学 · 统计学 2018-05-03 Ping Li

We consider a problem of ecological inference, in which individual-level covariates are known, but labeled data is available only at the aggregate level. The intended application is modeling voter preferences in elections. In Rosenman and…

机器学习 · 统计学 2019-07-23 Evan Rosenman

We consider a model for logistic regression where only a subset of features of size $p$ is used for training a linear classifier over $n$ training samples. The classifier is obtained by running gradient descent (GD) on logistic loss. For…

机器学习 · 统计学 2020-05-12 Zeyu Deng , Abla Kammoun , Christos Thrampoulidis

We consider a linear model where the coefficients - intercept and slopes - are random with a law in a nonparametric class and independent from the regressors. Identification often requires the regressors to have a support which is the whole…

统计理论 · 数学 2020-06-22 Christophe Gaillac , Eric Gautier

The metric properties of the set in which random variables take their values lead to relevant probabilistic concepts. For example, the mean of a random variable is a best predictor in that it minimizes the standard Euclidean distance or…

概率论 · 数学 2018-09-21 Henryk Gzyl

The goal of regression analysis is to predict the value of a numeric outcome variable y given a vector of joint values of other (predictor) variables x. Usually a particular x-vector does not specify a repeatable value for y, but rather a…

机器学习 · 统计学 2020-01-29 Jerome H. Friedman

In this paper we discuss how to evaluate the differences between fitted logistic regression models across sub-populations. Our motivating example is in studying computerized diagnosis for learning disabilities, where sub-populations based…

统计方法学 · 统计学 2023-03-24 Guy Ashiri-Prossner , Yuval Benjamini

Nonparametric regression with random design is considered. Estimates are defined by minimzing a penalized empirical $L_2$ risk over a suitably chosen class of neural networks with one hidden layer via gradient descent. Here, the gradient…

统计理论 · 数学 2019-12-10 Alina Braun , Michael Kohler , Harro Walk

An important problem in the field of bioinformatics is to identify interactive effects among profiled variables for outcome prediction. In this paper, a logistic regression model with pairwise interactions among a set of binary covariates…

人工智能 · 计算机科学 2016-12-30 Easton Li Xu , Xiaoning Qian , Tie Liu , Shuguang Cui

The beta regression model is a useful framework to model response variables that are rates or proportions, that is to say, response variables which are continuous and restricted to the interval (0,1). As with any other regression model,…

统计方法学 · 统计学 2024-06-27 Luis Firinguetti , Manuel González-Navarrete , Romer Machaca-Aguilar

Predictions are often probabilities; e.g., a prediction could be for precipitation tomorrow, but with only a 30% chance. Given such probabilistic predictions together with the actual outcomes, "reliability diagrams" help detect and diagnose…

统计理论 · 数学 2022-11-15 Imanol Arrieta-Ibarra , Paman Gujral , Jonathan Tannen , Mark Tygert , Cherie Xu