English
Related papers

Related papers: Enhancing Model Fit Evaluation in SEM: Practical T…

200 papers

A growing body of work has been querying LLMs with political questions to evaluate their potential biases. However, this probing method has limited stability, making comparisons between models unreliable. In this paper, we argue that LLMs…

Computation and Language · Computer Science 2025-06-30 Patrick Haller , Jannis Vamvas , Rico Sennrich , Lena A. Jäger

Deviations between the form of trajectory assumed in a fit to a set of measurements and the actual form of the trajectory can give rise to sequential correlations in the residuals from the fit. These correlations can provide a more powerful…

High Energy Physics - Experiment · Physics 2009-10-31 Robert V. Kowalewski , Paul D. Jackson

System-on-Chip (SoC) designs are used in every aspect of computing and their optimization is a difficult but essential task in today's competitive market. Data taken from SoCs to achieve this is often characterised by very long concurrent…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-09-25 Dave McEwan , Jose Nunez-Yanez

The Gini score is a popular tool in statistical modeling and machine learning for model validation and model selection. It is a purely rank based score that allows one to assess risk rankings. The Gini score for statistical modeling has…

Machine Learning · Statistics 2025-11-20 Alexej Brauer , Mario V. Wüthrich

Recommender systems have become crucial in the modern digital landscape, where personalized content, products, and services are essential for enhancing user experience. This paper explores statistical models for recommender systems,…

Methodology · Statistics 2024-08-13 Disha Ghandwani , Trevor Hastie

As machine learning models grow increasingly competent, their predictions can supplement scarce or expensive data in various important domains. In support of this paradigm, algorithms have emerged to combine a small amount of high-fidelity…

Machine Learning · Computer Science 2025-07-08 Zhun Deng , Thomas P Zollo , Benjamin Eyre , Amogh Inamdar , David Madras , Richard Zemel

Sample splitting is widely used in statistical applications, including classically in classification and more recently for inference post model selection. Motivating by problems in the study of diet, physical activity, and health, we…

Methodology · Statistics 2019-08-13 Eli S. Kravitz , Raymond J. Carroll , David Ruppert

Linear regression models are among the models most used in practice, although the practitioners are often not sure whether their assumed linear regression model is at least approximately true. In such situations, only designs for which the…

Statistics Theory · Mathematics 2007-06-13 Wolfgang Bischoff , Frank Miller

Determining the best model or models for a particular data set, a process known as Bayesian model comparison, is a critical part of probabilistic inference. Typically, this process assumes a fixed model-space (that is, a fixed set of…

Quantitative Methods · Quantitative Biology 2019-01-08 Thomas HB FitzGerald , Dorothea Hammerer , Thomas D Sambrook , Will D Penny

In semi-supervised learning, the prevailing understanding suggests that observing additional unlabeled samples improves estimation accuracy for linear parameters only in the case of model misspecification. In this work, we challenge such a…

Methodology · Statistics 2025-09-03 Kai Chen , Yuqian Zhang

Modern foundation models rely heavily on using scaling laws to guide crucial training decisions. Researchers often extrapolate the optimal architecture and hyper parameters settings from smaller training runs by describing the relationship…

Machine Learning · Computer Science 2025-02-27 Margaret Li , Sneha Kudugunta , Luke Zettlemoyer

This paper is an extension of the work about the exponential increase of the power of two non-parametric tests: the $ Z $-test and the chi-square goodness-of-fit test. Subject to having auxiliary information, it is possible to improve…

Statistics Theory · Mathematics 2021-09-03 Mickael Albertus

This article introduces the Multidimensional Research Assessment Matrix of scientific output. Its base notion holds that the choice of metrics to be applied in a research assessment process depends upon the unit of assessment, the research…

Digital Libraries · Computer Science 2014-06-24 Henk F. Moed , Gali Halevi

This paper considers the empirical likelihood (EL) construction of confidence intervals for a linear functional based on right censored lifetime data. Many of the results in literature show that log EL has a limiting scaled chi-square…

Statistics Theory · Mathematics 2012-03-28 Shuyuan He , Wei Liang , Junshan Shen , Grace Yang

Complex systems can be modelled at various levels of detail. Ideally, causal models of the same system should be consistent with one another in the sense that they agree in their predictions of the effects of interventions. We formalise…

High-dimensional tests are applied to find relevant sets of variables and relevant models. If variables are selected by analyzing the sums of products matrices and a corresponding mean-value test is performed, there is the danger that the…

Methodology · Statistics 2012-02-10 Juergen Laeuter , Maciej Rosolowski , Ekkehard Glimm

Objectives: Text categorization has been used in biomedical informatics for identifying documents containing relevant topics of interest. We developed a simple method that uses a chi-square-based scoring function to determine the likelihood…

Information Retrieval · Computer Science 2018-04-03 Andrej Kastrin , Borut Peterlin , Dimitar Hristovski

Consider a problem of predicting a response variable using a set of covariates in a linear regression model. If it is \emph{a priori} known or suspected that a subset of the covariates do not significantly contribute to the overall fit of…

Applications · Statistics 2011-09-13 SM Enayetur Raheem , S. Ejaz Ahmed

At every phase of scientific research, scientists must decide how to allocate limited resources to pursue the research inquiries with the greatest potential. This prioritization dictates which controlled interventions are studied, awarded…

Methodology · Statistics 2022-10-11 Bruce A. Corliss , Yaotian Wang , Heman Shakeri , Philip E. Bourne

In this paper, we consider the extent of the biases that may arise when an unmeasured confounder is omitted from a structural equation model (SEM) and we propose sensitivity analysis techniques to correct for such biases. We give an…

Methodology · Statistics 2021-03-11 Adam J. Sullivan , Tyler J. VanderWeele
‹ Prev 1 4 5 6 7 8 10 Next ›