English
Related papers

Related papers: The 2008 election: A preregistered replication ana…

200 papers

The promising performances of CNNs often overshadow the need to examine whether they are doing in the way we are actually interested. We show through experiments that even over-parameterized models would still solve a dataset by recklessly…

Computer Vision and Pattern Recognition · Computer Science 2022-05-24 Xuan Cheng , Tianshu Xie , Xiaomin Wang , Jiali Deng , Minghui Liu , Ming Liu

Data analysis based on information from several sources is common in economic and biomedical studies. This setting is often referred to as the data fusion problem, which differs from traditional missing data problems since no complete data…

Methodology · Statistics 2022-04-07 Wei Li , Shanshan Luo , Wangli Xu

In observational surveys, post-stratification is used to reduce bias resulting from differences between the survey population and the population under investigation. However, this can lead to inflated post-stratification weights and,…

Applications · Statistics 2016-06-24 Yannick Vandendijck , Christel Faes , Niel Hens

Imputing missing values is an important preprocessing step in data analysis, but the literature offers little guidance on how to choose between different imputation models. This letter suggests adopting the imputation model that generates a…

Methodology · Statistics 2021-07-13 Moritz Marbach

Language model agents reason from scratch on every query, discarding their chain of thought after each run. The result is lower accuracy and high run-to-run variance. We introduce reasoning graphs, which persist the per-evidence chain of…

Artificial Intelligence · Computer Science 2026-05-07 Matthew Penaroza

Understanding representational similarity between neural recordings and computational models is essential for neuroscience, yet remains challenging to measure reliably due to the constraints on the number of neurons that can be recorded…

Disordered Systems and Neural Networks · Physics 2025-10-27 Hyunmo Kang , Abdulkadir Canatar , SueYeon Chung

The Implicit Association Test, IAT, is widely used to measure hidden (subconscious) human biases, implicit bias, of many topics: race, gender, age, ethnicity, religion stereotypes. There is a need to understand the reliability of these…

Applications · Statistics 2023-12-27 S. Stanley Young , Warren B. Kindzierski

We consider numerical schemes for root finding of noisy responses through generalizing the Probabilistic Bisection Algorithm (PBA) to the more practical context where the sampling distribution is unknown and location-dependent. As in…

Machine Learning · Statistics 2017-11-03 Sergio Rodriguez , Michael Ludkovski

Graphs and networks are common ways of depicting biological information. In biology, many different biological processes are represented by graphs, such as regulatory networks, metabolic pathways and protein--protein interaction networks.…

Applications · Statistics 2010-11-16 Caiyan Li , Hongzhe Li

Large language models (LLMs) acquire general linguistic knowledge from massive-scale pretraining. However, pretraining data mainly comprised of web-crawled texts contain undesirable social biases which can be perpetuated or even amplified…

Computation and Language · Computer Science 2025-09-04 Takuma Udagawa , Yang Zhao , Hiroshi Kanayama , Bishwaranjan Bhattacharjee

The forecasting of political, economic, and public health indicators using internet activity has demonstrated mixed results. For example, while some measures of explicitly surveyed public opinion correlate well with social media proxies,…

Social and Information Networks · Computer Science 2021-07-14 Yi Li , Asieh Ahani , Haimao Zhan , Kevin Foley , Thayer Alshaabi , Kelsey Linnell , Peter Sheridan Dodds , Christopher M. Danforth , Adam Fox

Regularization is often used in high-dimensional regression settings to generate a sparse model, which can save tremendous computing resources and identify predictors that are most strongly associated with the response. When the predictors…

Machine Learning · Statistics 2026-05-07 Jia Wei He , R. Ayesha Ali , Gerarda Darlington

Gerrymandering voting districts is one of the most salient concerns of contemporary American society, and the creation of new voting maps, along with their subsequent legal challenges, speaks for much of our modern political discourse. The…

Physics and Society · Physics 2023-07-31 Casey Garner , Allen Holder

Bayesian Improved Surname Geocoding (BISG) is the most popular method for proxying race/ethnicity in voter registration files that do not contain it. This paper benchmarks BISG against a range of previously untested machine learning…

Machine Learning · Computer Science 2022-08-02 Ari Decter-Frain

Overestimation of turnout has long been an issue in election surveys, with nonresponse bias or voter overrepresentation identified as major sources of bias. However, adjusting for nonignorable nonresponse bias is substantially challenging.…

Methodology · Statistics 2026-04-07 Xinyu Li , Naiwen Ying , Kendrick Qijun Li , Xu Shi , Wang Miao

Artificial intelligence (AI) systems increasingly inform medical decision-making, yet concerns about algorithmic bias and inequitable outcomes persist, particularly for historically marginalized populations. This paper introduces the…

Machine Learning · Computer Science 2025-07-22 Andrés Morales-Forero , Lili J. Rueda , Ronald Herrera , Samuel Bassetto , Eric Coatanea

Election rules are formal processes that aggregate voters preferences, typically to select a single candidate, called the winner. Most of the election rules studied in the literature require the voters to rank the candidates from the most…

Data Structures and Algorithms · Computer Science 2019-01-31 Matthias Bentert , Piotr Skowron

Randomized controlled trials (RCTs) are the benchmark for causal inference, yet field implementation can drift from the registered design or, by chance, yield imbalances. We introduce a remote audit -- a preregistrable, design-based…

Methodology · Statistics 2025-10-21 Connor T. Jerzak , Adel Daoud

Replication helps ensure that a genotype-phenotype association observed in a genome-wide association (GWA) study represents a credible association and is not a chance finding or an artifact due to uncontrolled biases. We discuss…

Methodology · Statistics 2010-10-26 Peter Kraft , Eleftheria Zeggini , John P. A. Ioannidis

Replicability is central to scientific progress, and the partial conjunction (PC) hypothesis testing framework provides an objective tool to quantify it across disciplines. Existing PC methods assume independent studies. Yet many modern…

Methodology · Statistics 2025-12-30 Monitirtha Dey , Trambak Banerjee , Prajamitra Bhuyan , Arunabha Majumdar
‹ Prev 1 8 9 10 Next ›