中文
相关论文

相关论文: When Doesn't Cokriging Outperform Kriging?

200 篇论文

We consider how to make probability forecasts of binary labels. Our main mathematical result is that for any continuous gambling strategy used for detecting disagreement between the forecasts and the actual labels, there exists a…

机器学习 · 计算机科学 2007-05-23 Vladimir Vovk , Akimichi Takemura , Glenn Shafer

Researchers are increasingly subjecting artificial intelligence systems to psychological testing. But to rigorously compare their cognitive capacities with humans and other animals, we must avoid both over- and under-stating our…

人工智能 · 计算机科学 2025-03-05 Konstantinos Voudouris , Lucy G. Cheke , Eric Schulz

Prediction becomes more challenging with missing covariates. What method is chosen to handle missingness can greatly affect how models perform. In many real-world problems, the best prediction performance is achieved by models that can…

Mining 29,000 accounting ratios for t-statistics $> 2.0$ leads to cross-sectional return predictability similar to the peer review process. For both, $\approx50\%$ of predictability remains after the original sample periods. This finding…

综合金融 · 定量金融 2026-01-01 Andrew Y. Chen , Alejandro Lopez-Lira , Tom Zimmermann

A long noted difficulty when assessing the reliability (or calibration) of forecasting systems is that reliability, in general, is a hypothesis not about a finite dimensional parameter but about an entire functional relationship. A…

数据分析、统计与概率 · 物理学 2020-12-09 Jochen Bröcker

A common method to reduce the uncertainty of causal inferences from experiments is to assign treatments in fixed proportions within groups of similar units: blocking. Previous results indicate that one can expect substantial reductions in…

统计方法学 · 统计学 2015-08-31 Fredrik Sävje

Recent investigations in noise contrastive estimation suggest, both empirically as well as theoretically, that while having more "negative samples" in the contrastive loss improves downstream classification performance initially, beyond a…

机器学习 · 计算机科学 2022-06-24 Pranjal Awasthi , Nishanth Dikkala , Pritish Kamath

Test log-likelihood is commonly used to compare different models of the same data or different approximate inference algorithms for fitting the same probabilistic model. We present simple examples demonstrating how comparisons based on test…

机器学习 · 统计学 2024-01-22 Sameer K. Deshpande , Soumya Ghosh , Tin D. Nguyen , Tamara Broderick

Combining measurements which have "theoretical uncertainties" is a delicate matter, due to an unclear statistical basis. We present an algorithm based on the notion that a theoretical uncertainty represents an estimate of bias.

数据分析、统计与概率 · 物理学 2011-08-05 F. C. Porter

This report examines whether advanced AIs that perform well in training will be doing so in order to gain power later -- a behavior I call "scheming" (also sometimes called "deceptive alignment"). I conclude that scheming is a disturbingly…

计算机与社会 · 计算机科学 2023-11-29 Joe Carlsmith

We study the relationship between performance and practice by analyzing the activity of many players of a casual online game. We find significant heterogeneity in the improvement of player performance, given by score, and address this by…

计算机与社会 · 计算机科学 2017-03-16 Tushar Agarwal , Keith A. Burghardt , Kristina Lerman

Correctly evaluating defenses against adversarial examples has proven to be extremely difficult. Despite the significant amount of recent work attempting to design defenses that withstand adaptive attacks, few have succeeded; most papers…

Experiments deliver credible treatment-effect estimates but, because they are costly, are often restricted to specific sites, small populations, or particular mechanisms. A common practice across several fields is therefore to combine…

计量经济学 · 经济学 2025-12-30 Aristotelis Epanomeritakis , Davide Viviano

We study the problem of prediction with expert advice with adversarial corruption where the adversary can at most corrupt one expert. Using tools from viscosity theory, we characterize the long-time behavior of the value function of the…

机器学习 · 计算机科学 2021-03-02 Erhan Bayraktar , Ibrahim Ekren , Xin Zhang

Patterns of wins and losses in pairwise contests, such as occur in sports and games, consumer research and paired comparison studies, and human and animal social hierarchies, are commonly analyzed using probabilistic models that allow one…

物理与社会 · 物理学 2025-11-03 Maximilian Jerdee , M. E. J. Newman

Graph contrastive learning has shown great promise when labeled data is scarce, but large unlabeled datasets are available. However, it often does not take uncertainty estimation into account. We show that a variational Bayesian neural…

机器学习 · 计算机科学 2023-12-04 Alexander Möllers , Alexander Immer , Elvin Isufi , Vincent Fortuin

We study how to perform tests on samples of pairs of observations and predictions in order to assess whether or not the predictions are prudent. Prudence requires that that the mean of the difference of the observation-prediction pairs can…

风险管理 · 定量金融 2022-10-03 Dirk Tasche

We conduct an incentivized lab experiment to test participants' ability to understand the DA matching mechanism and the strategyproofness property, conveyed in different ways. We find that while many participants can (using a novel GUI)…

综合经济学 · 经济学 2024-09-30 Yannai A. Gonczarowski , Ori Heffetz , Guy Ishai , Clayton Thomas

Global explanations of a reinforcement learning (RL) agent's expected behavior can make it safer to deploy. However, such explanations are often difficult to understand because of the complicated nature of many RL policies. Effective human…

机器学习 · 计算机科学 2022-11-16 Sanjana Narayanan , Isaac Lage , Finale Doshi-Velez

Collaborative filtering is a rapidly advancing research area. Every year several new techniques are proposed and yet it is not clear which of the techniques work best and under what conditions. In this paper we conduct a study comparing…

信息检索 · 计算机科学 2012-05-16 Joonseok Lee , Mingxuan Sun , Guy Lebanon