中文
相关论文

相关论文: "Medium-n studies" in computing education conferen…

200 篇论文

Randomized trials are considered the gold standard for making informed decisions in medicine, yet they often lack generalizability to the patient populations in clinical practice. Observational studies, on the other hand, cover a broader…

统计方法学 · 统计学 2026-04-14 Piersilvio De Bartolomeis , Javier Abad , Konstantin Donhauser , Fanny Yang

We study a game theoretic model of standardized testing for college admissions. Students are of two types; High and Low. There is a college that would like to admit the High type students. Students take a potentially costly standardized…

计算机科学与博弈论 · 计算机科学 2021-02-17 Sampath Kannan , Mingzi Niu , Aaron Roth , Rakesh Vohra

The randomized $p$-value, (nonrandomized) mid-$p$-value and abstract randomized $p$-value have all been recommended for testing a null hypothesis whenever the test statistic has a discrete distribution. This paper provides a unifying…

统计计算 · 统计学 2014-12-02 Joshua D Habiger

Probability forecasts for binary events play a central role in many applications. Their quality is commonly assessed with proper scoring rules, which assign forecasts a numerical score such that a correct forecast achieves a minimal…

统计方法学 · 统计学 2022-07-04 Alexander Henzi , Johanna F. Ziegel

An important aspect of multiple hypothesis testing is controlling the significance level, or the level of Type I error. When the test statistics are not independent it can be particularly challenging to deal with this problem, without…

统计理论 · 数学 2009-03-04 Sandy Clarke , Peter Hall

Data analysis is a powerful tool in all experimental sciences. Statistical methods, such as sampling theory, computer technologies necessary for handling large amounts of data, skill in analysing information contained in different types of…

物理教育 · 物理学 2012-06-20 Vera Montalbano

Consistently checking the statistical significance of experimental results is the first mandatory step towards reproducible science. This paper presents a hitchhiker's guide to rigorous comparisons of reinforcement learning algorithms.…

统计方法学 · 统计学 2022-08-30 Cédric Colas , Olivier Sigaud , Pierre-Yves Oudeyer

Standard multiple testing procedures are designed to report a list of discoveries, or suspected false null hypotheses, given the hypotheses' p-values or test scores. Recently there has been a growing interest in enhancing such procedures by…

统计方法学 · 统计学 2025-10-29 Jack Freestone , William Stafford Noble , Uri Keich

In multiple hypothesis testing, the volume of data, defined as the number of replications per null times the total number of nulls, usually defines the amount of resource required. On the other hand, power is an important measure of…

统计理论 · 数学 2009-06-05 Zhiyi Chi

Simulations play a crucial role in the modern scientific process. Yet despite (or due to) this ubiquity, the Data Science community shares neither a comprehensive definition for a "high-quality" study nor a consolidated guide to designing…

统计计算 · 统计学 2025-05-16 Corrine F Elliott , James PC Duncan , Tiffany M Tang , Merle Behr , Karl Kumbier , Bin Yu

Since the emergence of deep learning and its adoption in steganalysis fields, most of the reference articles kept using small to medium size CNN, and learn them on relatively small databases. Therefore, benchmarks and comparisons between…

密码学与安全 · 计算机科学 2021-01-01 Hugo Ruiz , Marc Chaumont , Mehdi Yedroudj , Ahmed Oulad Amara , Frédéric Comby , Gérard Subsol

In many empirical studies of a large two-sided matching market (such as in a college admissions problem), the researcher performs statistical inference under the assumption that they observe a random sample from a large matching market. In…

计量经济学 · 经济学 2024-04-02 Jacob Schwartz , Kyungchul Song

Simulations are valuable tools for empirically evaluating the properties of statistical methods and are primarily employed in methodological research to draw general conclusions about methods. In addition, they can often be useful to…

其他统计学 · 统计学 2025-10-08 Anne-Laure Boulesteix , Patrick Callahan , Luzia Hanssum , Vincent Gaertner , Eva Hoster

Trials enroll a large number of subjects in order to attain power, making them expensive and time-consuming. Sample size calculations are often performed with the assumption of an unadjusted analysis, even if the trial analysis plan…

统计方法学 · 统计学 2021-07-06 Alejandro Schuler

Having a sufficient quantity of quality data is a critical enabler of training effective machine learning models. Being able to effectively determine the adequacy of a dataset prior to training and evaluating a model's performance would be…

机器学习 · 计算机科学 2026-04-28 Arya Hatamian , Lionel Levine , Haniyeh Ehsani Oskouie , Majid Sarrafzadeh

Developing state-of-the-art approaches for specific tasks is a major driving force in our research community. Depending on the prestige of the task, publishing it can come along with a lot of visibility. The question arises how reliable are…

机器学习 · 计算机科学 2018-03-28 Nils Reimers , Iryna Gurevych

Computing Education Research (CER) is critical for supporting the increasing number of students who need to learn computing skills. To systematically advance knowledge, publications must be clear enough to support replications,…

计算机与社会 · 计算机科学 2021-10-20 Sarah Heckman , Jeffrey C. Carver , Mark Sherriff , Ahmed Al-Zubidy

The Leiden Rankings can be used for grouping research universities by considering universities which are not statistically significantly different as homogeneous sets. The groups and intergroup relations can be analyzed and visualized using…

数字图书馆 · 计算机科学 2018-10-16 Loet Leydesdorff , Lutz Bornmann , John Mingers

For clinical studies with continuous outcomes, when the data are potentially skewed, researchers may choose to report the whole or part of the five-number summary (the sample median, the first and third quartiles, and the minimum and…

统计方法学 · 统计学 2023-05-09 Jiandong Shi , Dehui Luo , Xiang Wan , Yue Liu , Jiming Liu , Zhaoxiang Bian , Tiejun Tong

In multiple testing several criteria to control for type I errors exist. The false discovery rate, which evaluates the expected proportion of false discoveries among the rejected null hypotheses, has become the standard approach in this…

统计方法学 · 统计学 2023-11-03 Jacobo de Uña-Álvarez