中文
相关论文

相关论文: "Medium-n studies" in computing education conferen…

200 篇论文

The conference peer review process involves three constituencies with different objectives: authors want their papers accepted at prestigious venues (and quickly), conferences want to present a program with many high-quality and few…

计算机科学与博弈论 · 计算机科学 2025-09-30 Yichi Zhang , Fang-Yi Yu , Grant Schoenebeck , David Kempe

This paper raises concerns about the advantages of using statistical significance tests in research assessments as has recently been suggested in the debate about proper normalization procedures for citation indicators. Statistical…

数字图书馆 · 计算机科学 2012-09-26 Jesper W. Schneider

Many real-world classification problems are significantly class-imbalanced to detriment of the class of interest. The standard set of proper evaluation metrics is well-known but the usual assumption is that the test dataset imbalance equals…

机器学习 · 计算机科学 2020-04-16 Jan Brabec , Tomáš Komárek , Vojtěch Franc , Lukáš Machlica

Randomized controlled trials (RCTs) are increasingly prevalent in education research, and are often regarded as a gold standard of causal inference. Two main virtues of randomized experiments are that they (1) do not suffer from…

In the big data era, the need to reevaluate traditional statistical methods is paramount due to the challenges posed by vast datasets. While larger samples theoretically enhance accuracy and hypothesis testing power without increasing false…

统计方法学 · 统计学 2026-01-09 Xuekui Zhang , Li Xing , Jing Zhang , Soojeong Kim

Introductory statistical inference texts and courses treat the point estimation, hypothesis testing, and interval estimation problems separately, with primary emphasis on large-sample approximations. Here I present an alternative approach…

其他统计学 · 统计学 2017-07-14 Ryan Martin

We point out that the ideas underlying some test procedures recently proposed for testing post-model-selection (and for some other test problems) in the econometrics literature have been around for quite some time in the statistics…

统计理论 · 数学 2017-08-30 Hannes Leeb , Benedikt M. Pötscher

In many settings, robust data analysis involves computational methods for uncertainty quantification and statistical inference. To design frequentist studies that leverage robust analysis methods, suitable sample sizes to achieve desired…

统计方法学 · 统计学 2025-12-19 Luke Hagar , Andrew J. Martin

In multigroup data settings with small within-group sample sizes, standard $F$-tests of group-specific linear hypotheses can have low power, particularly if the within-group sample sizes are not large relative to the number of explanatory…

统计方法学 · 统计学 2022-03-25 Andrew McCormack , Peter Hoff

Attacks on the P-value are nothing new, but the recent attacks are increasingly more serious. They come from more mainstream sources, with widening targets such as a call to retire the significance testing altogether. While well meaning, I…

其他统计学 · 统计学 2022-01-11 Yudi Pawitan

In this paper, we study the hypothesis testing problem of, among $n$ random variables, determining $k$ random variables which have different probability distributions from the rest $(n-k)$ random variables. Instead of using separate…

信息论 · 计算机科学 2013-05-28 Weiyu Xu , Lifeng Lai

Most scientific disciplines use significance testing to draw conclusions about experimental or observational data. This classical approach provides a theoretical guarantee for controlling the number of false positives across a set of…

应用统计 · 统计学 2023-03-06 Stanley E. Lazic

Contemporary statistical publications rely on simulation to evaluate performance of new methods and compare them with established methods. In the context of meta-analysis of log-odds-ratios, we investigate how the ways in which simulations…

统计方法学 · 统计学 2020-07-06 Elena Kulinskaya , David C. Hoaglin , Ilyas Bakbergenuly

Distributed frameworks are widely used to handle massive data, where sample size $n$ is very large, and data are often stored in $k$ different machines. For a random vector $X\in \mathbb{R}^p$ with expectation $\mu$, testing the mean vector…

统计方法学 · 统计学 2021-10-07 Bin Du , Junlong Zhao

It is generally believed that more observations provide more information. However, we observe that in the independence test for rare events, the power of the test is, surprisingly, determined by the number of rare events rather than the…

统计方法学 · 统计学 2025-06-17 Danyang Huang , Liyuan Wang , Liping Zhu

The purpose of this project was to collect and analyse data about the comparability and real-life applicability of published results focusing on Microsoft Windows malware, more specifically the impact of dataset size and testing dataset…

密码学与安全 · 计算机科学 2022-06-14 David Illes

Consistently checking the statistical significance of experimental results is one of the mandatory methodological steps to address the so-called "reproducibility crisis" in deep reinforcement learning. In this tutorial paper, we explain how…

机器学习 · 计算机科学 2018-07-06 Cédric Colas , Olivier Sigaud , Pierre-Yves Oudeyer

Prior research has raised concerns about students' over-reliance on large language models (LLMs) in higher education. This paper examines how Computer Science students and instructors engage with LLMs across five scenarios: "Writing",…

人机交互 · 计算机科学 2026-02-06 Xinrui Lin , Heyan Huang , Shumin Shi , John Vines

In our chapter we address the statistical analysis of percentiles: How should the citation impact of institutions be compared? In educational and psychological testing, percentiles are already used widely as a standard to evaluate an…

数字图书馆 · 计算机科学 2014-04-16 Richard Williams , Lutz Bornmann

Is it possible to reliably evaluate the quality of peer reviews? We study this question driven by two primary motivations -- incentivizing high-quality reviewing using assessed quality of reviews and measuring changes to review quality in…

数字图书馆 · 计算机科学 2024-11-08 Alexander Goldberg , Ivan Stelmakh , Kyunghyun Cho , Alice Oh , Alekh Agarwal , Danielle Belgrave , Nihar B. Shah