中文
相关论文

相关论文: Bias-corrected methods for estimating the receiver…

200 篇论文

Model evaluation is of crucial importance in modern statistics application. The construction of ROC and calculation of AUC have been widely used for binary classification evaluation. Recent research generalizing the ROC/AUC analysis to…

机器学习 · 统计学 2024-04-23 Liang Wang , Luis Carvalho

The receiver operating characteristic (ROC) curve and its summary measure, the Area Under the Curve (AUC), are well-established tools for evaluating the efficacy of biomarkers in biomedical studies. Compared to the traditional ROC curve,…

统计方法学 · 统计学 2025-10-20 Ziad Akram Ali Hammouri , Yating Zou , Rahul Ghosal , Juan C. Vidal , Marcos Matabuena

Positivity violations can complicate estimation and interpretation of causal dose-response curves (CDRCs) for continuous interventions. Weighting-based methods are designed to handle limited overlap, but the resulting weighted targets can…

统计方法学 · 统计学 2026-02-13 Han Bao , Michael Schomaker

The ground truth used for training image, video, or speech quality prediction models is based on the Mean Opinion Scores (MOS) obtained from subjective experiments. Usually, it is necessary to conduct multiple experiments, mostly with…

音频与语音处理 · 电气工程与系统科学 2021-12-15 Gabriel Mittag , Saman Zadtootaghaj , Thilo Michael , Babak Naderi , Sebastian Möller

Accurately predicting faulty software units helps practitioners target faulty units and prioritize their efforts to maintain software quality. Prior studies use machine-learning models to detect faulty software code. We revisit past studies…

软件工程 · 计算机科学 2019-01-08 Libo Li , Stefan Lessmann , Bart Baesens

We consider inference on a scalar regression coefficient under a constraint on the magnitude of the control coefficients. A class of estimators based on a regularized propensity score regression is shown to exactly solve a tradeoff between…

计量经济学 · 经济学 2023-08-11 Timothy B. Armstrong , Michal Kolesár , Soonwoo Kwon

Numerous multimodal misinformation benchmarks exhibit bias toward specific modalities, allowing detectors to make predictions based solely on one modality. While previous research has quantified bias at the dataset level or manually…

人工智能 · 计算机科学 2025-11-11 Hehai Lin , Hui Liu , Shilei Cao , Jing Li , Haoliang Li , Wenya Wang

Speed-of-sound (SoS) is a biomechanical characteristic of tissue, and its imaging can provide a promising biomarker for diagnosis. Reconstructing SoS images from ultrasound acquisitions can be cast as a limited-angle computed-tomography…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Sonia Laguna , Lin Zhang , Can Deniz Bezek , Monika Farkas , Dieter Schweizer , Rahel A. Kubik-Huch , Orcun Goksel

With the widespread adoption of machine learning in the real world, the impact of the discriminatory bias has attracted attention. In recent years, various methods to mitigate the bias have been proposed. However, most of them have not…

机器学习 · 计算机科学 2025-03-26 Kenji Kobayashi , Yuri Nakao

In spite of considerable practical importance, current algorithmic fairness literature lacks technical methods to account for underlying geographic dependency while evaluating or mitigating bias issues for spatial data. We initiate the…

应用统计 · 统计学 2022-01-31 Subhabrata Majumdar , Cheryl Flynn , Ritwik Mitra

Robustness checks are routine in empirical work, but there is no standard statistical procedure to formally measure what one can learn from them. I propose a "robustness radius" measure to quantify the amount by which the robustness checks…

计量经济学 · 经济学 2026-02-24 Brenda Prallon

Many fields use the ROC curve and the PR curve as standard evaluations of binary classification methods. Analysis of ROC and PR, however, often gives misleading and inflated performance evaluations, especially with an imbalanced ground…

机器学习 · 统计学 2020-06-23 Chang Cao , Davide Chicco , Michael M. Hoffman

Inter-rater reliability (IRR) is one of the commonly used tools for assessing the quality of ratings from multiple raters. However, applicant selection procedures based on ratings from multiple raters usually result in a binary outcome; the…

统计方法学 · 统计学 2025-06-17 František Bartoš , Patrícia Martinková

Despite the growing demand for accurate surface normal estimation models, existing methods use general-purpose dense prediction models, adopting the same inductive biases as other tasks. In this paper, we discuss the inductive biases needed…

计算机视觉与模式识别 · 计算机科学 2024-03-04 Gwangbin Bae , Andrew J. Davison

Diagnostic tests are almost never perfect. Studies quantifying their performance use knowledge of the true health status, measured with a reference diagnostic test. Researchers commonly assume that the reference test is perfect, which is…

应用统计 · 统计学 2024-08-20 Filip Obradović

Conducting a randomization test is a common method for testing causal null hypotheses in randomized experiments. The popularity of randomization tests is largely because their statistical validity only depends on the randomization design,…

统计方法学 · 统计学 2025-01-15 Siyu Heng , Pamela A. Shaw

Most computer vision application rely on algorithms finding local correspondences between different images. These algorithms detect and compare stable local invariant descriptors centered at scale-invariant keypoints. Because of the…

计算机视觉与模式识别 · 计算机科学 2014-09-10 Ives Rey-Otero , Mauricio Delbracio , Jean-Michel Morel

Considerable interest has recently been focused on studying multiple phenotypes simultaneously in both epidemiological and genomic studies, either to capture the multidimensionality of complex disorders or to understand shared etiology of…

统计方法学 · 统计学 2015-11-26 Denis Agniel , Katherine P. Liao , Tianxi Cai

We propose an evaluation framework for class probability estimates (CPEs) in the presence of label uncertainty, which is commonly observed as diagnosis disagreement between experts in the medical domain. We also formalize evaluation metrics…

机器学习 · 统计学 2021-03-23 Takahiro Mimori , Keiko Sasada , Hirotaka Matsui , Issei Sato

We consider the problem of estimating the false-/ true-positive-rate (FPR/TPR) for a binary classification model when there are incorrect labels (label noise) in the validation set. Our motivating application is fraud prevention where…

机器学习 · 计算机科学 2023-08-08 Justin Tittelfitz
‹ 上一页 1 8 9 10 下一页 ›