English
Related papers

Related papers: Multiple tests for restricted mean time lost with …

200 papers

Many studies employ the analysis of time-to-event data that incorporates competing risks and right censoring. Most methods and software packages are geared towards analyzing data that comes from a continuous failure time distribution.…

Methodology · Statistics 2025-06-06 Tomer Meir , Malka Gorfine

Thul et al. (2020) called attention to problems that arise when chronometric experiments implementing specific factorial designs are analysed with the generalized additive mixed model (GAMM), using factor smooths to capture trial-to-trial…

Methodology · Statistics 2021-11-19 R. Harald Baayen , Matteo Fasiolo , Simon Wood , Yu-Ying Chuang

Recently, there as been an increasing interest in the use of heavily restricted randomization designs which enforces balance on observed covariates in randomized controlled trials. However, when restrictions are strict, there is a risk that…

Methodology · Statistics 2021-10-15 Mattias Nordin , Mårten Schultzberg

Experimentation is widely utilized for causal inference and data-driven decision-making across disciplines. In an A/B experiment, for example, an online business randomizes two different treatments (e.g., website designs) to their customers…

Methodology · Statistics 2025-01-15 Wenxuan Guo , JungHo Lee , Panos Toulis

Balancing between computational efficiency and sample efficiency is an important goal in reinforcement learning. Temporal difference (TD) learning algorithms stochastically update the value function, with a linear time complexity in the…

Machine Learning · Computer Science 2016-11-21 Clement Gehring , Yangchen Pan , Martha White

Considering a regression model, we address the question of testing the nullity of the regression function. The testing procedure is available when the variance of the observations is unknown and does not depend on any prior information on…

Statistics Theory · Mathematics 2019-04-08 Thi Thien Trang Bui

Power and sample size calculations for Wald tests in generalized linear models (GLMs) are often limited to specific cases like logistic regression. More general methods typically require detailed study parameters that are difficult to…

Methodology · Statistics 2026-01-21 Amy L Cochran , Shijie Yuan , Paul J Rathouz

We study a class of iterated empirical risk minimization (ERM) procedures in which two successive ERMs are performed on the same dataset, and the predictions of the first estimator enter as an argument in the loss function of the second.…

Machine Learning · Statistics 2026-02-02 Hugo Cui , Yue M. Lu

Multi-task learning (MTL) is an active field in deep learning in which we train a model to jointly learn multiple tasks by exploiting relationships between the tasks. It has been shown that MTL helps the model share the learned features…

Computer Vision and Pattern Recognition · Computer Science 2021-11-30 Akihiro Nakano , Shi Chen , Kazuyuki Demachi

Accelerated life-tests (ALTs) are used for inferring lifetime characteristics of highly reliable products. In particular, step-stress ALTs increase the stress level at which units under test are subject at certain pre-fixed times, thus…

Statistics Theory · Mathematics 2024-02-12 Narayanaswamy Balakrishnan , Maria Jaenada , Leandro Pardo

Recently, authors have studied inequalities involving expectations of selected functions viz. failure rate, mean residual life, aging intensity function and log-odds rate which are defined for left truncated random variables in reliability…

Statistics Theory · Mathematics 2016-04-27 Chanchal Kundu , Amit Ghosh

As evaluation designs of large language models may shape our trajectory toward artificial general intelligence, comprehensive and forward-looking assessment is essential. Existing benchmarks primarily assess static knowledge, while…

Computation and Language · Computer Science 2025-08-07 Jiayin Wang , Zhiquang Guo , Weizhi Ma , Min Zhang

With some regularity conditions maximum likelihood estimators (MLEs) always produce asymptotically optimal (in the sense of consistency, efficiency, sufficiency, and unbiasedness) estimators. But in general, the MLEs lead to non-robust…

Methodology · Statistics 2024-02-22 Chudamani Poudyal

This paper presents the first study for temporal relation extraction in a zero-shot setting focusing on biomedical text. We employ two types of prompts and five LLMs (GPT-3.5, Mixtral, Llama 2, Gemma, and PMC-LLaMA) to obtain responses…

Computation and Language · Computer Science 2024-06-18 Vasiliki Kougia , Anastasiia Sedova , Andreas Stephan , Klim Zaporojets , Benjamin Roth

The concept of mean inactivity time plays a crucial role in reliability, risk theory and life testing. In this regard, we introduce a weighted mean inactivity time function by considering a non-negative weight function. Based on this…

Probability · Mathematics 2021-03-16 Antonio Di Crescenzo , Abdolsaeed Toomaj

There is a well-known problem in Null Hypothesis Significance Testing: many statistically significant results fail to replicate in subsequent experiments. We show that this problem arises because standard `point-form null' significance…

Methodology · Statistics 2025-02-06 Fintan Costello , Paul Watts

Given an input query, generative models such as large language models produce a random response drawn from a response distribution. Given two input queries, it is natural to ask if their response distributions are the same. While…

Statistics Theory · Mathematics 2025-09-16 Aranyak Acharyya , Carey E. Priebe , Hayden S. Helm

This paper studies inference in randomized controlled trials with multiple treatments, where treatment status is determined according to a "matched tuples" design. Here, by a matched tuples design, we mean an experimental design where units…

Econometrics · Economics 2023-11-06 Yuehao Bai , Jizhou Liu , Max Tabord-Meehan

Prediction performance of a risk scoring system needs to be carefully assessed before its adoption in clinical practice. Clinical preventive care often uses risk scores to screen asymptomatic population. The primary clinical interest is to…

Methodology · Statistics 2018-06-22 Yan Yuan , Qian M. Zhou , Bingying Li , Hengrui Cai , Eric J. Chow , Gregory T. Armstrong

Reinforcement learning (RL) has been successfully used to solve many continuous control tasks. Despite its impressive results however, fundamental questions regarding the sample complexity of RL on continuous problems remain open. We study…

Machine Learning · Computer Science 2017-12-27 Stephen Tu , Benjamin Recht