English
Related papers

Related papers: A Win-Expectancy Framework for Contextualizing Run…

200 papers

A new model, which uses the frequency of individuals' annual home run totals, is employed to predict future home run totals and records in Major League Baseball. Complete home run frequency data from 1903--2005 is analyzed, resulting in…

Popular Physics · Physics 2007-05-23 D. J. Kelley , J. R. Mureika , J. A. Phillips

Control has long been recognized as a critical component of pitcher performance, reflecting a pitcher's ability to execute pitches in alignment with his intended targets. However, accurately inferring a pitcher's intentions presents a…

Applications · Statistics 2025-08-27 Matt Ludwig , Ryan S. Brill , Abraham J. Wyner

Composite endpoints are widely used in cardiovascular clinical trials to improve statistical efficiency while preserving clinical relevance. The Win Ratio (WR) measure and more general frameworks of Win Statistics have emerged as…

Methodology · Statistics 2025-11-24 Yunhan Mou , Fan Li , Denise Esserman , Yuan Huang

Results in contact sports like Rugby are mainly interpreted in terms of the ability and/or luck of teams. But this neglects the important role of the {\em motivation} of players, reflected in the effort exerted in the game. Here we present…

Applications · Statistics 2022-08-30 Federico Fioravanti , Fernando Delbianco , Fernando Tohmé

We propose a new learning to rank algorithm, named Weighted Margin-Rank Batch loss (WMRB), to extend the popular Weighted Approximate-Rank Pairwise loss (WARP). WMRB uses a new rank estimator and an efficient batch training algorithm. The…

Machine Learning · Statistics 2017-11-15 Kuan Liu , Prem Natarajan

Learning-based trajectory prediction models have encountered great success, with the promise of leveraging contextual information in addition to motion history. Yet, we find that state-of-the-art forecasting methods tend to overly rely on…

Computer Vision and Pattern Recognition · Computer Science 2022-04-22 Hédi Ben-Younes , Éloi Zablocki , Mickaël Chen , Patrick Pérez , Matthieu Cord

Foundation Models (FMs) trained on Electronic Health Records (EHRs) have achieved state-of-the-art results on numerous clinical prediction tasks. However, most existing EHR FMs have context windows of <1k tokens. This prevents them from…

When answering user queries, LLMs often retrieve knowledge from external sources stored in retrieval-augmented generation (RAG) databases. These are often populated from unvetted sources, e.g. the open web, and can contain maliciously…

Cryptography and Security · Computer Science 2026-03-27 Hao Wu , Prateek Saxena

Evaluating offensive linemen and pass rushers at the player level is difficult because observable outcomes are sparse, opponent-dependent, and strongly shaped by surrounding context. Using 2021 regular-season Hudl tracking data, we…

Traditional security scanners fail when facing new attack patterns they haven't seen before. They rely on fixed rules and predetermined signatures, making them blind to novel threats. We present a fundamentally different approach: instead…

Cryptography and Security · Computer Science 2025-11-21 Ayush Chaudhary

In limited overs cricket, the team batting first posts a target score for the team batting second to achieve in order to win the match. The team batting second is constrained by decreasing resources in terms of number of balls left and…

Applications · Statistics 2026-01-21 Rhitankar Bandyopadhyay , Dibyojyoti Bhattacharjee

Agentic systems are evaluated on benchmarks where agents interact with environments to solve tasks. Most papers report a pass@1 score computed from a single run per task, assuming this gives a reliable performance estimate. We test this…

Machine Learning · Computer Science 2026-03-26 Bjarni Haukur Bjarnason , André Silva , Martin Monperrus

We present a framework that gives a deep insight into the link between physical and technical-tactical aspects of soccer and it allows associating physical performance with value generation thanks to a top-down approach. First, we estimate…

Machine Learning · Statistics 2022-04-06 Sergio Llana , Borja Burriel , Pau Madrero , Javier Fernández

A new metric \texttt{BaryScore} to evaluate text generation based on deep contextualized embeddings e.g., BERT, Roberta, ELMo) is introduced. This metric is motivated by a new framework relying on optimal transport tools, i.e., Wasserstein…

Computation and Language · Computer Science 2021-09-10 Pierre Colombo , Guillaume Staerman , Chloe Clavel , Pablo Piantanida

In the summer of 2017, the National Basketball Association reduced the number of total timeouts, along with other rule changes, to regulate the flow of the game. With these rule changes, it becomes increasingly important for coaches to…

Applications · Statistics 2022-08-01 Connor Gibbs , Ryan Elmore , Bailey Fosdick

Several performance metrics for quantifying the in-game performances of individual football players have been proposed in recent years. Although the majority of the on-the-ball actions during games constitutes of passes, many of the…

Applications · Statistics 2018-10-05 Lotte Bransen , Jan Van Haaren

Iterated reference games - in which players repeatedly pick out novel referents using language - present a test case for agents' ability to perform context-sensitive pragmatic reasoning in multi-turn linguistic environments. We tested…

Computation and Language · Computer Science 2025-11-07 Alvin Wei Ming Tan , Ben Prystawski , Veronica Boyce , Michael C. Frank

In many sports, player re-identification is crucial for automatic video processing and analysis. However, most of the current studies on player re-identification in multi- or single-view sports videos focus on re-identification in the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-19 Tomohiro Suzuki , Kazushi Tsutsui , Kazuya Takeda , Keisuke Fujii

The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes domains suggest otherwise. We hypothesize that the primary difference between the two is…

Computation and Language · Computer Science 2026-04-10 Berkin Durmus , Chen Cen , Eduardo Pacheco , Arda Okan , Atila Orhon

Traditional Relative Efficiency (RE), based solely on variance, has limitations in evaluating estimator performance, particularly in planned missing data designs. We introduce Bhirkuti's Relative Efficiency (BRE), a novel metric that…

Methodology · Statistics 2025-05-01 Aneel Bhusal , Todd D. Little