中文
相关论文

相关论文: RiLACS: Risk-Limiting Audits via Confidence Sequen…

200 篇论文

We propose a simple common framework for Risk-Limiting and Bayesian (polling) audits for two-candidate plurality elections. Using it, we derive an expression for the general Bayesian audit; in particular, we do not restrict the prior to a…

密码学与安全 · 计算机科学 2019-08-06 Poorvi L. Vora

This paper presents DiffSum, a simple post-election risk-limiting ballot-polling audit for two-candidate plurality elections. DiffSum sequentially draws ballots (without replacement) until the numbers $a$, $b$, of votes for candidates $A$,…

计算机与社会 · 计算机科学 2015-09-02 Ronald L. Rivest

The City and County of San Francisco, CA, has used Instant Runoff Voting (IRV) for some elections since 2004. This report describes the first ever process pilot of Risk Limiting Audits for IRV, for the San Francisco District Attorney's race…

计算机与社会 · 计算机科学 2020-04-02 Michelle Blom , Andrew Conway , Dan King , Laurent Sandrolini , Philip B. Stark , Peter J. Stuckey , Vanessa Teague

Tabulation audits for an election provide statistical evidence that a reported contest outcome is "correct" (meaning that the tabulation of votes was properly performed), or else the tabulation audit determines the correct outcome. Stark…

密码学与安全 · 计算机科学 2018-02-13 Ronald L. Rivest

We propose an approach for preventing unsafe or otherwise low-quality large language model (LLM) outputs by leveraging the stochasticity of LLMs, an approach we call Repeated Checking with Regeneration (RCR). In this system, LLM checkers…

人工智能 · 计算机科学 2025-09-30 Jake R. Watts , Joel Sokol

For more than a century, election officials across the United States have inspected voting machines before elections using a procedure called Logic and Accuracy Testing (LAT). This procedure consists of election officials casting a test…

密码学与安全 · 计算机科学 2024-10-01 Braden L. Crimmins , J. Alex Halderman , Bradley Sturt

Sequential estimators are proposed for the relative risk, odds ratio, log relative risk or log odds ratio of a dichotomous attribute in two populations. The estimators take the same number of observations from each population, and guarantee…

统计方法学 · 统计学 2026-04-07 Luis Mendo

Test collections are information-retrieval tools that allow researchers to quickly and easily evaluate ranking algorithms. While test collections have become an integral part of IR research, the process of data creation involves significant…

信息检索 · 计算机科学 2025-07-15 Rikiya Takehi , Ellen M. Voorhees , Tetsuya Sakai , Ian Soboroff

Runtime assurance (RTA) addresses the problem of keeping an autonomous system safe while using an untrusted (or experimental) controller. This can be done via logic that explicitly switches between the untrusted controller and a safety…

计算机科学中的逻辑 · 计算机科学 2023-06-08 Kristina Miller , Christopher K. Zeitler , William Shen , Mahesh Viswanathan , Sayan Mitra

Instant-runoff voting (IRV) is used in several countries around the world. It requires voters to rank candidates in order of preference, and uses a counting algorithm that is more complex than systems such as first-past-the-post or scoring…

We present ATLAS-RTC, a runtime control system for autoregressive language models that enforces structured output during decoding. ATLAS-RTC monitors generation at each step, detects drift from output contracts using lightweight signals,…

机器学习 · 计算机科学 2026-04-07 Christopher Cruz

We propose confidence sequences -- sequences of confidence intervals which are valid uniformly over time -- for quantiles of any distribution over a complete, fully-ordered set, based on a stream of i.i.d. observations. We give methods both…

统计理论 · 数学 2022-07-08 Steven R. Howard , Aaditya Ramdas

Many voter-verifiable, coercion-resistant schemes have been proposed, but even the most carefully designed systems necessarily leak information via the announced result. In corner cases, this may be problematic. For example, if all the…

密码学与安全 · 计算机科学 2019-08-15 Wojciech Jamroga , Peter B. Roenne , Peter Y. A. Ryan , Philip B. Stark

The extensive scope of large language models (LLMs) across various domains underscores the critical importance of responsibility in their application, beyond natural language processing. In particular, the randomized nature of LLMs, coupled…

计算与语言 · 计算机科学 2024-04-19 Sana Ebrahimi , Nima Shahbazi , Abolfazl Asudeh

The main risk-limiting ballot polling audit in use today, BRAVO, is designed for use when single ballots are drawn at random and a decision regarding whether to stop the audit or draw another ballot is taken after each ballot draw…

密码学与安全 · 计算机科学 2021-02-23 Filip Zagórski , Grant McClearn , Sarah Morin , Neal McBurnett , Poorvi L. Vora

Large language models generate outputs stochastically and may produce fluent but invalid responses, including factual hallucinations. Existing mitigation strategies reduce average error rates but do not provide explicit control over the…

机器学习 · 计算机科学 2026-02-03 Sreenivasan Mohandas

Linguistic cues such as "I believe" and "probably" offer an intuitive interface for communicating confidence, yet a generalisable, principled calibration framework for linguistic confidence expressions remains underexplored. In particular,…

计算与语言 · 计算机科学 2026-05-20 Yi-Fan Yeh , Linwei Tao , Minjing Dong , Tao Huang , Jialin Yu , Philip Torr , Chang Xu

Randomized smoothing (RS) is one of the prominent techniques to ensure the correctness of machine learning models, where point-wise robustness certificates can be derived analytically. While RS is well understood for classification, its…

机器学习 · 计算机科学 2025-09-22 Emmanouil Seferis , Changshun Wu , Stefanos Kollias , Saddek Bensalem , Chih-Hong Cheng

Meta-analysis allows rigorous aggregation of estimates and uncertainty across multiple studies. When a given study reports multiple estimates, such as log odds ratios (ORs) or log relative risks (RRs) across exposure groups, accounting for…

统计方法学 · 统计学 2024-07-02 Alexander Johnson-Vázquez , Alexander W. Hsu , Peng Zheng , Aleksandr Aravkin

Detecting the presence of anomalies in regression models is a crucial task in machine learning, as anomalies can significantly impact the accuracy and reliability of predictions. Random Sample Consensus (RANSAC) is one of the most popular…

机器学习 · 统计学 2024-10-22 Le Hong Phong , Ho Ngoc Luat , Vo Nguyen Le Duy