English
Related papers

Related papers: Information-Theoretic Limits of Safety Verificatio…

200 papers

We study best-policy identification for finite-horizon risk-sensitive reinforcement learning under the entropic risk measure. Recent work established a constant gap in the exponential horizon dependence between lower and upper bounds on the…

Machine Learning · Computer Science 2026-05-14 Amer Essakine , Claire Vernade

We prove that no continuous, utility-preserving wrapper defense-a function $D: X\to X$ that preprocesses inputs before the model sees them-can make all outputs strictly safe for a language model with connected prompt space, and we…

Cryptography and Security · Computer Science 2026-04-14 Manish Bhatt , Sarthak Munshi , Vineeth Sai Narajala , Idan Habler , Ammar Al-Kahfah , Ken Huang , Joel Webb , Blake Gatto , Md Tamjidul Hoque

We investigate the problem of safety verification of infinite-state parameterized programs that are formed based on a rich class of topologies. We introduce a new proof system, called parametric proof spaces, which exploits the underlying…

Logic in Computer Science · Computer Science 2026-01-27 Ruotong Cheng , Azadeh Farzan

Differentially private (DP) mechanisms are difficult to interpret and calibrate because existing methods for mapping standard privacy parameters to concrete privacy risks -- re-identification, attribute inference, and data reconstruction --…

In this paper we consider the problem of uniformity testing with limited memory. We observe a sequence of independent identically distributed random variables drawn from a distribution $p$ over $[n]$, which is either uniform or is…

Information Theory · Computer Science 2022-06-22 Tomer Berg , Or Ordentlich , Ofer Shayevitz

We consider the problem of provably optimal exploration in reinforcement learning for finite horizon MDPs. We show that an optimistic modification to value iteration achieves a regret bound of $\tilde{O}( \sqrt{HSAT} + H^2S^2A+H\sqrt{T})$…

Machine Learning · Statistics 2017-07-04 Mohammad Gheshlaghi Azar , Ian Osband , Rémi Munos

Proposition. Let $f$ be a predictor trained on a distribution $P$ and evaluated on a shifted distribution $Q$. Under verifiable regularity and complexity constraints, the excess risk under shift admits an explicit upper bound determined by…

Machine Learning · Computer Science 2026-02-23 Chandrasekhar Gokavarapu , Sudhakar Gadde , Y. Rajasekhar , S. R. Bhargava

We study the matrix completion problem that leverages hierarchical similarity graphs as side information in the context of recommender systems. Under a hierarchical stochastic block model that well respects practically-relevant social…

Information Theory · Computer Science 2021-09-14 Junhyung Ahn , Adel Elmahdy , Soheil Mohajer , Changho Suh

We prove that no reinforcement learning policy with confidence-gated autonomy can simultaneously achieve maximum helpfulness, optimal calibration, and full autonomy under rational oversight, whenever some tasks exceed the agent's reliable…

Machine Learning · Computer Science 2026-05-26 Lauri Lovén , Nam Do , Hassan Mehmood , Dinesh Kumar Sah , Sasu Tarkoma

Constructing confidence intervals that are simultaneously valid across a class of estimates is central to tasks such as multiple mean estimation, generalization guarantees, and adaptive experimental design. We frame this as an ``error…

Machine Learning · Computer Science 2026-02-05 Sanath Kumar Krishnamurthy , Anna Lyubarskaja , Emma Brunskill , Susan Athey

This paper studies the problem of enforcing safety of a stochastic dynamical system over a finite time horizon. We use stochastic barrier functions as a means to quantify the probability that a system exits a given safe region of the state…

Systems and Control · Computer Science 2019-05-30 Cesar Santoyo , Maxence Dutreix , Samuel Coogan

Let $K$ be a field of degree $n$ and discriminant with absolute value $\Delta$. Under the assumption of the validity of the Generalized Riemann Hypothesis, we provide a new algorithm to compute a set of generators of the class group of $K$…

Number Theory · Mathematics 2025-06-19 Loïc Grenié , Giuseppe Molteni

We address the issue of safety in reinforcement learning. We pose the problem in an episodic framework of a constrained Markov decision process. Existing results have shown that it is possible to achieve a reward regret of…

Machine Learning · Computer Science 2023-01-26 Tao Liu , Ruida Zhou , Dileep Kalathil , P. R. Kumar , Chao Tian

This letter presents an approach to guarantee online safety of a cyber-physical system under multiple state and input constraints. Our proposed framework, called gatekeeper, recursively guarantees the existence of an infinite-horizon…

Systems and Control · Electrical Eng. & Systems 2025-08-14 Devansh R. Agrawal , Dimitra Panagou

Self-play reinforcement learning trains language models on their own generated tasks, co-evolving a proposer and solver without human labels. Recent systems report strong reasoning gains, but collapse and instability are widely observed and…

Machine Learning · Computer Science 2026-05-22 Sophia Xiao Pu , Zhaotian Weng , Chengzhi Liu , Jayanth Srinivasa , Gaowen Liu , William Yang Wang , Xin Eric Wang

Quantum mechanics provides extraordinarily accurate probabilistic predictions, yet the framework remains silent on what distinguishes quantum systems from definite measurement outcomes. This paper develops a measurement-theoretic framework…

History and Philosophy of Physics · Physics 2026-02-02 Jonathon Sendall

This paper addresses the fundamental question of determining the minimum number of distinct control laws required for global controllability of nonlinear systems that exhibit singularities in their feedback linearising controllers. We…

Computational Engineering, Finance, and Science · Computer Science 2026-03-20 Nikolaos D. Tantaroudas

Interpretable and explainable machine learning has seen a recent surge of interest. We focus on safety as a key motivation behind the surge and make the relationship between interpretability and safety more quantitative. Toward assessing…

Machine Learning · Computer Science 2022-11-04 Dennis Wei , Rahul Nair , Amit Dhurandhar , Kush R. Varshney , Elizabeth M. Daly , Moninder Singh

To apply reinforcement learning to safety-critical applications, we ought to provide safety guarantees during both policy training and deployment. In this work, we present theoretical results that place a bound on the probability of…

Machine Learning · Computer Science 2025-06-24 Jacques Cloete , Nikolaus Vertovec , Alessandro Abate

The MDL two-part coding $ \textit{index of resolvability} $ provides a finite-sample upper bound on the statistical risk of penalized likelihood estimators over countable models. However, the bound does not apply to unpenalized maximum…

Statistics Theory · Mathematics 2018-01-01 W. D. Brinda , Jason M. Klusowski