English
Related papers

Related papers: A Fine-grained Analysis of Fitted Q-evaluation: Be…

200 papers

Quality-Diversity (QD) algorithms have shown remarkable success in discovering diverse, high-performing solutions, but rely heavily on hand-crafted behavioral descriptors that constrain exploration to predefined notions of diversity.…

Machine Learning · Computer Science 2026-03-05 Saeed Hedayatian , Stefanos Nikolaidis

We study the problem of off-policy evaluation (OPE) in Reinforcement Learning (RL), where the aim is to estimate the performance of a new policy given historical data that may have been generated by a different policy, or policies. In…

Machine Learning · Computer Science 2019-12-16 Aurélien F. Bibaut , Ivana Malenica , Nikos Vlassis , Mark J. van der Laan

Estimating quantum entropies and divergences is an important problem in quantum physics, information theory, and machine learning. Quantum neural estimators (QNEs), which utilize a hybrid classical-quantum architecture, have recently…

Quantum Physics · Physics 2026-05-27 Sreejith Sreekumar , Ziv Goldfeld , Mark M. Wilde

This paper studies the non-parametric estimation and uniform inference for the conditional quantile regression function (CQRF) with covariates exposed to measurement errors. We consider the case that the distribution of the measurement…

Methodology · Statistics 2025-04-03 Haoze Hou , Wei Huang , Zheng Zhang

Modelling agent preferences has applications in a range of fields including economics and increasingly, artificial intelligence. These preferences are not always known and thus may need to be estimated from observed behavior, in which case…

Computer Science and Game Theory · Computer Science 2023-03-02 Daniel Chui , Jason Hartline , James R. Wright

Greedy-GQ with linear function approximation, originally proposed in \cite{maei2010toward}, is a value-based off-policy algorithm for optimal control in reinforcement learning, and it has a non-linear two timescale structure with the…

Machine Learning · Computer Science 2024-05-03 Yue Wang , Yi Zhou , Shaofeng Zou

Quantum-enhanced metrology aims to estimate an unknown parameter such that the precision scales better than the shot-noise bound. Single-shot adaptive quantum-enhanced metrology (AQEM) is a promising approach that uses feedback to tweak the…

Quantum Physics · Physics 2016-10-17 Pantita Palittapongarnpim , Peter Wittek , Barry C. Sanders

We study policy evaluation of offline contextual bandits subject to unobserved confounders. Sensitivity analysis methods are commonly used to estimate the policy value under the worst-case confounding over a given uncertainty set. However,…

Machine Learning · Statistics 2026-01-13 Kei Ishikawa , Niao He , Takafumi Kanamori

Value-based reinforcement-learning algorithms have shown strong results in games, robotics, and other real-world applications. Overestimation bias is a known threat to those algorithms and can sometimes lead to dramatic performance…

Machine Learning · Computer Science 2024-08-13 Martin Waltz , Ostap Okhrin

Diffusion models have shown impressive performance in capturing complex and multi-modal action distributions for game agents, but their slow inference speed prevents practical deployment in real-time game environments. While consistency…

Artificial Intelligence · Computer Science 2025-03-24 Ruoqi Zhang , Ziwei Luo , Jens Sjölund , Per Mattsson , Linus Gisslén , Alessandro Sestini

Problems in fields such as healthcare, robotics, and finance requires reasoning about the value both of what decision or action to take and when to take it. The prevailing hope is that artificial intelligence will support such decisions by…

Machine Learning · Computer Science 2025-03-21 Yoav Wald , Mark Goldstein , Yonathan Efroni , Wouter A. C. van Amsterdam , Rajesh Ranganath

Quantum estimation theory provides optimal observations for various estimation problems for unknown parameters in the state of the system under investigation. However, the theory has been developed under the assumption that every observable…

Quantum Physics · Physics 2009-11-10 M. Hotta , M. Ozawa

Quantum error correction (QEC) is essential for realizing scalable quantum computation. However, when evaluating its benefits, most analyses assume idealized components, overlooking the imperfections inherent in realistic fault-tolerant…

Quantum Physics · Physics 2026-05-26 Lorenzo Valentini , Diego Forlivesi , Marco Chiani

We revisit the problem of estimating an unknown parameter of a pure quantum state, and investigate `null-measurement' strategies in which the experimenter aims to measure in a basis that contains a vector close to the true system state.…

Quantum Physics · Physics 2023-10-11 Federico Girotti , Alfred Godley , Mădălin Guţă

Analogy-Based Estimation (ABE) is a popular method for non-algorithmic estimation due to its simplicity and effectiveness. The Analogy-Based Estimation (ABE) model was proposed by researchers, however, no optimal approach for reliable…

Software Engineering · Computer Science 2025-12-02 Tarun Chintada , Uday Kiran Cheera

Practitioners use feature importance to rank and eliminate weak predictors during model development in an effort to simplify models and improve generality. Unfortunately, they also routinely conflate such feature importance measures with…

Machine Learning · Computer Science 2020-06-09 Terence Parr , James D. Wilson , Jeff Hamrick

As machine learning models grow increasingly competent, their predictions can supplement scarce or expensive data in various important domains. In support of this paradigm, algorithms have emerged to combine a small amount of high-fidelity…

Machine Learning · Computer Science 2025-07-08 Zhun Deng , Thomas P Zollo , Benjamin Eyre , Amogh Inamdar , David Madras , Richard Zemel

For observational studies, we study the sensitivity of causal inference when treatment assignments may depend on unobserved confounders. We develop a loss minimization approach for estimating bounds on the conditional average treatment…

Methodology · Statistics 2022-03-11 Steve Yadlowsky , Hongseok Namkoong , Sanjay Basu , John Duchi , Lu Tian

Retriever-augmented instruction-following models are attractive alternatives to fine-tuned approaches for information-seeking tasks such as question answering (QA). By simply prepending retrieved documents in its input along with an…

Computation and Language · Computer Science 2024-04-18 Vaibhav Adlakha , Parishad BehnamGhader , Xing Han Lu , Nicholas Meade , Siva Reddy

The purpose of this article is to develop a general parametric estimation theory that allows the derivation of the limit distribution of estimators in non-regular models where the true parameter value may lie on the boundary of the…

Statistics Theory · Mathematics 2022-11-28 Junichiro Yoshida , Nakahiro Yoshida
‹ Prev 1 8 9 10 Next ›