English
Related papers

Related papers: Renewal processes with costs and rewards

200 papers

This paper develops an algorithmic-based approach for proving inductive properties of propositional sequent systems such as admissibility, invertibility, cut-elimination, and identity expansion. Although undecidable in general, these…

Logic in Computer Science · Computer Science 2021-01-11 Carlos Olarte , Elaine Pimentel , Camilo Rocha

In constrained reinforcement learning (RL), a learning agent seeks to not only optimize the overall reward but also satisfy the additional safety, diversity, or budget constraints. Consequently, existing constrained RL solutions require…

Machine Learning · Computer Science 2021-07-13 Sobhan Miryoosefi , Chi Jin

This paper provides a rigorous and gap-free proof of the index theorem used in the theory of regular economy. In the index theorem that is the subject of this paper, the assumptions for the excess demand function are only several usual…

Theoretical Economics · Economics 2023-06-27 Yuhki Hosoya

In this paper, we obtain some additional probabilistic properties of the renewal process $\{\hat{N}_{\alpha}(t)\}_{t\ge0}$, $0<\alpha\le 1$ introduced by Beghin and Orsingher (2010). A time-changed relationship connecting…

Probability · Mathematics 2026-04-09 Mostafizar Khandakar , Bratati Pal

The standard feedback model of reinforcement learning requires revealing the reward of every visited state-action pair. However, in practice, it is often the case that such frequent feedback is not available. In this work, we take a first…

Machine Learning · Computer Science 2021-03-08 Yonathan Efroni , Nadav Merlis , Shie Mannor

Repeated recursion unfolding is a new approach that repeatedly unfolds a recursion with itself and simplifies it while keeping all unfolded rules. Each unfolding doubles the number of recursive steps covered. This reduces the number of…

Programming Languages · Computer Science 2020-09-14 Thom Fruehwirth

Desirability can be understood as an extension of Anscombe and Aumann's Bayesian decision theory to sets of expected utilities. At the core of desirability lies an assumption of linearity of the scale in which rewards are measured. It is a…

Artificial Intelligence · Computer Science 2022-11-21 Enrique Miranda , Marco Zaffalon

The most promising recent methods for AI reasoning require applying variants of reinforcement learning (RL) either on rolled out trajectories from the LLMs, even for the step-wise rewards, or large quantities of human-annotated trajectory…

Artificial Intelligence · Computer Science 2025-06-25 Sara Rajaee , Kumar Pratik , Gabriele Cesa , Arash Behboodi

We describe a new framework for causal inference and its application to return time series. In this system, causal relationships are represented as logical formulas, allowing us to test arbitrarily complex hypotheses in a computationally…

Statistical Finance · Quantitative Finance 2010-06-14 Samantha Kleinberg , Petter N. Kolm , Bud Mishra

Survey of several forms of updating, with a practical illustrative example. We study several updating (conditioning) schemes that emerge naturally from a common scenarion to provide some insights into their meaning. Updating is a subtle…

Artificial Intelligence · Computer Science 2013-03-26 Philippe Smets

Long chain-of-thought (CoT) significantly enhances the reasoning capabilities of large language models (LLMs). However, extensive reasoning traces lead to inefficiencies and increased time-to-first-token (TTFT). We propose a training…

Computation and Language · Computer Science 2026-01-08 Roy Xie , David Qiu , Deepak Gopinath , Dong Lin , Yanchao Sun , Chong Wang , Saloni Potdar , Bhuwan Dhingra

This paper discusses the semantics and proof theory of Nilsson's probabilistic logic, outlining both the benefits of its well-defined model theory and the drawbacks of its proof theory. Within Nilsson's semantic framework, we derive a set…

Artificial Intelligence · Computer Science 2013-04-11 Peter Haddawy , Alan M. Frisch

Recent advances in applying reinforcement learning (RL) to large language models (LLMs) have led to substantial progress. In particular, a series of remarkable yet often counterintuitive phenomena have been reported in LLMs, exhibiting…

Machine Learning · Computer Science 2025-09-03 Haoze Wu , Cheng Wang , Wenshuo Zhao , Junxian He

We propose a useful approach for investigating the statistical properties of foreign currency exchange rates. Our approach is based on queueing theory, particularly, the so-called renewal-reward theorem. For the first passage processes of…

Data Analysis, Statistics and Probability · Physics 2008-12-02 Jun-ichi Inoue , Naoya Sazuka

This paper is a brief and informal presentation of cirquent calculus, a novel proof system for resource-conscious logics. As such, it is a refinement of sequent calculus with mechanisms that allow to explicitly account for the possibility…

Logic in Computer Science · Computer Science 2021-08-31 Giorgi Japaridze , Bikal Lamichhane

Many economic theory models incorporate finiteness assumptions that, while introduced for simplicity, play a real role in the analysis. We provide a principled framework for scaling results from such models by removing these finiteness…

Computer Science and Game Theory · Computer Science 2023-04-11 Yannai A. Gonczarowski , Scott Duke Kominers , Ran I. Shorrer

Transfer of recent advances in deep reinforcement learning to real-world applications is hindered by high data demands and thus low efficiency and scalability. Through independent improvements of components such as replay buffers or more…

Machine Learning · Computer Science 2022-11-28 André Eberhard , Houssam Metni , Georg Fahland , Alexander Stroh , Pascal Friederich

We study the large-time asymptotic of renewal-reward processes with a heavy-tailed waiting time distribution. It is known that the heavy tail of the distribution produces an extremely slow dynamics, resulting in a singular large deviation…

Mathematical Physics · Physics 2022-01-05 Hiroshi Horii , Raphael Lefevere , Takahiro Nemoto

The axiom of recovery, while capturing a central intuition regarding belief change, has been the source of much controversy. We argue briefly against putative counterexamples to the axiom--while agreeing that some of their insight deserves…

Artificial Intelligence · Computer Science 2007-05-23 Samir Chopra , Aditya Ghose , Thomas Meyer

In this letter we examine a model recently proposed to produce phase synchronization [K. Wood et al, Phys. Rev. Lett. 96, 145701 (2006)] and we show that the onset to synchronization corresponds to the emergence of an intermittent process…

Statistical Mechanics · Physics 2007-05-23 Simone Bianco , Elvis Geneston , Paolo Grigolini , Massimiliano Ignaccolo
‹ Prev 1 8 9 10 Next ›