English
Related papers

Related papers: MRL order, log-concavity and an application to pea…

200 papers

Option-critic learning is a general-purpose reinforcement learning (RL) framework that aims to address the issue of long term credit assignment by leveraging temporal abstractions. However, when dealing with extended timescales, discounting…

Machine Learning · Computer Science 2019-11-21 Akshay Dharmavaram , Matthew Riemer , Shalabh Bhatnagar

Many examples of exactly solvable birth and death processes, a typical stationary Markov chain, are presented together with the explicit expressions of the transition probabilities. They are derived by similarity transforming exactly…

Mathematical Physics · Physics 2015-05-13 Ryu Sasaki

Random processes with stationary increments and intrinsic random processes are two concepts commonly used to deal with non-stationary random processes. They are broader classes than stationary random processes and conceptually closely…

Probability · Mathematics 2025-12-05 Jongwook Kim

A simple model of the new notion of "Markov up" processes is proposed; its positive recurrence and ergodic properties are shown under the appropriate conditions.

Probability · Mathematics 2023-01-02 Alexander Veretennikov , Maria Veretennikova

Most results regarding Skorokhod embedding problems (SEP) so far rely on the assumption that the corresponding stopped process is uniformly integrable, which is equivalent to the convex ordering condition…

Probability · Mathematics 2020-01-01 Jiajie Wang

This work is a continuation of [Kalikaeva, MPRF, 23(2):225-240]. The object of study is ``Markov-up processes'' on $\mathbb Z_+$ and the moment of downcrossing a certain barrier. The processes considered in this paper differ from Markov…

Probability · Mathematics 2024-07-01 Diana Kalikaeva

Learned representations are a central component in modern ML systems, serving a multitude of downstream tasks. When training such representations, it is often the case that computational and statistical constraints for each downstream task…

We study reinforcement learning by combining recent advances in regularized linear programming formulations with the classical theory of stochastic approximation. Motivated by the challenge of designing algorithms that leverage off-policy…

Optimization and Control · Mathematics 2026-04-15 Axel Friedrich Wolter , Tobias Sutter

Symbolic indefinite integration in Computer Algebra Systems such as Maple involves selecting the most effective algorithm from multiple available methods. Not all methods will succeed for a given problem, and when several do, the results,…

Symbolic Computation · Computer Science 2025-08-11 Rashid Barket , Matthew England , Jürgen Gerhard

Multi-objective reinforcement learning (MORL) is the generalization of standard reinforcement learning (RL) approaches to solve sequential decision making problems that consist of several, possibly conflicting, objectives. Generally, in…

Artificial Intelligence · Computer Science 2019-10-08 Xi Chen , Ali Ghadirzadeh , Mårten Björkman , Patric Jensfelt

Non-linear Hawkes processes with memory kernels given by the sum of Erlang kernels are considered. It is shown that their stability properties can be studied in terms of an associated class of piecewise deterministic Markov processes,…

Probability · Mathematics 2018-11-27 Aline Duarte , Eva Löcherbach , Guilherme Ost

We consider a class of semi-Markov processes (SMP) such that the embedded discrete time Markov chain may be non-homogeneous. The corresponding augmented processes are represented as semi-martingales using stochastic integral equation…

Probability · Mathematics 2022-07-14 Anindya Goswami , Subhamay Saha , Ravishankar Kapildev Yadav

We provide a statistical analysis of regularization-based continual learning on a sequence of linear regression tasks, with emphasis on how different regularization terms affect the model performance. We first derive the convergence rate…

Machine Learning · Computer Science 2024-06-11 Xuyang Zhao , Huiyuan Wang , Weiran Huang , Wei Lin

We propose trace logic, an instance of many-sorted first-order logic, to automate the partial correctness verification of programs containing loops. Trace logic generalizes semantics of program locations and captures loop semantics by…

Logic in Computer Science · Computer Science 2020-08-07 Pamina Georgiou , Bernhard Gleiss , Laura Kovács

We use the abstract method of (local) martingale problems in order to give criteria for convergence of stochastic processes. Extending previous notions, the formulation we use is neither restricted to Markov processes (or semimartingales),…

Probability · Mathematics 2021-08-27 David Criens , Peter Pfaffelhuber , Thorsten Schmidt

The proliferation of fast, dense, byte-addressable nonvolatile memory suggests that data might be kept in pointer-rich "in-memory" format across program runs and even process and system crashes. For full generality, such data requires…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-03-17 Wentao Cai , Haosen Wen , H. Alan Beadle , Chris Kjellqvist , Mohammad Hedayati , Michael L. Scott

We consider mark-recapture-recovery (MRR) data of animals where the model parameters are a function of individual time-varying continuous covariates. For such covariates, the covariate value is unobserved if the corresponding individual is…

Methodology · Statistics 2013-12-02 Roland Langrock , Ruth King

We consider Model-Agnostic Meta-Learning (MAML) methods for Reinforcement Learning (RL) problems, where the goal is to find a policy using data from several tasks represented by Markov Decision Processes (MDPs) that can be updated by one…

Machine Learning · Computer Science 2021-11-18 Alireza Fallah , Kristian Georgiev , Aryan Mokhtari , Asuman Ozdaglar

Recent years have seen a rise in interest in terms of using machine learning, particularly reinforcement learning (RL), for production scheduling problems of varying degrees of complexity. The general approach is to break down the…

Machine Learning · Computer Science 2023-02-16 Alexandru Rinciog , Anne Meyer

We present a new approach to termination analysis of logic programs. The essence of the approach is that we make use of general term-orderings (instead of level mappings), like it is done in transformational approaches to logic program…

Programming Languages · Computer Science 2007-05-23 Alexander Serebrenik , Danny De Schreye
‹ Prev 1 8 9 10 Next ›