English
Related papers

Related papers: Taming the Heavy Tail: Age-Optimal Preemption

200 papers

We introduce and test an algorithm that adaptively estimates large deviation functions characterizing the fluctuations of additive functionals of Markov processes in the long-time limit. These functions play an important role for predicting…

Statistical Mechanics · Physics 2023-03-30 Grégoire Ferré , Hugo Touchette

Distributed optimization has become the default training paradigm in modern machine learning due to the growing scale of models and datasets. To mitigate communication overhead, local updates are often applied before global aggregation,…

Machine Learning · Computer Science 2025-08-15 Su Hyeong Lee , Manzil Zaheer , Tian Li

We study the policy evaluation problem in multi-agent reinforcement learning where a group of agents, with jointly observed states and private local actions and rewards, collaborate to learn the value function of a given policy via local…

Optimization and Control · Mathematics 2021-11-08 Dongsheng Ding , Xiaohan Wei , Zhuoran Yang , Zhaoran Wang , Mihailo R. Jovanović

We consider a modulated process S which, conditional on a background process X, has independent increments. Assuming that S drifts to -infinity and that its increments (jumps) are heavy-tailed (in a sense made precise in the paper), we…

Probability · Mathematics 2017-11-29 Sergey Foss , Takis Konstantopoulos , Stan Zachary

We study the problem of estimating the mean of a distribution in high dimensions when either the samples are adversarially corrupted or the distribution is heavy-tailed. Recent developments in robust statistics have established efficient…

Data Structures and Algorithms · Computer Science 2021-01-20 Samuel B. Hopkins , Jerry Li , Fred Zhang

This study presents a robust optimization algorithm for automated highway merge. The merging scenario is one of the challenging scenes in automated driving, because it requires adjusting ego vehicle's speed to match other vehicles before…

Robotics · Computer Science 2025-03-20 Takeru Goto , Kosuke Toda , Takayasu Kumano

Age of Information (AoI) has been recognized as an important metric to measure the freshness of information. Central to this consensus is that minimizing AoI can enhance the freshness of information, thereby facilitating the accuracy of…

Information Theory · Computer Science 2024-08-09 Aimin Li , Shaohua Wu , Gary C. F. Lee , Xiaomeng Cheng , Sumei Sun

We consider piecewise-deterministic optimal control problems in which the environment randomly switches among several deterministic modes, and the goal is to optimize the expected cost up to the termination while taking the likelihood of…

Optimization and Control · Mathematics 2015-12-31 Zhengdi Shen , Alexander Vladimirsky

We consider Piecewise Deterministic Markov Processes (PDMPs) with a finite set of discrete states. In the regime of fast jumps between discrete states, we prove a law of large number and a large deviation principle. In the regime of fast…

Probability · Mathematics 2008-09-16 A. Faggionato , D. Gabrielli , M. Ribezzi Crivellari

Age of Information (AoI) is a crucial metric for quantifying information freshness in real-time systems where the sampling rate of data packets is time-varying. Evaluating AoI under such conditions is challenging, as system states become…

Information Theory · Computer Science 2025-07-08 Jin Xu , Weiqi Wang , Natarajan Gautam

While the notion of age of information (AoI) has recently emerged as an important concept for analyzing ultra-reliable low-latency communications (URLLC), the majority of the existing works have focused on the average AoI measure. However,…

Networking and Internet Architecture · Computer Science 2019-11-28 Mohamed K. Abdel-Aziz , Chen-Feng Liu , Sumudu Samarakoon , Mehdi Bennis , Walid Saad

New sampling algorithms based on simulating continuous-time stochastic processes called piece-wise deterministic Markov processes (PDMPs) have shown considerable promise. However, these methods can struggle to sample from multi-modal or…

Methodology · Statistics 2022-05-31 Matthew Sutton , Robert Salomone , Augustin Chevallier , Paul Fearnhead

The purpose of this paper is to introduce a new Markov chain Monte Carlo method and exhibit its efficiency by simulation and high-dimensional asymptotic theory. Key fact is that our algorithm has a reversible proposal transition kernel,…

Methodology · Statistics 2014-12-22 Kengo Kamatani

In this note, we study distributed time-varying optimization for a multi-agent system. We first focus on a class of time-varying quadratic cost functions, and develop a new distributed algorithm that integrates an average estimator and an…

Systems and Control · Electrical Eng. & Systems 2024-08-06 Liangze Jiang , Zheng-Guang Wu , Lei Wang

An implicit mass-matrix penalization (IMMP) of Hamiltonian dynamics is proposed, and associated dynamical integrators, as well as sampling Monte-Carlo schemes, are analyzed for systems with multiple time scales. The penalization is based on…

Numerical Analysis · Mathematics 2009-06-01 Petr Plechac , Mathias Rousset

In offline reinforcement learning (RL), the absence of active exploration calls for attention on the model robustness to tackle the sim-to-real gap, where the discrepancy between the simulated and deployed environments can significantly…

Machine Learning · Computer Science 2024-06-28 He Wang , Laixi Shi , Yuejie Chi

We present the first finite-sample analysis of policy evaluation in robust average-reward Markov Decision Processes (MDPs). Prior work in this setting have established only asymptotic convergence guarantees, leaving open the question of…

Machine Learning · Statistics 2025-12-11 Yang Xu , Washim Uddin Mondal , Vaneet Aggarwal

Despite the successes of probabilistic models based on passing noise through neural networks, recent work has identified that such methods often fail to capture tail behavior accurately, unless the tails of the base distribution are…

Machine Learning · Statistics 2023-06-16 Feynman Liang , Liam Hodgkinson , Michael W. Mahoney

Output-length prediction is important for efficient LLM serving, as it directly affects batching, memory reservation, and scheduling. For prompt-only length prediction, most existing methods use a one-shot sampled length as the label,…

Machine Learning · Computer Science 2026-04-10 Jing Wang , Yu-Yang Qian , Ke Xue , Chao Qian , Peng Zhao , Zhi-Hua Zhou

A new method for stochastic control based on neural networks and using randomisation of discrete random variables is proposed and applied to optimal stopping time problems. The method models directly the policy and does not need the…

Computational Finance · Quantitative Finance 2021-01-11 Thomas Deschatre , Joseph Mikael