中文
相关论文

相关论文: Asymptotics and Renewal Approximation in the Onlin…

200 篇论文

Let $v_n$ be the maximum expected length of an increasing subsequence, which can be selected by an online nonanticipating policy from a random sample of size $n$. Refining known estimates, we obtain an asymptotic expansion of $v_n$ up to a…

最优化与控制 · 数学 2018-08-21 Amirlan Seksenbayev

Given a sequence of independent random variables with a common continuous distribution, we consider the online decision problem where one seeks to minimize the expected value of the time that is needed to complete the selection of a…

概率论 · 数学 2016-09-05 Alessandro Arlotto , Elchanan Mossel , J. Michael Steele

Consider a sequence of $n$ independent random variables with a common continuous distribution $F$, and consider the task of choosing an increasing subsequence where the observations are revealed sequentially and where an observation must be…

概率论 · 数学 2016-08-02 Alessandro Arlotto , Vinh V. Nguyen , J. Michael Steele

This note is motivated by connections between the online and offline problems of selecting a possibly long subsequence from a Poisson-paced sequence of uniform marks under either a monotonicity or a sum constraint. The offline problem with…

概率论 · 数学 2020-07-16 Alexander Gnedin

Given a sequence of $n$ independent random variables with common continuous distribution, we propose a simple adaptive online policy that selects a monotone increasing subsequence. We show that the expected number of monotone increasing…

概率论 · 数学 2019-10-22 Alessandro Arlotto , Yehua Wei , Xinchang Xie

This paper considers online optimization for a system that performs a sequence of back-to-back tasks. Each task can be processed in one of multiple processing modes that affect the duration of the task, the reward earned, and an additional…

最优化与控制 · 数学 2024-01-17 Michael J. Neely

We provide asymptotic approximations to the distribution of statistics that are obtained from network data for limiting sequences that let the number of nodes (agents) in the network grow large. Network formation is permitted to be…

综合经济学 · 经济学 2021-11-03 Konrad Menzel

We find a two term asymptotic expansion for the optimal expected value of a sequentially selected monotone subsequence from a random permutation of length n. A striking feature of this expansion is that tells us that the expected value of…

概率论 · 数学 2015-09-16 Peichao Peng , J. Michael Steele

This paper studies a long-term resource allocation problem over multiple periods where each period requires a multi-stage decision-making process. We formulate the problem as an online allocation problem in an episodic finite-horizon…

数据结构与算法 · 计算机科学 2023-10-20 Duksang Lee , William Overman , Dabeen Lee

We survey the theory of increasing and decreasing subsequences of permutations. Enumeration problems in this area are closely related to the RSK algorithm. The asymptotic behavior of the expected value of the length is(w) of the longest…

组合数学 · 数学 2007-05-23 Richard P. Stanley

We revisit the problem of estimating the parameters of a partially observed diffusion process, consisting of a hidden state process and an observed process, with a continuous time parameter. The estimation is to be done online, i.e. the…

最优化与控制 · 数学 2018-10-16 Simone Carlo Surace , Jean-Pascal Pfister

We study the $(\varepsilon, \delta)$-PAC policy identification problem in finite-horizon episodic Markov Decision Processes. Existing approaches provide finite-time guarantees for approximate settings ($\varepsilon>0$) but suffer from high…

机器学习 · 计算机科学 2026-05-06 Cyrille Kone , Kevin Jamieson

We propose a novel randomized linear programming algorithm for approximating the optimal policy of the discounted Markov decision problem. By leveraging the value-policy duality and binary-tree data structures, the algorithm adaptively…

最优化与控制 · 数学 2019-06-04 Mengdi Wang

Scaled type Markov renewal processes generalize classical renewal processes: renewal times come from a one parameter family of probability laws and the sequence of the parameters is the trajectory of an ergodic Markov chain. Our primary…

概率论 · 数学 2015-03-17 Zsolt Pajor-Gyulai , Domokos Szász

We resolve the fundamental problem of online decoding with general $n^{th}$ order ergodic Markov chain models. Specifically, we provide deterministic and randomized algorithms whose performance is close to that of the optimal offline…

机器学习 · 计算机科学 2019-05-31 Vikas K. Garg , Tamar Pichkhadze

We consider the problem of computing the value and an optimal strategy for minimizing the expected termination time in one-counter Markov decision processes. Since the value may be irrational and an optimal strategy may be rather…

形式语言与自动机理论 · 计算机科学 2012-05-08 Tomáš Brázdil , Antonín Kučera , Petr Novotný , Dominik Wojtczak

The Poisson compound decision problem is a long-standing problem in statistics, where empirical Bayes methodologies are commonly used to estimate Poisson's means in static or batch domains. In this paper, we study the Poisson compound…

统计方法学 · 统计学 2025-06-10 Stefano Favaro , Sandra Fortini

We analyze the optimal policy for the sequential selection of an alternating subsequence from a sequence of $n$ independent observations from a continuous distribution $F$, and we prove a central limit theorem for the number of selections…

概率论 · 数学 2016-09-05 Alessandro Arlotto , J. Michael Steele

In this paper, we study a continuous-time discounted jump Markov decision process with both controlled actions and observations. The observation is only available for a discrete set of time instances. At each time of observation, one has to…

最优化与控制 · 数学 2019-07-16 Yunhan Huang , Veeraruna Kavitha , Quanyan Zhu

This paper presents a novel algorithm for efficient online estimation of the filter derivatives in general hidden Markov models. The algorithm, which has a linear computational complexity and very limited memory requirements, is furnished…

统计计算 · 统计学 2019-01-10 Jimmy Olsson , Johan Westerborn Alenlöv
‹ 上一页 1 2 3 10 下一页 ›