English

Computational Hardness of Static Distributionally Robust Markov Decision Processes

Optimization and Control 2026-05-08 v5

Abstract

We present some hardness results on finding the optimal policy for the static formulation of distributionally robust Markov decision processes. We construct problem instances such that when the considered policy class is Markovian and non-randomized, finding the optimal policy is NP-hard. When the considered policy class is Markovian and randomized, the robust value function possesses sub-optimal strict local minimizers, and finding the optimal policy is also NP-hard. The considered instances involve an ambiguity set with only two transition kernels.

Keywords

Cite

@article{arxiv.2511.02224,
  title  = {Computational Hardness of Static Distributionally Robust Markov Decision Processes},
  author = {Yan Li},
  journal= {arXiv preprint arXiv:2511.02224},
  year   = {2026}
}
R2 v1 2026-07-01T07:20:32.788Z