Unrestricted supremum for target discounted-sum objectives in MDPs
Determine the unrestricted optimal probability of satisfying a target discounted-sum objective in a finite Markov decision process, without restricting strategies to finite memory.
References
For the supremum, computing the unrestricted optimal probability remains open.
— Target Discounted Sum Problem on Markov Chains with Applications to Markov Decision Processes
(2609.03670 - Bertrand et al., 3 Sep 2026) in Section 1, Applications