Pages that link to "Item:Q3119642"
From MaRDI portal
The following pages link to Estimate and approximate policy iteration algorithm for discounted Markov decision models with bounded costs and Borel spaces (Q3119642):
Displaying 12 items.
- A perturbation approach to a class of discounted approximate value iteration algorithms with Borel spaces (Q330284) (← links)
- Suboptimal policy determination for large-scale Markov decision processes. I: Description and bounds (Q799497) (← links)
- Approximation of discounted minimax Markov control problems and zero-sum Markov games using Hausdorff and Wasserstein distances (Q1741211) (← links)
- Block-successive approximation for a discounted Markov decision model (Q2265958) (← links)
- Complexity bounds for approximately solving discounted MDPs by value iterations (Q2661516) (← links)
- PAC Bounds for Discounted MDPs (Q3164829) (← links)
- Approximate Fixed Point Iteration with an Application to Infinite Horizon Markov Decision Processes (Q3399246) (← links)
- (Q3496174) (← links)
- (Q3730373) (← links)
- A perturbation approach to approximate value iteration for average cost Markov decision processes with Borel spaces and bounded costs (Q5227201) (← links)
- Time-varying Markov decision processes with state-action-dependent discount factors and unbounded costs (Q5227206) (← links)
- A New Policy Evaluation Algorithm for Markov Decision Processes with Quasi Birth-Death Structure (Q5462817) (← links)