Pages that link to "Item:Q3968777"
From MaRDI portal
The following pages link to Bounds for the regret loss in dynamic programming under adaptive control (Q3968777):
Displaying 7 items.
- Adaptive control of Markov processes with incomplete state information and unknown parameters (Q1071659) (← links)
- Continuous dependence of stochastic control models on the noise distribution (Q1100128) (← links)
- Nonparametric adaptive control of discrete-time partially observable stochastic systems (Q1122548) (← links)
- First-order sensitivity of the optimal value in a Markov decision model with respect to deviations in the transition probability function (Q2216181) (← links)
- On truncations and perturbations of Markov decision problems with an application to queueing network overflow control (Q2638971) (← links)
- Estimation and control in discounted stochastic dynamic programming (Q3758580) (← links)
- Generalized Lipschitz-continuity of integrals with respect to a parameter of the intergrating probability measure (Q3759593) (← links)