Pages that link to "Item:Q4537821"
From MaRDI portal
The following pages link to Stochastic Approximation for Nonexpansive Maps: Application to <i>Q</i>-Learning Algorithms (Q4537821):
Displaying 14 items.
- Value iteration and adaptive dynamic programming for data-driven adaptive optimal control design (Q313259) (← links)
- Stochastic approximation with long range dependent and heavy tailed noise (Q383264) (← links)
- Stabilization of stochastic approximation by step size adaptation (Q450652) (← links)
- Reinforcement learning for long-run average cost. (Q1427588) (← links)
- On the convergence of stochastic approximations under a subgeometric ergodic Markov dynamic (Q2044347) (← links)
- The O.D.E. Method for Convergence of Stochastic Approximation and Reinforcement Learning (Q4943730) (← links)
- Some Limit Properties of Markov Chains Induced by Recursive Stochastic Algorithms (Q5037552) (← links)
- Risk-Sensitive Reinforcement Learning via Policy Gradient Search (Q5102286) (← links)
- Technical Note—Consistency Analysis of Sequential Learning Under Approximate Bayesian Inference (Q5130497) (← links)
- Continuous-Time Robust Dynamic Programming (Q5205609) (← links)
- Empirical Q-Value Iteration (Q5856670) (← links)
- Analyzing Approximate Value Iteration Algorithms (Q5868951) (← links)
- Concentration of Contractive Stochastic Approximation and Reinforcement Learning (Q5870773) (← links)
- Stochastic Fixed-Point Iterations for Nonexpansive Maps: Convergence and Error Bounds (Q6180255) (← links)