The following pages link to (Q5405224):
Displaying 7 items.
- Time-varying policy rule under learning (Q500481) (← links)
- Reduced complexity dynamic programming based on policy iteration (Q1206904) (← links)
- Kernel dynamic policy programming: applicable reinforcement learning to robot systems with high dimensional states (Q2292214) (← links)
- On linear and super-linear convergence of natural policy gradient algorithm (Q2670744) (← links)
- Applications of variable discounting dynamic programming to iterated function systems and related problems (Q4621340) (← links)
- (Q4999027) (← links)
- (Q4999029) (← links)