The following pages link to Richard S. Sutton (Q1049133):
Displaying 16 items.
- Temporal-difference search in Computer Go (Q420936) (← links)
- Natural actor-critic algorithms (Q1049136) (← links)
- Associative search network: A reinforcement learning associative memory (Q1149264) (← links)
- Landmark learning: An illustration of associative search (Q1154810) (← links)
- Synthesis of nonlinear control surfaces by a layered associative search network (Q1170041) (← links)
- Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning (Q1606316) (← links)
- Reinforcement learning with replacing eligibility traces (Q1911343) (← links)
- Reward is enough (Q2238710) (← links)
- Policy iterations for reinforcement learning problems in continuous time and space -- fundamental theory and methods (Q2664203) (← links)
- An emphatic approach to the problem of off-policy temporal-difference learning (Q2810885) (← links)
- True online temporal-difference learning (Q2834469) (← links)
- On Generalized Bellman Equations and Temporal-Difference Learning (Q3305109) (← links)
- Stimulus Representation and the Timing of Reward-Prediction Errors in Models of the Dopamine System (Q3544332) (← links)
- (Q4626283) (← links)
- (Q5477862) (← links)
- Reward-respecting subtasks for model-based reinforcement learning (Q6088325) (← links)