The following pages link to 10.1162/153244303765208377 (Q3044104):
Displaying 44 items.
- R-MAX (Q15078) (← links)
- Belief and truth in hypothesised behaviours (Q274416) (← links)
- A synthesis of automated planning and reinforcement learning for efficient, robust decision-making (Q334800) (← links)
- Knows what it knows: a framework for self-aware learning (Q413843) (← links)
- Learning to compete, coordinate, and cooperate in repeated games using reinforcement learning (Q413845) (← links)
- Reducing reinforcement learning to KWIK online regression (Q616761) (← links)
- Statistical estimation with bounded memory (Q693352) (← links)
- Efficient learning equilibrium (Q814627) (← links)
- Stochastic feature mapping for PAC-Bayes classification (Q890313) (← links)
- An analysis of model-based interval estimation for Markov decision processes (Q959899) (← links)
- If multi-agent learning is the answer, what is the question? (Q1028919) (← links)
- Perspectives on multiagent learning (Q1028921) (← links)
- Decentralized reinforcement learning of robot behaviors (Q1748471) (← links)
- Efficient exploration through active learning for value function approximation in reinforcement learning (Q1784573) (← links)
- A joint Gaussian process model for active visual recognition with expertise estimation in crowdsourcing (Q1800023) (← links)
- Dimension reduction and its application to model-based exploration in continuous spaces (Q1959610) (← links)
- Adaptive-resolution reinforcement learning with polynomial exploration in deterministic domains (Q1959632) (← links)
- Multi-agent reinforcement learning: a selective overview of theories and algorithms (Q2094040) (← links)
- Controller exploitation-exploration reinforcement learning architecture for computing near-optimal policies (Q2318167) (← links)
- Guiding exploration by pre-existing knowledge without modifying reward (Q2383522) (← links)
- AWESOME: a general multiagent learning algorithm that converges in self-play and learns a best response against stationary opponents (Q2384141) (← links)
- Relational reinforcement learning with guided demonstrations (Q2407440) (← links)
- Bayesian optimistic Kullback-Leibler exploration (Q2425228) (← links)
- Robust Algorithms via PAC-Bayes and Laplace Distributions (Q2805741) (← links)
- Reinforcement learning in robust Markov decision processes (Q2833106) (← links)
- (Q4636987) (← links)
- Bayesian Exploration for Approximate Dynamic Programming (Q4971589) (← links)
- Cooperative learning with joint state value approximation for multi-agent systems (Q4980275) (← links)
- (Q4998915) (← links)
- (Q4999029) (← links)
- (Q5053336) (← links)
- Fictitious Play in Zero-Sum Stochastic Games (Q5093269) (← links)
- (Q5149240) (← links)
- (Q5214215) (← links)
- Induction and Exploitation of Subgoal Automata for Reinforcement Learning (Q5856492) (← links)
- Model-based Reinforcement Learning: A Survey (Q5870792) (← links)
- Robust Control for Dynamical Systems with Non-Gaussian Noise via Formal Abstractions (Q5881801) (← links)
- Explicit explore, exploit, or escape \((E^4)\): near-optimal safety-constrained reinforcement learning in polynomial time (Q6106432) (← links)
- Recent advances in reinforcement learning in finance (Q6146668) (← links)
- From Reinforcement Learning to Deep Reinforcement Learning: An Overview (Q6162303) (← links)
- Independent learning in stochastic games (Q6200215) (← links)
- Exploiting action impact regularity and exogenous state variables for offline reinforcement learning (Q6488780) (← links)
- Transferable dynamics models for efficient object-oriented reinforcement learning (Q6494370) (← links)
- Finding the optimal exploration-exploitation trade-off online through Bayesian risk estimation and minimization (Q6566614) (← links)