Pages that link to "Item:Q2527418"
From MaRDI portal
The following pages link to On the linear model with two absorbing barriers (Q2527418):
Displaying 16 items.
- Distributed dynamic reinforcement of efficient outcomes in multiagent coordination and network formation (Q367468) (← links)
- On ergodic two-armed bandits (Q417067) (← links)
- Convergence in models with bounded expected relative hazard rates (Q472194) (← links)
- Nonconvergence to saddle boundary points under perturbed reinforcement learning (Q495761) (← links)
- On a general class of absorbing-barrier learning algorithms (Q1153692) (← links)
- Learning behavior of stochastic automata in the last stage of learning (Q1225030) (← links)
- epsilon-optimality of a general class of learning algorithms (Q1838322) (← links)
- When can the two-armed bandit algorithm be trusted? (Q1879915) (← links)
- Slow learning with small drift in two-absorbing-barrier models (Q2544129) (← links)
- Achieving Unbounded Resolution in<i>Finite</i>Player Goore Games Using Stochastic Automata, and Its Applications (Q2888572) (← links)
- Learning automaton as a model to simulate and analyse learning behaviour in rats (Q3493428) (← links)
- How Fast Is the Bandit? (Q3506304) (← links)
- Learning behaviours of hierarchical structure stochastic automata operating in a non-stationary multi-teacher environment (Q3795235) (← links)
- An application of the stochastic automaton to the investment game (Q3892111) (← links)
- Theoretical considerations of the parameter self-optimization by stochastic automata (Q4145648) (← links)
- Regret bounds for Narendra-Shapiro bandit algorithms (Q5086451) (← links)