Skip to content
Kudos AI

Tagged “temporal-difference”

1 article.

9 min readReinforcement Learning

Reinforcement Learning and Q-Learning

Learning to act well without a model of the world: temporal-difference updates, the Q-learning rule, exploration versus exploitation, and a run that recovers the planned optimum from experience alone.

Reinforcement LearningMachine LearningArtificial Intelligence