CANONICAL HISTORY
TD-Gammon — AAAI Fall Symposium Publication Milestone
Gerald Tesauro published a TD-Gammon paper for AAAI-FS 1993 on 22 October 1993. IBM Research describes TD-Gammon as a neural network that learned to play backgammon through self-play using the TD(lambda) reinforcement-learning algorithm and reports that a version with hand-crafted input features was estimated to play at a strong master level.
Evidence / resource
This page preserves the public LINEAiGE record and its first-party source relationship.
Record identity
LINEAiGE IDtd-gammon-1993