LINEAiGE
CANONICAL HISTORY

TD-Gammon — AAAI Fall Symposium Publication Milestone

Gerald Tesauro published a TD-Gammon paper for AAAI-FS 1993 on 22 October 1993. IBM Research describes TD-Gammon as a neural network that learned to play backgammon through self-play using the TD(lambda) reinforcement-learning algorithm and reports that a version with hand-crafted input features was estimated to play at a strong master level.

22 October 1993CANONICAL HISTORY

Evidence / resource

This page preserves the public LINEAiGE record and its first-party source relationship.

Record identity

LINEAiGE ID
td-gammon-1993