Temporal difference learning and TD-Gammon
No author records on this work.
The source holds an abstract for this work, but its best open-access copy is under no open licence, which does not permit us to republish the text. Read it at the source below.
this paper
works it cites
works citing it
node size = global citations · hover for the full title
What this paper cites, inside the corpus
| Paper | Year | Cited |
|---|---|---|
| Multilayer feedforward networks are universal approximators | 1989 | 21,633 |
| Parallel Distributed Processing | 1986 | 15,411 |
| Some Studies in Machine Learning Using the Game of Checkers | 1959 | 4,416 |
| Learning to Predict by the Methods of Temporal Differences | 1988 | 3,970 |
What cites it, inside the corpus
| Paper | Year | Cited |
|---|---|---|
| Playing Atari with Deep Reinforcement Learning | 2013 | 5,104 |
| Deep Reinforcement Learning: A Brief Survey | 2017 | 4,434 |
| Deep Reinforcement Learning with Double Q-Learning | 2016 | 3,516 |
Links
Topics
| Artificial Intelligence in Games | Computer Science |
| Reinforcement Learning in Robotics | Computer Science |
| Advanced Bandit Algorithms Research | Decision Sciences |
Is this record sound?
partial
One field of this record is missing or disagrees with another. What is shown below is what the source publishes.
- weakensThe source lists no authors for this work at all, so there is nobody to attribute it to and it appears on no author page.
- supports7 reference(s) recorded.
- neutralThe DOI carries no year to check against.
- supportsA title is present.
Provenance
sha256 88a60cdbb94c6a13…