Richard S. Sutton
Richard S. Sutton
collaborator
line weight = collaboration weight
Works in this corpus
| Title | Position | Year | Cited |
|---|---|---|---|
| Reinforcement Learning: An Introduction | first | 1998 | 27,409 |
| Introduction to Reinforcement Learning | first | 1998 | 6,917 |
| Policy Gradient Methods for Reinforcement Learning with Function Approximation | first | 1999 | 4,966 |
| Learning to Predict by the Methods of Temporal Differences | first | 1988 | 3,970 |
| Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning | first | 1999 | 3,201 |
| Learning to predict by the methods of temporal differences | first | 1988 | 2,818 |
| Integrated Architectures for Learning, Planning, and Reacting Based on Approximating Dynamic Programming | first | 1990 | 1,375 |
Is this one person?
This record carries a person-claimed ORCID and no signal that it mixes two people.
- supportsAn ORCID is claimed by a person, not inferred, so it is the only identity assertion here that a human made.
- supports2 distinct name form(s) across this row's works, counting a spelled-out given name and its initial as one form.
- supportsAt most 1 distinct institution(s) inside any five-year window, which is a normal career.
- supports84% of this row's works sit in its single largest field.
- neutralConfidence is judged on the 7 work(s) this corpus holds, not on the author's whole output. A single-work row carries little evidence either way.