Actor-critic algorithms
Vijay R. Konda low, John N. Tsitsiklis
The source holds an abstract for this work, but its best open-access copy is under cc-by-nc, which does not permit us to republish the text. Read it at the source below.
this paper
works it cites
works citing it
node size = global citations · hover for the full title
What this paper cites, inside the corpus
What cites it, inside the corpus
| Paper | Year | Cited |
|---|---|---|
| Policy Gradient Methods for Reinforcement Learning with Function Approximation | 1999 | 4,966 |
| Deep Reinforcement Learning: A Brief Survey | 2017 | 4,434 |
| Counterfactual Multi-Agent Policy Gradients | 2018 | 1,727 |
Links
Topics
| Reinforcement Learning in Robotics | Computer Science |
| Advanced Control Systems Optimization | Engineering |
| Adaptive Dynamic Programming Control | Computer Science |
Is this record sound?
complete
Nothing in this record contradicts itself and no field we check is missing.
- supports2 author record(s) attached.
- supports17 reference(s) recorded.
- neutralThe DOI carries no year to check against.
- supportsA title is present.
Provenance
sha256 db1645b78a57e29a…