arXiv (Cornell University)· 2016· Preprint
Continuous control with deep reinforcement learning
- 6,787citations
- 2016year
Short summary
A novel actor-critic, model-free algorithm, the Deterministic Policy Gradient (DPG), robustly solves over 20 simulated physics tasks with continuous action spaces, matching planning algorithm performance and learning end-to-end from raw pixels.
AI-generated from the title and abstract; the full text is not read.
TakeawaysIn the app
Key pointsIn the app
Ask the paperIn the app
The rest is in the Pofolia app
Takeaways, key points and questions to the paper; new summaries every day for your field. Free.
Sign in on the web to openField: Artificial Intelligence
Artificial IntelligenceComputer Science