Agent57
Agent57 is a model-free distributed reinforcement learning algorithm developed by Google DeepMind and reported in 2020.
Explore Reinforcement Learning through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Reinforcement Learning.
Showing 1-5 of 5 articles
Agent57 is a model-free distributed reinforcement learning algorithm developed by Google DeepMind and reported in 2020.
Intrinsic Discovery is an experimental training method introduced by Induction Labs in August 2026.
RAGEN-2 is a 2026 research paper and public code extension for diagnosing and mitigating reasoning collapse during reinforcement learning of multi-turn large language model agents.
SPADE, short for Self-Play in Adaptive Synthetic Executable Environments, is a reinforcement-learning framework in which one language model alternates between designing executable training environments and…
Tim Rocktäschel is a German computer scientist known for his work on reinforcement learning, open-ended learning, and language-based AI agents.