HuggingFace TRL
TRL (Transformer Reinforcement Learning, now stylized as Transformers Reinforcement Learning) is an open-source Python library maintained by Hugging Face for post-training large language models with…
Explore Reinforcement Learning through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Reinforcement Learning.
Showing 1-6 of 6 articles
TRL (Transformer Reinforcement Learning, now stylized as Transformers Reinforcement Learning) is an open-source Python library maintained by Hugging Face for post-training large language models with…
Microduck is a small biped robot developed by Pollen Robotics, the robotics team within Hugging Face.
MuJoCo (short for Multi-Joint dynamics with Contact) is an open-source physics simulator designed for fast and accurate simulation of articulated mechanical systems with rich contact interactions.
OpenAI Baselines is a collection of open-source, high-quality reference implementations of reinforcement learning (RL) algorithms released by OpenAI.
RAGEN-2 is a 2026 research paper and public code extension for diagnosing and mitigating reasoning collapse during reinforcement learning of multi-turn large language model agents.
Tülu 3 is a fully open post-training recipe and a corresponding family of instruction-tuned language models released by the Allen Institute for AI (Ai2) on November 21, 2024.