AgileRL
Streamlines reinforcement learning with RLOps: evolutionary hyperparameter optimization removes exhaustive tuning runs for up to 10x faster training.
Relax Omni-Modal RL Engine
An asynchronous distributed reinforcement-learning engine for post-training text, vision, and audio models.
HUD
A platform for building RL environments and evals for AI agents — define an environment once, then evaluate and train any model on it.
MetaClaw
A local proxy that turns personal-agent conversations into skills, memory, and optional scheduled LoRA learning.
MobileGym
A programmable mobile simulator for deterministic GUI-agent evaluation and online RL training.
Agentic RL: The Most Detailed Intro
A hands-on learning hub for Agentic RL — foundations, runnable code, an experiment-visualization dashboard, and readings of frontier base-model papers.
Tongyi DeepResearch
An open model and inference system for long-horizon web research, evidence gathering, and research question answering.
AWorld Agent Harness
Turn domain expertise into multi-agent workflows that build, evaluate, and iteratively improve concrete outputs.
MARO Resource Optimization Platform
Simulate and optimize real-world resource decisions with reinforcement learning and distributed components.
Memora Agent Memory
A structured memory layer that stores rich agent history and retrieves it through abstractions and semantic cues.
IR-SIM: Lightweight Robot Simulator
A Python-based lightweight simulator for navigation, control, and learning, configured via simple YAML files.
skrl: Modular Reinforcement Learning Library
A highly modular and readable RL library implemented in PyTorch, JAX, and NVIDIA Warp.
gym-pybullet-drones: Quadcopter RL Environment
PyBullet-based Gymnasium environments for single and multi-agent reinforcement learning of quadcopter control.
Agent Lightning
Capture agent trajectories and train improved prompts or policy resources for existing AI agents.
MARTI: Multi-Agent Reinforced Training & Inference for LLMs
Train LLM-based multi-agent systems with reinforcement learning and tree-search-augmented reasoning across debate, chain-of-agents, and mixture-of-agents workflows.
RLinf RL Infrastructure
Scalable reinforcement-learning infrastructure for embodied robotics and agentic AI training.
ART Agent Reinforcement Trainer
Train multi-step agents from real task experience with GRPO.
AI Engineering from Scratch
A free, open-source curriculum of 503 lessons across 20 phases, teaching you to build AI end-to-end from math to agents.
ChatArena
A Python framework for building and running multi-agent language-game experiments with LLM players.
veRL Agent Training
Multi-turn reinforcement learning training for long-horizon LLM and vision-language agents.
AgentFlow
Train an in-system planner with Flow-GRPO for more reliable multi-tool, long-horizon reasoning.
Hands-On Modern RL
A practice-first reinforcement learning course from CartPole to LLM post-training and agentic systems.
Mobile-Agent: A Family of Powerful GUI Agents
A multi-platform GUI agent family by Alibaba's Tongyi Lab, automating mobile, desktop, and web tasks via multimodal models.
AgentsMeetRL — Awesome List of Agentic Reinforcement Learning
A curated, categorized collection of open-source RL-training projects for LLM agents, with technical details and an interactive dashboard.