Agents tagged "grpo"

9 results
grpo ×
Data & Analysis

Relax Omni-Modal RL Engine

An asynchronous distributed reinforcement-learning engine for post-training text, vision, and audio models.

★ 621 FS 64 Some gaps 6d ago Apache-2.0
Dev & Engineering

HUD

A platform for building RL environments and evals for AI agents — define an environment once, then evaluate and train any model on it.

★ 302 FS 63 Some gaps 4d ago MIT
Dev & Engineering

MARTI: Multi-Agent Reinforced Training & Inference for LLMs

Train LLM-based multi-agent systems with reinforcement learning and tree-search-augmented reasoning across debate, chain-of-agents, and mixture-of-agents workflows.

★ 560 FS 36 Major gaps 1mo ago MIT
Data & Analysis

ART Agent Reinforcement Trainer

Train multi-step agents from real task experience with GRPO.

★ 11k FS 34 Major gaps today Apache-2.0
Dev & Engineering

Hands-On Modern RL

A practice-first reinforcement learning course from CartPole to LLM post-training and agentic systems.

★ 4.4k FS 30 Major gaps 20d ago NOASSERTION
Dev & Engineering

AgentsMeetRL — Awesome List of Agentic Reinforcement Learning

A curated, categorized collection of open-source RL-training projects for LLM agents, with technical details and an interactive dashboard.

★ 1.8k FS 29 Major gaps 8d ago
Data & Analysis

Open-AgentRL

RL training framework for reasoning, tool use, and interactive agent environments.

★ 641 FS 28 Major gaps 3mo ago Apache-2.0
Data & Analysis

AReaL Async RL Platform

Asynchronous RL infrastructure for training reasoning models and tool-using agent workflows.

★ 5.8k FS 0 Major gaps today Apache-2.0
Dev & Engineering

rLLM: Reinforcement Learning Framework for LLM Agents

A unified framework for RL training of language agents, supporting any harness, any sandbox, and one-flag backend switching.

★ 5.8k FS 0 Major gaps 11d ago Apache-2.0