OrcaReplay
Time travel for AI coding agents: record, replay offline, and fork any run from any checkpoint onto different models — byte-for-byte, to find out exactly why it failed.
AgentSight
eBPF-based system-level observability for AI agents — see what agents actually do on your machine with zero SDK, proxy, or vendor integration.
Archestra
A unified control plane for enterprise model access, MCP tools, agent execution, security, and observability.
Cohort
A local-first, evidence-driven Agent Runtime that records verifiable execution traces, forks historical runs for counterfactual experiments, and governs context, cost, and tool risk.
Google Cloud Agent Starter Pack
Scaffold, evaluate, and deploy production-oriented GenAI agents on Google Cloud with ready-made templates and infrastructure.
Pydantic AI
Build composable generative-AI applications in Python with type hints and Pydantic validation.
NVIDIA NeMo Agent Toolkit
Adds intelligence to AI agents across any framework, enhancing speed, accuracy, and decision-making through enterprise-grade instrumentation, observability, and continuous learning.
Runtm
Open-source sandboxes where coding agents build and deploy: run Claude Code, Cursor, and other agents in isolated environments with live URLs, logs, and previews.
Agentspan
A durable runtime that keeps AI agents running through crashes, long waits, and human approvals.
Arize Phoenix
Open-source AI observability platform for tracing, evaluating, and troubleshooting LLM applications.
Adrian
Open-source runtime security engine for AI agents: it watches actions and reasoning in real time and intercepts malicious tool use, prompt injection and policy drift before the agent acts.
Trigger.dev
Build, deploy, and observe durable AI agents and background workflows.
HolmesGPT SRE Agent
An SRE agent for investigating production incidents and finding root causes.
CAPA
Declare skills, tools, rules, sub-agents, and MCP servers once in capabilities.yaml — CAPA syncs them to 35+ AI coding agents and runs a local MCP gateway, ending scattered agent config.
Google Agents CLI
CLI commands and coding-agent skills for building, evaluating, and deploying ADK agents on Google Cloud.
Claude DevTools
The missing debugging tool for Claude Code — visualize hidden session logs, tool calls, token usage, and context window in a clean UI.
Agents Observe
Real-time observability dashboard for Claude Code and multi-agent sessions, making every subagent, tool call, and token cost visible as it happens.
Ongrid
An ops AI Agent that understands your infrastructure, finds the root cause, and fixes it — right from Slack, Telegram, Lark or DingTalk.
Mastra
A TypeScript framework for building, orchestrating, evaluating, and deploying AI applications and agents.
Generative AI with LangChain
Learn to build production-oriented LLM applications and multi-agent systems with Python, LangChain, and LangGraph.
Claude Code Router (CCR)
One local control plane for every AI agent: route, fail over, extend, and observe from a single app.
AgentOps
Observability, session replay, and cost tracking for AI-agent applications.
VibePod — Containerized AI Coding Agent Runner
One CLI that runs Claude, Codex, Copilot and other coding agents inside isolated Docker or Podman containers, with local metrics and traffic analytics.
Langflow
Visual platform for building and deploying AI agents and workflows