Dev & Engineering AI Agents
Codex Autoresearch
Run measurable optimization loops inside your codebase: give it a goal, a benchmark, and a boundary, and it experiments, keeps evidence, and leaves a reviewable patch.
Zoetrope
Watch a Claude Code session as a live flow graph — main agent, subagents and tool calls — in your terminal or browser.
elizaOS
An open-source TypeScript system for building, running, and extending autonomous AI agents.
Unlazy Completion Discipline
Keep AI agents from finishing early with deep decomposition, acceptance ledgers, and executable gates.
Pydantic AI Skills
Agent Skills support for Pydantic AI with progressive disclosure, remote skill registries, and sandboxed execution of bundled scripts — so your skill library can grow without bloating the prompt.
Better Harness
Diagnose and improve coding-agent workflows with verifiable evidence.
JS Reverse MCP
Continuous browser-side JavaScript debugging and replay for AI coding assistants.
Apache Maka
A local-first agent workspace that performs project work while preserving recoverable execution records.
Kungfu
Keep one body of work moving across agents, sessions, and failures.
NotebookLM Python Automation
Automate Gemini Notebook research, generation, and exports through Python, CLI, or coding agents.
ADK Agent Recipes
Start ADK agent projects from small, runnable reference implementations.
Pilotfish Multi-Model Orchestration
Frontier models plan, cheaper models execute, and fresh-context verification guards quality — a one-prompt-install orchestration policy for Claude Code.
Osaurus
A native Mac harness for persistent, tool-using AI agents that can run with local or cloud models.
git-lrc Commit-Time Code Review
Automatically review Git diffs and surface risks before each commit lands.
Stately Agent
State-machine-powered LLM agents with XState: the machine owns control flow, the model only picks legal events, so invalid agent actions are impossible by construction.
Chrome DevTools MCP
Give coding agents direct Chrome automation, debugging, and performance analysis.
Harmonist
An IDE-hooked multi-agent workflow that enforces review, memory, and integrity checks for AI-assisted coding.
Bernstein
Deterministically orchestrate parallel CLI coding agents with offline-verifiable run records.
LangWatch
Evaluate, test, trace, and monitor LLM applications and AI agents across development and production.
Gortex Code Intelligence
A local code graph that gives coding assistants precise context and cross-repository impact analysis.
Oh My Hermes (OMH)
A professional operating layer for Hermes Agent: per-model routing, parallel coding, 108 specialist skills, and reviewable long-term memory — install once, keep Hermes.
Surf CLI
A CLI that lets AI agents control Chrome — zero config, agent-agnostic, and battle-tested, so any agent that can run shell commands can drive the browser.
OpenHands Agent Canvas
A self-hosted control center for coding agents — run OpenHands, Claude Code, Codex, or any ACP agent
TanStack AI
Build type-safe streaming, tool-calling, and multimodal AI applications without committing the application core to one model provider.