mu Coding Agent

Offload routine session decisions to a fast judgment kernel so the main model can focus on coding.

Source repo
qybaihe/mu
Stars
★ 476
Last updated
1d ago
License
MIT
Primary language
TypeScript

At a glance

How it runs
Desktop appCLI
Works with
Universal · cross-platformChatGPT · Codex · Claude Code · Claude.ai · OpenAI API · Claude API (Partial support)
Cost
Free software; you pay for model usage
Setup effort
Medium · a few setup steps
You'll need
Node.js 22.19 or newerShell / CLINetwork accessLocal filesystemMCP Server
Typical use
Developers who want an interactive terminal coding agent to work in the current project and resume a previous session.
Not a fit if
  • Teams that require every judgment to stay on-device
  • Users who need a stable release rather than pre-release software
Source review
77/100 · Good

What does this agent do, and when should you use it?

mu is a coding agent built on pi, available as a command-line tool and as a desktop app with its runtime included. It places bounded judgment questions at points covering input, context selection, tool safety, progress, and multi-agent coordination. A small judge handles those calls while the main model reads the project, uses tools, and does the coding work; verdicts can change what enters context or whether an operation proceeds. The desktop app adds panels for the plain-language board, judgments, hive, files, preview, source, and browser. The project describes itself as early development, with 0.1.x pre-releases.

Run mu in a project directory for an interactive coding session, use mu -p for a one-shot prompt, or mu -c to continue the last session. mu can admit tool output chunk by chunk, fold exact repeats in test logs, screen web and MCP output for prompt injection, and route questions about command risk, constraints, approvals, task drift, and completion to configured judges. It uses model providers supported by pi and can route decision points to Jev, Laya, classifier models, or LLMs. The agent produces code changes and tool results while recording verdicts in a local ledger; hive sub-agents inspect code and run commands but do not edit, leaving changes to the main model.

  1. Developers who want an interactive terminal coding agent to work in the current project and resume a previous session.
  2. Engineers who need to reduce repetitive test output or select log sections according to the current debugging goal.
  3. Teams that want explicit checks around risky commands, stated constraints, and tool approvals.
  4. Maintainers who want read-only sub-agents to investigate different areas before the main agent makes changes.
  5. Developers who prefer a desktop workspace for reviewing judgments, progress, files, and browser actions.

How do you install or deploy this agent?

For the desktop app, download an installer for macOS, Windows, or Linux from GitHub Releases; the README says no separate Node installation is needed. The CLI requires Node.js 22.19 or newer:

How do you use this agent?

Install the CLI, connect a model, then start a session in the project directory:

npm i -g mu-agent
mu setup
mu

A judge key is not required to get started: the README says the default is the free Jev on OpenCode Zen. Configure another judge's credentials with mu setup or in the desktop app's judges page. Check connections or run a one-shot prompt with:

mu doctor
mu -p "Fix an issue in this project"

What are this agent's strengths and limitations?

Pros
  • Its judgment kernel separates context admission, tool risk, constraints, task drift, and completion into configurable decision points, including shadow mode.
  • Exact repeated test-log content can be folded without a model call; the README reports a 51% character reduction across its sample logs, with markers that expand losslessly.
  • Judgment options include hosted Jev, the 322M-parameter local Laya, classifier models, and general LLMs.
  • The desktop app bundles the mu runtime and includes judgment ledger, hive, progress board, and browser panels.
Limitations
  • The project is explicitly early development; its 0.1.x npm and desktop releases are pre-releases and names, settings, and formats may change.
  • Unless using Laya or disabling hosted judges, judgment requests may go to the selected service; the default free Jev sends required question fields to OpenCode Zen.
  • The CLI requires Node.js 22.19 or newer, and account setup differs between CLI and desktop.
  • The README's test-log measurements use a limited author-collected sample and report no end-to-end comparison with pi, Claude Code, or Codex on the same tasks.

How does this agent compare with similar options?

mu extends pi with a judgment kernel. The README says there is not yet an end-to-end comparison against pi, Claude Code, or Codex on the same tasks. mu can also import and continue Claude Code and Codex CLI conversations.

Key facts side by side with the most closely related agents.

Agent Source review Form / cost Stars Updated Language Full support on
mu Coding Agent This agent 77 · Good Desktop appFree + model costs ★ 476 1d ago TypeScript —
Understanding AI Agents: Design Principles and Engineering Practice 41 · Major gaps CLIFree + model costs ★ 53k 9d ago Python —
How Claude Code Works 0 · Major gaps Web appFree ★ 3.7k 1mo ago — Claude Code
senpi 74 · Some gaps CLIFree + model costs ★ 473 today TypeScript OpenAI API · Claude API

How does FollowAgents rate this agent?

FollowAgents source review · FARS-2.1
Good
77/ 100 5-point scale 3.9 / 5
Trust 22/29
Reliability 8/14
Adaptability 16/18
Convention 14/18
Effectiveness 10/13
Verifiability 7/8
Why each dimension lost points
Trust22 / 29 · 3.8/5

SECURITY.md says mu runs with the current user's permissions and is not a sandbox, recommending a container or VM when isolation is needed. Risk rules, constraint checks, approval for project MCP servers, and permission modes provide meaningful safeguards, but the evidence does not establish that they prevent every mistaken action, so this is not full marks. The documentation specifies where conversations and judge data go, describes credential masking, and says keys are not sent to another judge service. External commands and other effects are governed by permission modes, with checkpoints and rewind commands documented. Dependency pinning and CI checks are visible, but the supplied material shows no dependency audit or vulnerability mitigation record. pi and AionUi are named as upstream projects, and the repository scope and maintenance channel are clear; publisher identity is unverified and the team is small, so provenance is not independently established.

Reliability8 / 14 · 2.9/5

The README explains decision points, modes, and commands, but the root package.json identifies a pi-monorepo at version 0.0.3 while the README describes mu as a 0.1.x prerelease; the desktop app also depends on AionUi and its relationship to the repository-level manifest is not fully clear in the supplied evidence. CI lists build, type-check, and test workflows, but runtime decisions can depend on free OpenCode Zen or other external model services, and the local judge has stated limits. SECURITY.md provides a private vulnerability-report route and response expectations, while CI defines timeouts and failure artifacts; concrete user-facing error behavior is only partially evidenced.

Adaptability16 / 18 · 4.4/5

The documentation gives clear uses for interactive coding, context filtering, test logs, tool safety, and multi-agent work. Permission modes, judge choices, active/shadow/off decision points, sandbox boundaries, and judge fallibility are described, so capability limits and trigger scope are well covered. The CLI specifies Node.js 22.19 or newer and the desktop targets several operating systems and architectures; resource needs, environment differences, and offline behavior are not fully documented in the supplied material.

Convention14 / 18 · 3.9/5

The README has a contents list, organized tables, commands, configuration guidance, and links to reference documents. It clearly states installation steps, prerelease status, measurement scope, and limitations. Decision-point and product naming is structured, but the version evidence conflicts (README 0.1.x versus root manifest 0.0.3), and the README warns that names, settings, and formats may change, reducing naming stability. The LICENSE contains MIT text, but its copyright names Mario Zechner and the supplied materials do not explain how that relates to the repository or the given MIT metadata. SECURITY.md defines supported versions, maintenance responsibility, and a reporting path; the supplied materials do not show an actual change-history entry.

Effectiveness10 / 13 · 3.8/5

The product offers a readable status board, judge ledger, commands, and per-decision-point explanations that make outputs inspectable and actionable. The README reports quantified log-folding and goal-aware selection results, with methodology and limitations, supporting potential value; the evidence remains primarily project-authored and lacks an end-to-end comparison with comparable agents. Latency, cost, and context savings have concrete figures, but judge services add external dependency and possible cost, and independent comparative evidence is absent.

Verifiability7 / 8 · 4.4/5

The README connects major performance claims to an in-repository replay script and methodology document, and states the dataset, measurement date, label source, and uncovered cases. SECURITY.md and CI configuration provide additional evidence about boundaries and automated workflows. The supplied material includes a no-op placeholder test and workflow definitions, not proof that every claimed test or behavior is effective, so cross-source corroboration is moderate. The documentation distinguishes author measurements, scope, limitations, and unverified claims clearly.

Risks and how to mitigate them
  • mu is not a sandbox and may run commands or change files with the current user's permissions; use a container or VM when isolation is required.
  • Without a judge key, decision data goes to OpenCode Zen; external model use sends conversation data to the selected provider. Confirm those services and data flows meet your requirements.
  • The README and root package.json give conflicting version signals, and the license text's copyright attribution is unexplained in the supplied material; verify the intended release package, version, and license attribution before adopting it in a formal workflow.
Evidence confidence: Low Reviewed Oct 09, 2026 Reviewed revision 537d028f483e
See the full review method →

FAQ

Do I need a judge API key?
No. The README says mu uses the free Jev on OpenCode Zen when no judge key is configured; you can also configure another service or use local Laya.
Where do judgment requests go?
That depends on the judge. The default Jev sends the fields needed for a question to OpenCode Zen. With Laya or MU_JUDGE=off, judgments stay on the machine.
Will the agent run every command automatically?
No. Rules catch commands that look dangerous first, and risk, approval, and constraint decision points can ask the user when the request is unclear.
Can hive sub-agents edit my project?
The README says bees can read code and run commands, but never edit; the main model makes changes.
Is mu ready for teams that need stable interfaces?
Evaluate carefully: the project is in early development, and pre-release names, settings, and formats may still change.
View on GitHub ↗ Install ↓

Compare agents like this one

The same FARS review applied across the shortlist this agent qualifies for.

Related agents