score & fix your ai agent

Download it_

Get the full score.

Fix the gaps.

It's like a report card for your AI agent. Download AMC, open it, and in about a minute you'll see what's strong, what's weak, and exactly what to fix. Then AMC fixes the gaps for you and writes them into your agent. No account. No signup. No terminal. Free forever, bring your own AI key.

No Docker · No terminal · No signup · Free forever

Try it in your browser →
Prefer the terminal?
$curl -fsSL https://agentmaturity.co/install.sh | sh
244checks in the full score
264deep-dive checks
1,171CLI commands
142attack test packs
41industry packs
8,604passing tests

install

Get AMC in under a minute_

Two easy ways: download the app for macOS or Windows, or paste one line into your terminal. Both give you the same full score and one-click fix — every download is a pinned, SHA-256-verified GitHub release.

macOS + Linux

Verified release install

$ curl -fsSL https://agentmaturity.co/install.sh | sh
$ amc

Downloads the pinned platform archive and checks SHA-256 before installing its included AMC package.

copy command →
Windows

PowerShell install

PS> irm https://agentmaturity.co/install.ps1 | iex
PS> amc

Uses the same pinned release and checksum contract, then opens the same CLI and local Studio.

copy command →

desktop studio

Prefer buttons? Run Studio.

AMC includes macOS and Windows launcher apps for Agent Maturity Compass Studio, plus the same local control panel for scores, evidence, reports, assurance packs, Industry Packs, and run history.

macOS app Windows app no account score dashboard signed evidence
amc up
$ open "Agent Maturity Compass Studio.app"
Studiohttp://127.0.0.1:3212
WindowsAgent Maturity Compass Studio.cmd
Run full scorebutton + CLI
Assurance packs142 adversarial packs
Industry Packscommercial checkout not live
Reportsartifact + evidence status
appsLaunch locally

The macOS native WebKit app and Windows system-browser launcher use a version-pinned AMC runtime, without bundling Electron or trusting a stale global CLI.

scoreSee the trust baseline

Overall level, layer bars, evidence coverage, top gaps, and an explicit READY or blocked claim state.

packsBrowse free and paid depth

Assurance is free core. Industry Packs show locked and unlocked states.

reportsShare with the team

Generate outputs for engineering, compliance, and leadership review.

current capabilities

More than a score.
Proof you can operate_

AMC now connects the default score to proof receipts, provider-neutral action observation, runtime Watch signals, fleet topology, policy enforcement, compliance binders, and portable evidence. The first-run Activation path separates signed setup readiness from real runtime completion. These are AMC-owned surfaces, not source-specific wrappers.

activation + proofREADY setup is not COMPLETE

amc connect --status and Studio track connected agent, first observed action, first control decision, and first signed proof. Only verified runtime receipts complete the path.

domain proof laneamc proof check

Source-to-rule checks emit amcproof artifacts and fail closed as unsupported when correctness proof is missing.

watch + enforceObserved actions, signed decisions

Observe Claude Code and Gemini CLI actions by default, or opt into loopback control that reuses AMC policies and binds the exact native response to a signed receipt.

fleetOverview + trust graph

Use fleet overview, graph validation, SLO status, and trust-graph exports before scoring multi-agent systems.

enforce + shieldPolicy and safe proof

Runtime firewall decisions, resource snapshots, safe confirmation proof, and 142 assurance packs stay in the free core.

docs + API1,171 CLI paths

The generated command inventory, API reference, OpenAPI playground, and docs hub expose the same product surface users run locally.

self-reported
100
evidence-verified
16
=
trust gap
84

AMC separates what an agent claims from what execution evidence proves.

0.4xself-reported weight
1.0xobserved evidence weight
2.5xtrust multiplier

the trust stack

Eight named surfaces_
One complete trust platform

The top of the page is intentionally simple. This is the deeper AMC map: score, harden, enforce, prove, monitor, comply, govern fleets, and issue portable trust identity.

platform surfaces

Score trust before you ship

Execution-backed scoring, assurance, enforcement, evidence, monitoring, compliance, fleet management, and portable trust identity.

amc score

integrations

Works with your stack_

15 built-in adapters, one-command Claude Code and Gemini CLI hooks, and portable signed receipts that show exact events, controls, runtime version, effective mode, and lossiness. No second policy engine required.

LangChain

LangChain

CrewAI

CrewAI

Anthropic

Anthropic

Google

Google

Meta

Meta

Claude

Claude

Gemini

Gemini

Python SDK

Python SDK

Hugging Face

Hugging Face

Google Cloud

Google Cloud

Databricks

Databricks

Docker

Docker

PyTorch

PyTorch

GitHub

GitHub

our approach

Expertise built on
evidence_

01

Score and analyze

Point AMC at any agent. The default 244-question full score runs against execution evidence, while the optional 264-question lifecycle set adds runtime, proof, memory, and fleet coverage.

02

Fix and harden

AMC identifies high-signal gaps, generates targeted guardrail work, and uses 142 assurance packs for prompt injection, exfiltration, memory risk, sycophancy, and other adversarial failures.

03

Ship and monitor

Add CI gates, keep signed local evidence, prevent trust regressions, and generate compliance artifacts for EU AI Act, ISO 42001, NIST AI RMF, SOC 2, and OWASP LLM Top 10 review.

research foundation

Built on primary sources_

AMC is grounded in AI safety, security, governance, compliance, and agent-system research. The modules map papers and standards into scoring controls, assurance packs, and operational evidence.

NIST

NIST AI Risk Management Framework

Core governance structure for trustworthy AI deployment and evaluation.

maps to governance, risk, and cross-framework evidencenist.gov →
ICML 2024

Levels of AGI — Operationalizing Progress

Maturity levels framework for AI capabilities and autonomy.

maps to L0-L5 maturity and autonomy scoringarxiv 2311.02462 →
NeurIPS 2024

Connecting the Dots — LLMs Infer Latent Info

Frontier models infer information from distributed evidence.

maps to context leakage and information barriersarxiv 2406.14546 →
memory

Persistent Memory Injection

Cross-session memory poisoning through self-reinforcing payloads.

maps to memory integrity and persistence checksarxiv 2602.15654 →
agents

Monitor Bypass and Agent-as-a-Proxy Risk

Shows why monitoring-only defenses can fail without runtime controls.

maps to monitor bypass resistance and enforcementarxiv 2602.05066 →
security

MCP Security Bench

Evaluation coverage for MCP-specific attack vectors and tool-boundary failures.

maps to MCP compliance and security resilience packsarxiv 2510.15994 →
cost

Economic Denial of Service

Cost amplification through tool-calling and hidden runtime loops.

maps to budget controls and cost predictabilityarxiv 2601.10955 →
alignment

Sycophancy and Objective Drift

Agent behavior can optimize social approval while drifting from ground truth.

maps to sycophancy, alignment, and refusal checksarxiv 2602.08092 →

Agent Maturity Compass whitepaper → / see research notes and gap analysis →

pricing

Free forever. Pro packs for industries_

free / open source

AMC Core

Everything to score and fix your agent: the app, the CLI, the full score, one-click fix, assurance packs, reports, and CI gates. Bring your own AI key.

244 default questions · 264 lifecycle questions · 142 assurance packs · 14 adapters · MIT licensed
start free →
planned · $9.99 / month · per user

Industry Packs — preview

All 41 Industry Domain Packs for regulated verticals. Public checkout is not yet publicly live or available.

Health · Wealth · Education · Mobility · Technology · Governance · Environment
follow launch status →

faq

Common questions_

No. The default amc command generates the full 244-question score. Use amc run --question-set lifecycle for the 264-question lifecycle-expanded set.

No. No account, no API key, no cloud. AMC runs on your machine and your data stays there.

Yes — and it's the point. A fresh agent with no evidence starts at L0. The score grows as AMC watches your agent actually work.

Every report carries a cryptographic signature. Change one character and verification fails. That is why you can hand it to an auditor or a customer — they can check it without trusting you. Try it on a real bundle we published.

Self-reported documentation can claim 100/100 while execution-verified evidence shows a much weaker score. AMC closes that gap with observed evidence, signed artifacts, and repeatable proof chains.

AMC is a trusted observer. Self-reported evidence is capped at 0.4x weight; observed runtime evidence carries 1.0x weight. It scores maturity, finds gaps, runs assurance packs, and preserves verifiable evidence.

15 adapters: AutoGen, Claude Code, CrewAI, Gemini CLI, Generic CLI, Hermes Agent, LangChain for Node and Python, LangGraph, LlamaIndex, OpenAI Agents SDK, OpenClaw, OpenHands, Python AMC SDK, and Semantic Kernel. OpenAI-compatible endpoints use the Generic CLI adapter with the /openai gateway route. Run amc adapters capabilities for signed, current coverage instead of relying on a static badge.

Yes. Use gates such as amc gate --min-score L3 to fail builds below threshold and prevent trust regressions.

No account is needed for the free core. Industry Packs are a planned commercial add-on; checkout is not publicly live.

No. Use the verified macOS/Linux or Windows release installer above. Docker is optional.

Yes. Run amc up and use the local Studio dashboard.

Yes. Run amc fix — it scores your agent, explains the top gaps in plain words, asks once, then writes guardrail rules into your agent's own file (AGENTS.md, CLAUDE.md, .cursorrules, and 12 more). It never touches your API keys, and it leaves a hash-sealed receipt of exactly what changed.

The core trust stack is MIT licensed: Score, Shield, Enforce, Vault, Watch, Comply, Fleet, Passport, adapters, Studio, reports, and CI gates. AMC plans a $9.99/month per user Industry Packs add-on, but public checkout and automatic license issuance are not yet available.

start today

Score your agent with one command_

install amc →