Agent Failure Archive

186 real AI agent post-mortems, 107 of them measurement failures. Free tools, paid via x402.

Should I use this

Quality & Safety

B
Description quality
82%
Schema completeness
61%
Naming quality
84%
Poisoning risk
100%
Permission match
100%
Protocol compliance
100%

Findings (2)

  • LOWTool 'sample' description lacks action verbin sample
  • LOWTool 'catalog' description lacks action verbin catalog

Based on automated analysis of tool definitions and protocol compliance.

Context Cost

~742Tokens (tool definitions)
~496 BTypical response size
Moderate attention impact (0.58% of 128k context)

This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.

Install

One-Click Install

Add this to your `claude_desktop_config.json` file:

{
  "mcpServers": {
    "agent-failure-archive": {
      "url": "https://desktop-ai2ata5-1.tailfeb765.ts.net/mcp"
    }
  }
}

Remote endpoints

https://desktop-ai2ata5-1.tailfeb765.ts.net/mcpstreamable-http

What it can do

Tool inventory

Tools (9)

🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
🟢precheck(claim, evidence)

Free, no wallet. Check a conclusion against nine known ways of fooling yourself and get back which checks it trips plus the question each one asks. Use it before writing 'we found that'. The paid audit tool adds why each matters, the real incident with the numbers measured at the time, and what to run.

Input Schema

{
  "type": "object",
  "properties": {
    "claim": {
      "title": "Claim",
      "type": "string"
    },
    "evidence": {
      "default": "",
      "title": "Evidence",
      "type": "string"
    }
  },
  "required": [
    "claim"
  ],
  "title": "precheckArguments"
}

Output Schema

{
  "type": "object",
  "properties": {
    "result": {
      "title": "Result",
      "type": "string"
    }
  },
  "required": [
    "result"
  ],
  "title": "precheckOutput"
}
🟡sample

Free, no wallet. Two complete post-mortems from the archive.

Input Schema

{
  "type": "object",
  "properties": {},
  "title": "sampleArguments"
}

Output Schema

{
  "type": "object",
  "properties": {
    "result": {
      "title": "Result",
      "type": "string"
    }
  },
  "required": [
    "result"
  ],
  "title": "sampleOutput"
}
⚪catalog

Free, no wallet. What the archive holds, what each paid tool costs, and how payment works.

Input Schema

{
  "type": "object",
  "properties": {},
  "title": "catalogArguments"
}

Output Schema

{
  "type": "object",
  "properties": {
    "result": {
      "title": "Result",
      "type": "string"
    }
  },
  "required": [
    "result"
  ],
  "title": "catalogOutput"
}
🟢contents(theme)

Free, no wallet. Every case title in the archive, tagged with the trap it illustrates. Titles only, no bodies. Read this to see what the paid archive actually contains before deciding it is worth a dollar.

Input Schema

{
  "type": "object",
  "properties": {
    "theme": {
      "default": "",
      "title": "Theme",
      "type": "string"
    }
  },
  "title": "contentsArguments"
}

Output Schema

{
  "type": "object",
  "properties": {
    "result": {
      "title": "Result",
      "type": "string"
    }
  },
  "required": [
    "result"
  ],
  "title": "contentsOutput"
}
🟢audit(claim, evidence)

$0.02. Full audit of a claim: why each tripped check matters, the incident behind it with measured numbers, and what to run.

Input Schema

{
  "type": "object",
  "properties": {
    "claim": {
      "title": "Claim",
      "type": "string"
    },
    "evidence": {
      "default": "",
      "title": "Evidence",
      "type": "string"
    }
  },
  "required": [
    "claim"
  ],
  "title": "auditArguments"
}
🟡search(q)

$0.01. Three real agent post-mortems matching a symptom, each with root cause, the fix that worked, and the prevention rule.

Input Schema

{
  "type": "object",
  "properties": {
    "q": {
      "title": "Q",
      "type": "string"
    }
  },
  "required": [
    "q"
  ],
  "title": "searchArguments"
}
⚪brief(action)

$0.05. Pre-flight risk brief before an irreversible action, drawn from the ways that class of action actually failed.

Input Schema

{
  "type": "object",
  "properties": {
    "action": {
      "title": "Action",
      "type": "string"
    }
  },
  "required": [
    "action"
  ],
  "title": "briefArguments"
}
🟢research(q)

$0.25. The 107 measurement failures from an eight-month attempt to measure one person's individuation with embeddings.

Input Schema

{
  "type": "object",
  "properties": {
    "q": {
      "default": "",
      "title": "Q",
      "type": "string"
    }
  },
  "title": "researchArguments"
}
⚪archive

$1.00. Every case in one response. One payment, no subscription, no account.

Input Schema

{
  "type": "object",
  "properties": {},
  "title": "archiveArguments"
}

Community

Rate this Server

Evidence

Recent observations

verifiedversion not recorded9 tools
verifiedversion not recorded9 tools