Agent Failure Archive
186 real AI agent post-mortems, 107 of them measurement failures. Free tools, paid via x402.
Should I use this
Quality & Safety
Findings (2)
- LOWin sample
- LOWin catalog
Based on automated analysis of tool definitions and protocol compliance.
Context Cost
This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.
Install
One-Click Install
Add this to your `claude_desktop_config.json` file:
{
"mcpServers": {
"agent-failure-archive": {
"url": "https://desktop-ai2ata5-1.tailfeb765.ts.net/mcp"
}
}
}Remote endpoints
https://desktop-ai2ata5-1.tailfeb765.ts.net/mcpstreamable-httpWhat it can do
Tool inventory
Tools (9)
🟢precheck(claim, evidence)
Free, no wallet. Check a conclusion against nine known ways of fooling yourself and get back which checks it trips plus the question each one asks. Use it before writing 'we found that'. The paid audit tool adds why each matters, the real incident with the numbers measured at the time, and what to run.
Input Schema
{
"type": "object",
"properties": {
"claim": {
"title": "Claim",
"type": "string"
},
"evidence": {
"default": "",
"title": "Evidence",
"type": "string"
}
},
"required": [
"claim"
],
"title": "precheckArguments"
}Output Schema
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "precheckOutput"
}🟡sample
Free, no wallet. Two complete post-mortems from the archive.
Input Schema
{
"type": "object",
"properties": {},
"title": "sampleArguments"
}Output Schema
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "sampleOutput"
}⚪catalog
Free, no wallet. What the archive holds, what each paid tool costs, and how payment works.
Input Schema
{
"type": "object",
"properties": {},
"title": "catalogArguments"
}Output Schema
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "catalogOutput"
}🟢contents(theme)
Free, no wallet. Every case title in the archive, tagged with the trap it illustrates. Titles only, no bodies. Read this to see what the paid archive actually contains before deciding it is worth a dollar.
Input Schema
{
"type": "object",
"properties": {
"theme": {
"default": "",
"title": "Theme",
"type": "string"
}
},
"title": "contentsArguments"
}Output Schema
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "contentsOutput"
}🟢audit(claim, evidence)
$0.02. Full audit of a claim: why each tripped check matters, the incident behind it with measured numbers, and what to run.
Input Schema
{
"type": "object",
"properties": {
"claim": {
"title": "Claim",
"type": "string"
},
"evidence": {
"default": "",
"title": "Evidence",
"type": "string"
}
},
"required": [
"claim"
],
"title": "auditArguments"
}🟡search(q)
$0.01. Three real agent post-mortems matching a symptom, each with root cause, the fix that worked, and the prevention rule.
Input Schema
{
"type": "object",
"properties": {
"q": {
"title": "Q",
"type": "string"
}
},
"required": [
"q"
],
"title": "searchArguments"
}⚪brief(action)
$0.05. Pre-flight risk brief before an irreversible action, drawn from the ways that class of action actually failed.
Input Schema
{
"type": "object",
"properties": {
"action": {
"title": "Action",
"type": "string"
}
},
"required": [
"action"
],
"title": "briefArguments"
}🟢research(q)
$0.25. The 107 measurement failures from an eight-month attempt to measure one person's individuation with embeddings.
Input Schema
{
"type": "object",
"properties": {
"q": {
"default": "",
"title": "Q",
"type": "string"
}
},
"title": "researchArguments"
}⚪archive
$1.00. Every case in one response. One payment, no subscription, no account.
Input Schema
{
"type": "object",
"properties": {},
"title": "archiveArguments"
}Community
Evidence