Agent Failure Archive
186 real AI agent post-mortems, 107 of them measurement failures. Free tools, paid via x402.
我該用這個嗎
品質與安全性
發現項目(2)
- LOW在 sample 中
- LOW在 catalog 中
根據工具定義與協定合規性的自動化分析。
上下文成本
這是每次將伺服器的工具載入模型上下文時所消耗的約略 token 數量。數量越高,可用於其他工作的注意力就越少。
安裝
一鍵安裝
將以下內容加入你的 `claude_desktop_config.json` 檔案:
{
"mcpServers": {
"agent-failure-archive": {
"url": "https://desktop-ai2ata5-1.tailfeb765.ts.net/mcp"
}
}
}遠端端點
https://desktop-ai2ata5-1.tailfeb765.ts.net/mcpstreamable-http它能做什麼
工具清單
工具(9)
🟢precheck(claim, evidence)
Free, no wallet. Check a conclusion against nine known ways of fooling yourself and get back which checks it trips plus the question each one asks. Use it before writing 'we found that'. The paid audit tool adds why each matters, the real incident with the numbers measured at the time, and what to run.
輸入結構描述
{
"type": "object",
"properties": {
"claim": {
"title": "Claim",
"type": "string"
},
"evidence": {
"default": "",
"title": "Evidence",
"type": "string"
}
},
"required": [
"claim"
],
"title": "precheckArguments"
}輸出結構描述
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "precheckOutput"
}🟡sample
Free, no wallet. Two complete post-mortems from the archive.
輸入結構描述
{
"type": "object",
"properties": {},
"title": "sampleArguments"
}輸出結構描述
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "sampleOutput"
}⚪catalog
Free, no wallet. What the archive holds, what each paid tool costs, and how payment works.
輸入結構描述
{
"type": "object",
"properties": {},
"title": "catalogArguments"
}輸出結構描述
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "catalogOutput"
}🟢contents(theme)
Free, no wallet. Every case title in the archive, tagged with the trap it illustrates. Titles only, no bodies. Read this to see what the paid archive actually contains before deciding it is worth a dollar.
輸入結構描述
{
"type": "object",
"properties": {
"theme": {
"default": "",
"title": "Theme",
"type": "string"
}
},
"title": "contentsArguments"
}輸出結構描述
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "contentsOutput"
}🟢audit(claim, evidence)
$0.02. Full audit of a claim: why each tripped check matters, the incident behind it with measured numbers, and what to run.
輸入結構描述
{
"type": "object",
"properties": {
"claim": {
"title": "Claim",
"type": "string"
},
"evidence": {
"default": "",
"title": "Evidence",
"type": "string"
}
},
"required": [
"claim"
],
"title": "auditArguments"
}🟡search(q)
$0.01. Three real agent post-mortems matching a symptom, each with root cause, the fix that worked, and the prevention rule.
輸入結構描述
{
"type": "object",
"properties": {
"q": {
"title": "Q",
"type": "string"
}
},
"required": [
"q"
],
"title": "searchArguments"
}⚪brief(action)
$0.05. Pre-flight risk brief before an irreversible action, drawn from the ways that class of action actually failed.
輸入結構描述
{
"type": "object",
"properties": {
"action": {
"title": "Action",
"type": "string"
}
},
"required": [
"action"
],
"title": "briefArguments"
}🟢research(q)
$0.25. The 107 measurement failures from an eight-month attempt to measure one person's individuation with embeddings.
輸入結構描述
{
"type": "object",
"properties": {
"q": {
"default": "",
"title": "Q",
"type": "string"
}
},
"title": "researchArguments"
}⚪archive
$1.00. Every case in one response. One payment, no subscription, no account.
輸入結構描述
{
"type": "object",
"properties": {},
"title": "archiveArguments"
}社群
證據