ai-eval
Cloudflare Workers MCP server: ai-eval
Should I use this
Quality & Safety
B
Findings (2)
- LOWin score_response
- LOWin compare_responses
Based on automated analysis of tool definitions and protocol compliance.
Context Cost
~214Tokens (tool definitions)
~567 BTypical response size
Minimal attention impact (0.17% of 128k context)
This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.
Install
One-Click Install
Add this to your `claude_desktop_config.json` file:
{
"mcpServers": {
"ai-eval": {
"url": "https://api.lazy-mac.com/ai-eval/mcp"
}
}
}Remote endpoints
https://api.lazy-mac.com/ai-eval/mcpstreamable-httpWhat it can do
Tool inventory
Tools (3)
🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
⚪score_response(prompt, response, criteria)
Score an AI response against a prompt using heuristic metrics (length, relevance, structure, completeness)
Input Schema
{
"type": "object",
"properties": {
"prompt": {
"type": "string",
"description": "The original prompt/question"
},
"response": {
"type": "string",
"description": "The AI response to evaluate"
},
"criteria": {
"type": "array",
"items": {
"type": "string"
},
"description": "Optional keywords that should appear in response"
}
},
"required": [
"prompt",
"response"
]
}⚪compare_responses(prompt, responses)
Compare and rank multiple AI responses to the same prompt
Input Schema
{
"type": "object",
"properties": {
"prompt": {
"type": "string"
},
"responses": {
"type": "array",
"minItems": 2,
"items": {
"type": "string"
}
}
},
"required": [
"prompt",
"responses"
]
}🟢text_metrics(text)
Get text quality metrics: word count, sentence count, estimated tokens, readability grade
Input Schema
{
"type": "object",
"properties": {
"text": {
"type": "string"
}
},
"required": [
"text"
]
}Community
Evidence
Recent observations
verifiedversion not recorded3 tools
verifiedversion not recorded3 tools