ModelsAgree

Consensus 'best X for Y' rankings from ChatGPT, Claude, Gemini & Grok. Search, verdicts, history.

Should I use this

Quality & Safety

A
Description quality
92%
Schema completeness
86%
Naming quality
100%
Poisoning risk
100%
Permission match
100%
Protocol compliance
100%

Based on automated analysis of tool definitions and protocol compliance.

Context Cost

~542Tokens (tool definitions)
~513 BTypical response size
Minimal attention impact (0.42% of 128k context)

This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.

Install

One-Click Install

Add this to your `claude_desktop_config.json` file:

{
  "mcpServers": {
    "modelsagree": {
      "url": "https://modelsagree.com/mcp"
    }
  }
}

Remote endpoints

https://modelsagree.com/mcpstreamable-http

What it can do

Tool inventory

Tools (5)

🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
🟢search_best(query)

Find what AI models agree is the best product/tool/service for a need, or look up how a specific brand ranks. Returns matching categories and brands with a dated one-sentence verdict and source URLs. Use for any 'best X for Y' question. Excludes medical, financial, and legal advice.

Input Schema

{
  "type": "object",
  "properties": {
    "query": {
      "type": "string",
      "description": "A category, use-case, or brand name — e.g. 'best llm observability', 'ci/cd for cloud native', or 'Langfuse'."
    }
  },
  "required": [
    "query"
  ]
}
🟢get_best_in_category(slug)

Get the full ranked leaderboard and verdict for a category slug (obtained from search_best or list_categories), e.g. 'best-llm-observability'.

Input Schema

{
  "type": "object",
  "properties": {
    "slug": {
      "type": "string",
      "description": "Category slug, e.g. best-llm-observability"
    }
  },
  "required": [
    "slug"
  ]
}
🟢get_product(slug)

Get a single brand/product's record across every category it is ranked in, with verdicts and its homepage.

Input Schema

{
  "type": "object",
  "properties": {
    "slug": {
      "type": "string",
      "description": "Product slug, e.g. langfuse"
    }
  },
  "required": [
    "slug"
  ]
}
🟢get_poll_history(slug, model, limit)

Get the raw poll history — every model's pick over time — behind a category's verdict. The credibility/audit trail. Optionally filter by model.

Input Schema

{
  "type": "object",
  "properties": {
    "slug": {
      "type": "string",
      "description": "Category slug, e.g. best-llm-observability"
    },
    "model": {
      "type": "string",
      "enum": [
        "ChatGPT",
        "Claude",
        "Gemini",
        "Grok"
      ],
      "description": "Optional: only this model's picks."
    },
    "limit": {
      "type": "integer",
      "minimum": 1,
      "maximum": 500,
      "description": "Optional: max rows (default 100)."
    }
  },
  "required": [
    "slug"
  ]
}
🟢list_categories

List every category ModelsAgree ranks (slug + title). Large; prefer search_best when you have a specific need.

Input Schema

{
  "type": "object",
  "properties": {}
}

Community

Rate this Server

Evidence

Recent observations

verifiedversion not recorded5 tools
verifiedversion not recorded5 tools