mcp

Live LLM API pricing: token prices, comparisons, cheapest-model lookups. No key required.

Should I use this

Quality & Safety

A
Description quality
100%
Schema completeness
77%
Naming quality
95%
Poisoning risk
100%
Permission match
100%
Protocol compliance
100%

Based on automated analysis of tool definitions and protocol compliance.

Context Cost

~589Tokens (tool definitions)
~884 BTypical response size
Minimal attention impact (0.46% of 128k context)

This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.

Install

One-Click Install

Add this to your `claude_desktop_config.json` file:

{
  "mcpServers": {
    "mcp": {
      "command": "npx",
      "args": [
        "@modelpricewatch/mcp"
      ]
    }
  }
}

Runnable packages

npm@modelpricewatch/mcp1.0.0stdio

Remote endpoints

https://modelpricewatch.com/mcpstreamable-http

What it can do

Tool inventory

Tools (4)

🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
🟢search_models(query, category, open_source, limit)

Search the live LLM pricing database by model name, provider, or id. Returns matching models with current input/output prices (USD per 1M tokens), context window, modality, and category. Use this to answer 'how much does <model> cost' or 'what models does <provider> offer'.

Input Schema

{
  "type": "object",
  "properties": {
    "query": {
      "type": "string",
      "description": "Free-text match against model name, provider, or id (e.g. 'claude', 'gpt-5', 'gemini flash'). Omit to list all."
    },
    "category": {
      "type": "string",
      "description": "Filter by category, e.g. flagship, reasoning, budget, coding, embedding, fast, mid-tier."
    },
    "open_source": {
      "type": "boolean",
      "description": "If true, only open-source/open-weight models; if false, only proprietary."
    },
    "limit": {
      "type": "number",
      "description": "Max results (default 20, max 50)."
    }
  }
}
🟢get_model_pricing(model_id)

Get full pricing and capability details for one model by its id (from search_models). Returns input/output/cached price per 1M tokens, blended cost, context window, modality, release date, and the modelpricewatch.com page URL.

Input Schema

{
  "type": "object",
  "properties": {
    "model_id": {
      "type": "string",
      "description": "The model id, e.g. 'anthropic-claude-opus-4-8' or 'openai-gpt-5-5'. Get ids from search_models."
    }
  },
  "required": [
    "model_id"
  ]
}
🟢cheapest_models(sort_by, category, open_source, limit)

Find the cheapest current models, ranked by input price, output price, or a blended cost. The generic ranking covers generative text models (embeddings, OCR and realtime models are excluded — they price different work); pass category to rank a specific pool instead, e.g. 'embedding'. Use to answer 'what is the cheapest model for <use case>'.

Input Schema

{
  "type": "object",
  "properties": {
    "sort_by": {
      "type": "string",
      "description": "Ranking metric: 'input', 'output', or 'blended' (default 'blended')."
    },
    "category": {
      "type": "string",
      "description": "Filter by category, e.g. flagship, reasoning, budget, coding, embedding."
    },
    "open_source": {
      "type": "boolean",
      "description": "If true, only open-source models."
    },
    "limit": {
      "type": "number",
      "description": "How many to return (default 10, max 50)."
    }
  }
}
🟢list_providers

List all tracked AI model providers (OpenAI, Anthropic, Google, etc.) with a short description and their pricing page.

Input Schema

{
  "type": "object",
  "properties": {}
}

Community

Rate this Server

Evidence

Recent observations

verifiedversion not recorded4 tools
verifiedversion not recorded4 tools