inferenceindexer-mcp

AI inference pricing for agents: live and historical model prices, provider comparison.

¿Debería usar esto?

Calidad y seguridad

A
Calidad de la descripción
96%
Integridad del esquema
65%
Calidad de los nombres
94%
Riesgo de envenenamiento
100%
Coincidencia de permisos
100%
Cumplimiento del protocolo
100%

Basado en el análisis automatizado de las definiciones de herramientas y el cumplimiento del protocolo.

Costo de contexto

~1,798Tokens (definiciones de herramientas)
~870 BTamaño de respuesta típico
Impacto moderado en la atención (1.40% del contexto de 128k)

Este es el número aproximado de tokens que se consumen cada vez que las herramientas del servidor se cargan en el contexto de un modelo. Los recuentos más altos reducen la atención disponible para otras tareas.

Instalar

Instalación con un clic

Agrega esto a tu archivo `claude_desktop_config.json`:

{
  "mcpServers": {
    "inferenceindexer-mcp": {
      "command": "uvx",
      "args": [
        "inferenceindexer-mcp"
      ]
    }
  }
}

Paquetes ejecutables

pypiinferenceindexer-mcp0.1.1stdio

Puntos de conexión remotos

https://api.inferenceindexer.ai/mcpstreamable-http

Qué puede hacer

Inventario de herramientas

Herramientas (10)

🟢 Solo lectura🟡 Escritura🔴 Eliminación⚪ Desconocido
⚪recommend_models(budget_max_usd_per_m, context_min, modality, zdr, eu_sovereign, ...)

Recommend the best-value AI models for given constraints, ranked with receipts. The core answer endpoint: give it constraints and it returns the top models ranked by Cost/IQ (quality-adjusted price, lower is better), each with a plain-English 'why', a hot-swap endpoint_config (provider base_url + native model id, ready to call), as-of timestamps, and runner-ups. Args: budget_max_usd_per_m: Max blended price $/M (optional). context_min: Minimum context window in tokens (optional). modality: 'text' (default), 'vision', or 'any'. zdr: Require zero-data-retention providers (optional). eu_sovereign: Require EU-sovereign providers (optional). reasoning: Filter reasoning models (null = any, true/false). limit: Max recommendations (1-20, default 5). Returns: ranked recommendations with endpoint_config and ranking evidence.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "budget_max_usd_per_m": {
      "anyOf": [
        {
          "type": "number"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Budget Max Usd Per M"
    },
    "context_min": {
      "anyOf": [
        {
          "type": "integer"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Context Min"
    },
    "modality": {
      "default": "text",
      "title": "Modality",
      "type": "string"
    },
    "zdr": {
      "default": false,
      "title": "Zdr",
      "type": "boolean"
    },
    "eu_sovereign": {
      "default": false,
      "title": "Eu Sovereign",
      "type": "boolean"
    },
    "reasoning": {
      "anyOf": [
        {
          "type": "boolean"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Reasoning"
    },
    "limit": {
      "default": 5,
      "title": "Limit",
      "type": "integer"
    }
  },
  "title": "recommend_modelsArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "recommend_modelsDictOutput"
}
🟢explain_model(model_id, history_days)

Get everything about one model in a single call: the full picture. Returns current pricing (input/output/blended, Cost/IQ, 24h/7d changes), a price-history summary with trend, all provider endpoints, the cheapest hand-verified endpoint with its native model id (for hot-swapping), privacy flags (ZDR/EU availability), and the AA intelligence score. Everything is as-of stamped. Args: model_id: Canonical model id, e.g. 'anthropic/claude-sonnet-5'. history_days: Price-history window (default 30, max 365). Returns: complete model profile with pricing, endpoints, privacy, quality.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "model_id": {
      "title": "Model Id",
      "type": "string"
    },
    "history_days": {
      "default": 30,
      "title": "History Days",
      "type": "integer"
    }
  },
  "required": [
    "model_id"
  ],
  "title": "explain_modelArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "explain_modelDictOutput"
}
🟢search_models(query, tier, limit, sort)

Search and list AI inference models with current pricing. Args: query: Text search on model id/name (optional). tier: Filter by tier: frontier | standard | budget | micro | zdr | eu (optional). limit: Max results (1-100, default 25). sort: Sort key, e.g. 'blended' (price), 'sit' (SIT score) (optional). Returns: models with input/output/blended $/M pricing, provider, tier.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "query": {
      "anyOf": [
        {
          "type": "string"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Query"
    },
    "tier": {
      "anyOf": [
        {
          "type": "string"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Tier"
    },
    "limit": {
      "default": 25,
      "title": "Limit",
      "type": "integer"
    },
    "sort": {
      "anyOf": [
        {
          "type": "string"
        },
        {
          "type": "null"
        }
      ],
      "default": null,
      "title": "Sort"
    }
  },
  "title": "search_modelsArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "search_modelsDictOutput"
}
🟢get_model(model_id)

Get full detail + current pricing for one model by its id. Args: model_id: Canonical model id, e.g. 'openai/gpt-5.6' or 'anthropic/claude-sonnet-5'. Returns: pricing, tier, SIT score, quality-adjusted price (Cost/IQ).

Esquema de entrada

{
  "type": "object",
  "properties": {
    "model_id": {
      "title": "Model Id",
      "type": "string"
    }
  },
  "required": [
    "model_id"
  ],
  "title": "get_modelArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "get_modelDictOutput"
}
🟢get_model_history(model_id, days)

Get HISTORICAL price data / trends for one model. This is InferenceIndexer's differentiator: aggregators like OpenRouter expose only current price; this returns the price over time (input, output, blended $/M), enabling trend analysis. Args: model_id: Canonical model id, e.g. 'openai/gpt-5.6'. days: History window in days (1-365, default 30; plan-dependent). Returns: historical price series for the model.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "model_id": {
      "title": "Model Id",
      "type": "string"
    },
    "days": {
      "default": 30,
      "title": "Days",
      "type": "integer"
    }
  },
  "required": [
    "model_id"
  ],
  "title": "get_model_historyArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "get_model_historyDictOutput"
}
🟢list_providers

List all inference providers with model counts and price stats.

Esquema de entrada

{
  "type": "object",
  "properties": {},
  "title": "list_providersArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "list_providersDictOutput"
}
🟢get_provider(provider_name)

Get detail for one provider: models, tier breakdown, price range. Args: provider_name: Provider name, e.g. 'DeepInfra', 'Novita', 'Venice'. Returns: provider detail with model list and pricing.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "provider_name": {
      "title": "Provider Name",
      "type": "string"
    }
  },
  "required": [
    "provider_name"
  ],
  "title": "get_providerArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "get_providerDictOutput"
}
🟢get_composite_latest

Get the current SIT-Composite index value + per-tier breakdown. The SIT-Composite is a usage-weighted mean of the top-50 models by token volume, reflecting what developers actually pay for inference.

Esquema de entrada

{
  "type": "object",
  "properties": {},
  "title": "get_composite_latestArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "get_composite_latestDictOutput"
}
🟢get_composite_history(days)

Get SIT-Composite index history / trend over time. Args: days: History window in days (1-90, default 30). Returns: historical composite index values.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "days": {
      "default": 30,
      "title": "Days",
      "type": "integer"
    }
  },
  "title": "get_composite_historyArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "get_composite_historyDictOutput"
}
⚪compare_providers(model_id)

Compare the price of one model across the providers that host it. Args: model_id: Canonical model id, e.g. 'meta/muse-spark-1.1'. Returns: per-provider endpoints with pricing, showing where direct provider prices diverge (e.g. from OpenRouter's negotiated rate).

Esquema de entrada

{
  "type": "object",
  "properties": {
    "model_id": {
      "title": "Model Id",
      "type": "string"
    }
  },
  "required": [
    "model_id"
  ],
  "title": "compare_providersArguments"
}

Esquema de salida

{
  "type": "object",
  "additionalProperties": true,
  "title": "compare_providersDictOutput"
}

Comunidad

Califica este servidor

Evidencia

Observaciones recientes

verificadoversión no registrada10 herramientas
verificadoversión no registrada10 herramientas
verificadoversión no registrada10 herramientas