LLM Latency Tracker

Measured latency, time to first token and uptime for ~45 AI inference APIs, by region.

我该使用它吗

质量与安全性

A
描述质量
100%
模式完整度
90%
命名质量
100%
投毒风险
100%
权限匹配度
100%
协议合规性
100%

基于对工具定义和协议合规性的自动分析。

上下文开销

~183token 数(工具定义)
~422 B典型响应大小
对注意力的影响极小(占 128k 上下文窗口的 0.14%)

这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。

安装

一键安装

将以下内容添加到你的 `claude_desktop_config.json` 文件中:

{
  "mcpServers": {
    "llm-latency-tracker": {
      "url": "https://llmlatency.dev/mcp"
    }
  }
}

远程端点

https://llmlatency.dev/mcpstreamable-http

它能做什么

工具清单

工具(2)

🟢 只读🟡 写入🔴 删除⚪ 未知
🟢get_ai_api_latency(region)

Measured latency (TTFB p50/p95) and uptime rankings of AI inference API providers by region, from llmlatency.dev.

输入模式

{
  "type": "object",
  "properties": {
    "region": {
      "type": "string",
      "description": "eu-hetzner, us-central, ap-tokyo or sa-east; omit for all"
    }
  }
}
🟢get_model_deprecations(provider)

AI model deprecation calendar: announced and shutdown dates, replacement models, and how many days of migration notice each provider actually gives (median/min/max). Every entry is verified against the provider own deprecation page.

输入模式

{
  "type": "object",
  "properties": {
    "provider": {
      "type": "string",
      "description": "openai, anthropic, google, mistral, cohere, azure-openai; omit for all"
    }
  }
}

社区

评价此服务器

证据

最近观测

已验证未记录版本2 个工具
已验证未记录版本2 个工具