kurdish-tts-stt

Sorani & Kurmanji TTS+STT: Kurdish speech most APIs lack. 885 voices, free tier, no key to browse.

我该使用它吗

质量与安全性

A
描述质量
100%
模式完整度
75%
命名质量
93%
投毒风险
100%
权限匹配度
100%
协议合规性
100%

基于对工具定义和协议合规性的自动分析。

上下文开销

~2,462token 数(工具定义)
~3.2 KB典型响应大小
对注意力有中等影响(占 128k 上下文窗口的 1.92%)

这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。

安装

一键安装

将以下内容添加到你的 `claude_desktop_config.json` 文件中:

{
  "mcpServers": {
    "kurdish-tts-stt": {
      "url": "https://www.kurdishtts.com/api/mcp"
    }
  }
}

远程端点

https://www.kurdishtts.com/api/mcpstreamable-http

它能做什么

工具清单

工具(6)

🟢 只读🟡 写入🔴 删除⚪ 未知
🟢list_dialects

List the supported Kurdish dialects and their scripts. Static capability descriptor — costs nothing and needs no API key.

输入模式

{
  "type": "object",
  "properties": {}
}

输出模式

{
  "type": "object",
  "properties": {
    "dialects": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "name": {
            "type": "string"
          },
          "script": {
            "type": "string"
          },
          "direction": {
            "type": "string"
          },
          "speaker_id_prefix": {
            "type": "string"
          }
        },
        "required": [
          "id",
          "name",
          "script",
          "direction",
          "speaker_id_prefix"
        ],
        "additionalProperties": false
      }
    }
  },
  "required": [
    "dialects"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢list_voices(model_version, dialect, gender, limit, offset)

List available Kurdish text-to-speech voices. Returns speaker ids to use with synthesize_speech. Free metadata call — no credits consumed, no API key needed. Defaults to model_version "v3", the SAME default synthesize_speech uses — pass the same model_version to both, because a speaker id from one catalog does not exist in another. Results are paginated; the response reports total_count and which ids a free plan may render.

输入模式

{
  "type": "object",
  "properties": {
    "model_version": {
      "type": "string",
      "enum": [
        "v3",
        "v4",
        "v5"
      ],
      "description": "Voice catalog. Default \"v3\", matching synthesize_speech. v3 is the compatibility catalog; v4 is the large catalog (hundreds of voices); v5 is the curated catalog plus Cast and Studio voices, and requires a paid API plan."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "badini"
      ],
      "description": "Filter by dialect."
    },
    "gender": {
      "type": "string",
      "enum": [
        "male",
        "female"
      ],
      "description": "Filter by voice gender."
    },
    "limit": {
      "type": "integer",
      "minimum": 1,
      "maximum": 200,
      "description": "Maximum voices to return. Default 25."
    },
    "offset": {
      "type": "integer",
      "minimum": 0,
      "description": "Number of voices to skip, for paging. Default 0."
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

输出模式

{
  "type": "object",
  "properties": {
    "model_version": {
      "type": "string"
    },
    "model_version_source": {
      "type": "string"
    },
    "your_plan": {
      "type": [
        "string",
        "null"
      ]
    },
    "your_allowed_model_versions": {
      "anyOf": [
        {
          "type": "array",
          "items": {
            "type": "string"
          }
        },
        {
          "type": "null"
        }
      ]
    },
    "total_count": {
      "type": "number"
    },
    "unique_speaker_ids": {
      "type": "number"
    },
    "returned_count": {
      "type": "number"
    },
    "offset": {
      "type": "number"
    },
    "limit": {
      "type": "number"
    },
    "has_more": {
      "type": "boolean"
    },
    "free_plan_speaker_ids": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "note": {
      "type": "string"
    },
    "speakers": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "speaker_id": {
            "type": "string"
          },
          "name": {
            "type": "string"
          },
          "dialect": {
            "type": "string"
          },
          "gender": {
            "type": "string"
          },
          "free_plan": {
            "type": "boolean"
          },
          "usable_by_you": {
            "type": [
              "boolean",
              "null"
            ]
          },
          "requires_developer_plan": {
            "type": "boolean"
          },
          "reads_either_dialect": {
            "type": "boolean"
          },
          "dialect_note": {
            "type": "string"
          }
        },
        "required": [
          "free_plan",
          "usable_by_you",
          "requires_developer_plan",
          "reads_either_dialect"
        ],
        "additionalProperties": true
      }
    }
  },
  "required": [
    "model_version",
    "model_version_source",
    "your_plan",
    "your_allowed_model_versions",
    "total_count",
    "unique_speaker_ids",
    "returned_count",
    "offset",
    "limit",
    "has_more",
    "free_plan_speaker_ids",
    "note",
    "speakers"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢get_plan

Show what this connection may use. With an API key: your current plan, remaining allowance, which model_versions and speaker ids you may render. Without a key: the purchasable plan ladder and how to get a key. Free, no credits consumed. Call this FIRST when any tool returns a 401 or 403 — it is the fastest way to learn what went wrong.

输入模式

{
  "type": "object",
  "properties": {}
}

输出模式

{
  "type": "object",
  "properties": {
    "authenticated": {
      "type": "boolean"
    },
    "plan": {
      "type": "string"
    },
    "plan_name": {
      "type": "string"
    },
    "status": {
      "type": "string"
    },
    "tts_characters_remaining": {
      "type": [
        "number",
        "null"
      ]
    },
    "tts_characters_quota": {
      "type": [
        "number",
        "null"
      ]
    },
    "tts_characters_used": {
      "type": [
        "number",
        "null"
      ]
    },
    "allowed_model_versions": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "cast_and_studio_voices": {
      "type": "boolean"
    },
    "usable_speaker_ids": {
      "type": "object",
      "additionalProperties": {
        "type": "array",
        "items": {
          "type": "string"
        }
      }
    },
    "stt": {},
    "plans_url": {
      "type": "string"
    },
    "keys_url": {
      "type": "string"
    },
    "next_step": {
      "type": "string"
    },
    "catalog": {}
  },
  "required": [
    "authenticated",
    "plans_url",
    "keys_url",
    "next_step"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟡synthesize_speech(text, speaker_id, model_version, format, dialect, ...)

Convert Kurdish text (Sorani or Kurmanji) to speech audio. Requires a TTS API key; characters are billed against your plan. Get speaker_id from list_voices called with the SAME model_version you pass here (default "v3") — ids are not shared between catalogs. Returns one complete clip: MCP cannot stream, so for a live voice agent call POST https://www.kurdishtts.com/api/tts-stream directly instead (SSE, first audio in ~1s). Max 4000 characters per call in the default mp3 container, 600 with format "wav"; free plans are capped at 500 server-side. Note: speed is caller-facing (higher = faster).

输入模式

{
  "type": "object",
  "properties": {
    "text": {
      "type": "string",
      "minLength": 1,
      "maxLength": 4000,
      "description": "Kurdish text to synthesize. Max 4000 characters (only 600 if you set format \"wav\"); free plans are capped at 500."
    },
    "speaker_id": {
      "type": "string",
      "description": "Voice id from list_voices, e.g. \"sorani_85\" or \"kurmanji_6\" on the default \"v3\" catalog. Dialect is derived from the prefix. Must come from the same model_version you pass below."
    },
    "model_version": {
      "type": "string",
      "enum": [
        "v3",
        "v4",
        "v5"
      ],
      "description": "Voice catalog and plan entitlement. Default \"v3\". Explicit \"v5\" and the Cast/Studio voices require a paid API plan."
    },
    "format": {
      "type": "string",
      "enum": [
        "mp3",
        "opus",
        "wav"
      ],
      "description": "Audio container. Default \"mp3\". Audio is returned base64-encoded inside the tool result, so the container decides how much of your context it costs: for identical speech, mp3 is ~7.5x smaller than wav and opus ~11.5x. Choose \"wav\" only when you need uncompressed audio, and keep the text under 600 characters if you do."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "badini"
      ],
      "description": "Language the text is in. Normally inferred from the speaker_id prefix and safe to omit. REQUIRED to get a correct Kurmanji read from a Cast or Studio voice (cast_*, studio_*): those are tagged sorani after their reference clip but read Sorani and Kurmanji, so without this they pronounce Kurmanji with Sorani phonetics. \"badini\" is valid only with one of the six badini_ voices — pairing it with any other voice is refused with a 400."
    },
    "speed": {
      "type": "number",
      "minimum": 0.25,
      "maximum": 4,
      "description": "Playback speed, higher = faster. Default 1."
    },
    "include_timestamps": {
      "type": "boolean",
      "description": "Return JSON with word-level timestamps instead of an audio block. NOTE: the audio in that JSON is headerless raw PCM16 (24kHz mono), not a WAV file — add a WAV header before saving it, or it will not play."
    }
  },
  "required": [
    "text",
    "speaker_id"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟡transcribe_audio(audio_base64, dialect, filename, mime_type)

Transcribe Kurdish audio (Sorani or Kurmanji) to text. Requires an STT API key; usage is metered per audio minute against your plan. IMPORTANT: dialect selects the decoder and nothing detects it for you — transcribing Sorani audio as Kurmanji returns fluent, confident, WRONG text with no error. Pass dialect "auto" when you are not certain, and pick the coherent transcript from the two it returns. Max 3MB of decoded audio over MCP (~90s of 16kHz WAV, but ~25 minutes of 64kbps MP3 — send compressed audio to fit more); for larger files call POST https://www.kurdishtts.com/api/stt-proxy directly (multipart).

输入模式

{
  "type": "object",
  "properties": {
    "audio_base64": {
      "type": "string",
      "minLength": 1,
      "maxLength": 4194308,
      "description": "Base64-encoded audio file (WAV, MP3, FLAC, OGG or M4A). Max 3MB decoded — compressed formats fit far more speech in that budget than WAV does."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "auto"
      ],
      "description": "Which decoder to run. \"sorani\" or \"kurmanji\" when you know the dialect. \"auto\" transcribes with BOTH and returns both transcripts so you can choose — it bills the audio twice, so prefer a known dialect when you have one."
    },
    "filename": {
      "type": "string",
      "description": "Original filename, used to infer the format. Default audio.wav."
    },
    "mime_type": {
      "type": "string",
      "description": "Audio MIME type. Default audio/wav."
    }
  },
  "required": [
    "audio_base64",
    "dialect"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🔴start_streaming_transcription(dialect)

Open a live speech-to-text session for LISTENING — a microphone or audio stream you transcribe in real time. This does NOT make anything speak; to speak Kurdish use synthesize_speech, or POST https://www.kurdishtts.com/api/tts-stream for progressive audio. WARNING: calling this immediately consumes one streaming session from your STT plan quota — only call it when you are ready to connect. Returns a websocket_url: open it, stream PCM16 mono 16kHz audio chunks, send {"type": "finalize"} to flush and {"type": "done"} to close. Session duration is limited by plan.

输入模式

{
  "type": "object",
  "properties": {
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji"
      ],
      "description": "Dialect that will be spoken."
    }
  },
  "required": [
    "dialect"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

社区

评价此服务器

证据

最近观测

已验证未记录版本6 个工具
已验证未记录版本6 个工具
已验证未记录版本6 个工具