kurdish-tts-stt

Sorani & Kurmanji TTS+STT: Kurdish speech most APIs lack. 885 voices, free tier, no key to browse.

Sollte ich dies verwenden

Qualität und Sicherheit

A
Qualität der Beschreibung
100%
Vollständigkeit des Schemas
75%
Qualität der Benennung
93%
Risiko der Vergiftung
100%
Übereinstimmung der Berechtigungen
100%
Einhaltung des Protokolls
100%

Basierend auf einer automatisierten Analyse der Tool-Definitionen und der Einhaltung des Protokolls.

Kontextkosten

~2,462Tokens (Tool-Definitionen)
~3.2 KBTypische Antwortgröße
Mittlere Auswirkung auf die Aufmerksamkeit (1.92% von 128k Kontext)

Dies ist die ungefähre Anzahl der Tokens, die jedes Mal verbraucht werden, wenn die Tools des Servers in den Kontext eines Modells geladen werden. Höhere Werte verringern die Aufmerksamkeit, die für andere Aufgaben verfügbar ist.

Installieren

Installation mit einem Klick

Fügen Sie dies Ihrer Datei `claude_desktop_config.json` hinzu:

{
  "mcpServers": {
    "kurdish-tts-stt": {
      "url": "https://www.kurdishtts.com/api/mcp"
    }
  }
}

Remote-Endpunkte

https://www.kurdishtts.com/api/mcpstreamable-http

Was es kann

Tool-Inventar

Tools (6)

🟢 Nur lesen🟡 Schreiben🔴 Löschen⚪ Unbekannt
🟢list_dialects

List the supported Kurdish dialects and their scripts. Static capability descriptor — costs nothing and needs no API key.

Eingabe-Schema

{
  "type": "object",
  "properties": {}
}

Ausgabe-Schema

{
  "type": "object",
  "properties": {
    "dialects": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "name": {
            "type": "string"
          },
          "script": {
            "type": "string"
          },
          "direction": {
            "type": "string"
          },
          "speaker_id_prefix": {
            "type": "string"
          }
        },
        "required": [
          "id",
          "name",
          "script",
          "direction",
          "speaker_id_prefix"
        ],
        "additionalProperties": false
      }
    }
  },
  "required": [
    "dialects"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢list_voices(model_version, dialect, gender, limit, offset)

List available Kurdish text-to-speech voices. Returns speaker ids to use with synthesize_speech. Free metadata call — no credits consumed, no API key needed. Defaults to model_version "v3", the SAME default synthesize_speech uses — pass the same model_version to both, because a speaker id from one catalog does not exist in another. Results are paginated; the response reports total_count and which ids a free plan may render.

Eingabe-Schema

{
  "type": "object",
  "properties": {
    "model_version": {
      "type": "string",
      "enum": [
        "v3",
        "v4",
        "v5"
      ],
      "description": "Voice catalog. Default \"v3\", matching synthesize_speech. v3 is the compatibility catalog; v4 is the large catalog (hundreds of voices); v5 is the curated catalog plus Cast and Studio voices, and requires a paid API plan."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "badini"
      ],
      "description": "Filter by dialect."
    },
    "gender": {
      "type": "string",
      "enum": [
        "male",
        "female"
      ],
      "description": "Filter by voice gender."
    },
    "limit": {
      "type": "integer",
      "minimum": 1,
      "maximum": 200,
      "description": "Maximum voices to return. Default 25."
    },
    "offset": {
      "type": "integer",
      "minimum": 0,
      "description": "Number of voices to skip, for paging. Default 0."
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

Ausgabe-Schema

{
  "type": "object",
  "properties": {
    "model_version": {
      "type": "string"
    },
    "model_version_source": {
      "type": "string"
    },
    "your_plan": {
      "type": [
        "string",
        "null"
      ]
    },
    "your_allowed_model_versions": {
      "anyOf": [
        {
          "type": "array",
          "items": {
            "type": "string"
          }
        },
        {
          "type": "null"
        }
      ]
    },
    "total_count": {
      "type": "number"
    },
    "unique_speaker_ids": {
      "type": "number"
    },
    "returned_count": {
      "type": "number"
    },
    "offset": {
      "type": "number"
    },
    "limit": {
      "type": "number"
    },
    "has_more": {
      "type": "boolean"
    },
    "free_plan_speaker_ids": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "note": {
      "type": "string"
    },
    "speakers": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "speaker_id": {
            "type": "string"
          },
          "name": {
            "type": "string"
          },
          "dialect": {
            "type": "string"
          },
          "gender": {
            "type": "string"
          },
          "free_plan": {
            "type": "boolean"
          },
          "usable_by_you": {
            "type": [
              "boolean",
              "null"
            ]
          },
          "requires_developer_plan": {
            "type": "boolean"
          },
          "reads_either_dialect": {
            "type": "boolean"
          },
          "dialect_note": {
            "type": "string"
          }
        },
        "required": [
          "free_plan",
          "usable_by_you",
          "requires_developer_plan",
          "reads_either_dialect"
        ],
        "additionalProperties": true
      }
    }
  },
  "required": [
    "model_version",
    "model_version_source",
    "your_plan",
    "your_allowed_model_versions",
    "total_count",
    "unique_speaker_ids",
    "returned_count",
    "offset",
    "limit",
    "has_more",
    "free_plan_speaker_ids",
    "note",
    "speakers"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢get_plan

Show what this connection may use. With an API key: your current plan, remaining allowance, which model_versions and speaker ids you may render. Without a key: the purchasable plan ladder and how to get a key. Free, no credits consumed. Call this FIRST when any tool returns a 401 or 403 — it is the fastest way to learn what went wrong.

Eingabe-Schema

{
  "type": "object",
  "properties": {}
}

Ausgabe-Schema

{
  "type": "object",
  "properties": {
    "authenticated": {
      "type": "boolean"
    },
    "plan": {
      "type": "string"
    },
    "plan_name": {
      "type": "string"
    },
    "status": {
      "type": "string"
    },
    "tts_characters_remaining": {
      "type": [
        "number",
        "null"
      ]
    },
    "tts_characters_quota": {
      "type": [
        "number",
        "null"
      ]
    },
    "tts_characters_used": {
      "type": [
        "number",
        "null"
      ]
    },
    "allowed_model_versions": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "cast_and_studio_voices": {
      "type": "boolean"
    },
    "usable_speaker_ids": {
      "type": "object",
      "additionalProperties": {
        "type": "array",
        "items": {
          "type": "string"
        }
      }
    },
    "stt": {},
    "plans_url": {
      "type": "string"
    },
    "keys_url": {
      "type": "string"
    },
    "next_step": {
      "type": "string"
    },
    "catalog": {}
  },
  "required": [
    "authenticated",
    "plans_url",
    "keys_url",
    "next_step"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟡synthesize_speech(text, speaker_id, model_version, format, dialect, ...)

Convert Kurdish text (Sorani or Kurmanji) to speech audio. Requires a TTS API key; characters are billed against your plan. Get speaker_id from list_voices called with the SAME model_version you pass here (default "v3") — ids are not shared between catalogs. Returns one complete clip: MCP cannot stream, so for a live voice agent call POST https://www.kurdishtts.com/api/tts-stream directly instead (SSE, first audio in ~1s). Max 4000 characters per call in the default mp3 container, 600 with format "wav"; free plans are capped at 500 server-side. Note: speed is caller-facing (higher = faster).

Eingabe-Schema

{
  "type": "object",
  "properties": {
    "text": {
      "type": "string",
      "minLength": 1,
      "maxLength": 4000,
      "description": "Kurdish text to synthesize. Max 4000 characters (only 600 if you set format \"wav\"); free plans are capped at 500."
    },
    "speaker_id": {
      "type": "string",
      "description": "Voice id from list_voices, e.g. \"sorani_85\" or \"kurmanji_6\" on the default \"v3\" catalog. Dialect is derived from the prefix. Must come from the same model_version you pass below."
    },
    "model_version": {
      "type": "string",
      "enum": [
        "v3",
        "v4",
        "v5"
      ],
      "description": "Voice catalog and plan entitlement. Default \"v3\". Explicit \"v5\" and the Cast/Studio voices require a paid API plan."
    },
    "format": {
      "type": "string",
      "enum": [
        "mp3",
        "opus",
        "wav"
      ],
      "description": "Audio container. Default \"mp3\". Audio is returned base64-encoded inside the tool result, so the container decides how much of your context it costs: for identical speech, mp3 is ~7.5x smaller than wav and opus ~11.5x. Choose \"wav\" only when you need uncompressed audio, and keep the text under 600 characters if you do."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "badini"
      ],
      "description": "Language the text is in. Normally inferred from the speaker_id prefix and safe to omit. REQUIRED to get a correct Kurmanji read from a Cast or Studio voice (cast_*, studio_*): those are tagged sorani after their reference clip but read Sorani and Kurmanji, so without this they pronounce Kurmanji with Sorani phonetics. \"badini\" is valid only with one of the six badini_ voices — pairing it with any other voice is refused with a 400."
    },
    "speed": {
      "type": "number",
      "minimum": 0.25,
      "maximum": 4,
      "description": "Playback speed, higher = faster. Default 1."
    },
    "include_timestamps": {
      "type": "boolean",
      "description": "Return JSON with word-level timestamps instead of an audio block. NOTE: the audio in that JSON is headerless raw PCM16 (24kHz mono), not a WAV file — add a WAV header before saving it, or it will not play."
    }
  },
  "required": [
    "text",
    "speaker_id"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟡transcribe_audio(audio_base64, dialect, filename, mime_type)

Transcribe Kurdish audio (Sorani or Kurmanji) to text. Requires an STT API key; usage is metered per audio minute against your plan. IMPORTANT: dialect selects the decoder and nothing detects it for you — transcribing Sorani audio as Kurmanji returns fluent, confident, WRONG text with no error. Pass dialect "auto" when you are not certain, and pick the coherent transcript from the two it returns. Max 3MB of decoded audio over MCP (~90s of 16kHz WAV, but ~25 minutes of 64kbps MP3 — send compressed audio to fit more); for larger files call POST https://www.kurdishtts.com/api/stt-proxy directly (multipart).

Eingabe-Schema

{
  "type": "object",
  "properties": {
    "audio_base64": {
      "type": "string",
      "minLength": 1,
      "maxLength": 4194308,
      "description": "Base64-encoded audio file (WAV, MP3, FLAC, OGG or M4A). Max 3MB decoded — compressed formats fit far more speech in that budget than WAV does."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "auto"
      ],
      "description": "Which decoder to run. \"sorani\" or \"kurmanji\" when you know the dialect. \"auto\" transcribes with BOTH and returns both transcripts so you can choose — it bills the audio twice, so prefer a known dialect when you have one."
    },
    "filename": {
      "type": "string",
      "description": "Original filename, used to infer the format. Default audio.wav."
    },
    "mime_type": {
      "type": "string",
      "description": "Audio MIME type. Default audio/wav."
    }
  },
  "required": [
    "audio_base64",
    "dialect"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🔴start_streaming_transcription(dialect)

Open a live speech-to-text session for LISTENING — a microphone or audio stream you transcribe in real time. This does NOT make anything speak; to speak Kurdish use synthesize_speech, or POST https://www.kurdishtts.com/api/tts-stream for progressive audio. WARNING: calling this immediately consumes one streaming session from your STT plan quota — only call it when you are ready to connect. Returns a websocket_url: open it, stream PCM16 mono 16kHz audio chunks, send {"type": "finalize"} to flush and {"type": "done"} to close. Session duration is limited by plan.

Eingabe-Schema

{
  "type": "object",
  "properties": {
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji"
      ],
      "description": "Dialect that will be spoken."
    }
  },
  "required": [
    "dialect"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

Community

Diesen Server bewerten

Nachweis

Aktuelle Beobachtungen

verifiziertVersion nicht aufgezeichnet6 Tools
verifiziertVersion nicht aufgezeichnet6 Tools
verifiziertVersion nicht aufgezeichnet6 Tools