kurdish-tts-stt

Sorani & Kurmanji TTS+STT: Kurdish speech most APIs lack. 885 voices, free tier, no key to browse.

¿Debería usar esto?

Calidad y seguridad

A
Calidad de la descripción
100%
Integridad del esquema
75%
Calidad de los nombres
93%
Riesgo de envenenamiento
100%
Coincidencia de permisos
100%
Cumplimiento del protocolo
100%

Basado en el análisis automatizado de las definiciones de herramientas y el cumplimiento del protocolo.

Costo de contexto

~2,462Tokens (definiciones de herramientas)
~3.2 KBTamaño de respuesta típico
Impacto moderado en la atención (1.92% del contexto de 128k)

Este es el número aproximado de tokens que se consumen cada vez que las herramientas del servidor se cargan en el contexto de un modelo. Los recuentos más altos reducen la atención disponible para otras tareas.

Instalar

Instalación con un clic

Agrega esto a tu archivo `claude_desktop_config.json`:

{
  "mcpServers": {
    "kurdish-tts-stt": {
      "url": "https://www.kurdishtts.com/api/mcp"
    }
  }
}

Puntos de conexión remotos

https://www.kurdishtts.com/api/mcpstreamable-http

Qué puede hacer

Inventario de herramientas

Herramientas (6)

🟢 Solo lectura🟡 Escritura🔴 Eliminación⚪ Desconocido
🟢list_dialects

List the supported Kurdish dialects and their scripts. Static capability descriptor — costs nothing and needs no API key.

Esquema de entrada

{
  "type": "object",
  "properties": {}
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "dialects": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "name": {
            "type": "string"
          },
          "script": {
            "type": "string"
          },
          "direction": {
            "type": "string"
          },
          "speaker_id_prefix": {
            "type": "string"
          }
        },
        "required": [
          "id",
          "name",
          "script",
          "direction",
          "speaker_id_prefix"
        ],
        "additionalProperties": false
      }
    }
  },
  "required": [
    "dialects"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢list_voices(model_version, dialect, gender, limit, offset)

List available Kurdish text-to-speech voices. Returns speaker ids to use with synthesize_speech. Free metadata call — no credits consumed, no API key needed. Defaults to model_version "v3", the SAME default synthesize_speech uses — pass the same model_version to both, because a speaker id from one catalog does not exist in another. Results are paginated; the response reports total_count and which ids a free plan may render.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "model_version": {
      "type": "string",
      "enum": [
        "v3",
        "v4",
        "v5"
      ],
      "description": "Voice catalog. Default \"v3\", matching synthesize_speech. v3 is the compatibility catalog; v4 is the large catalog (hundreds of voices); v5 is the curated catalog plus Cast and Studio voices, and requires a paid API plan."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "badini"
      ],
      "description": "Filter by dialect."
    },
    "gender": {
      "type": "string",
      "enum": [
        "male",
        "female"
      ],
      "description": "Filter by voice gender."
    },
    "limit": {
      "type": "integer",
      "minimum": 1,
      "maximum": 200,
      "description": "Maximum voices to return. Default 25."
    },
    "offset": {
      "type": "integer",
      "minimum": 0,
      "description": "Number of voices to skip, for paging. Default 0."
    }
  },
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "model_version": {
      "type": "string"
    },
    "model_version_source": {
      "type": "string"
    },
    "your_plan": {
      "type": [
        "string",
        "null"
      ]
    },
    "your_allowed_model_versions": {
      "anyOf": [
        {
          "type": "array",
          "items": {
            "type": "string"
          }
        },
        {
          "type": "null"
        }
      ]
    },
    "total_count": {
      "type": "number"
    },
    "unique_speaker_ids": {
      "type": "number"
    },
    "returned_count": {
      "type": "number"
    },
    "offset": {
      "type": "number"
    },
    "limit": {
      "type": "number"
    },
    "has_more": {
      "type": "boolean"
    },
    "free_plan_speaker_ids": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "note": {
      "type": "string"
    },
    "speakers": {
      "type": "array",
      "items": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string"
          },
          "speaker_id": {
            "type": "string"
          },
          "name": {
            "type": "string"
          },
          "dialect": {
            "type": "string"
          },
          "gender": {
            "type": "string"
          },
          "free_plan": {
            "type": "boolean"
          },
          "usable_by_you": {
            "type": [
              "boolean",
              "null"
            ]
          },
          "requires_developer_plan": {
            "type": "boolean"
          },
          "reads_either_dialect": {
            "type": "boolean"
          },
          "dialect_note": {
            "type": "string"
          }
        },
        "required": [
          "free_plan",
          "usable_by_you",
          "requires_developer_plan",
          "reads_either_dialect"
        ],
        "additionalProperties": true
      }
    }
  },
  "required": [
    "model_version",
    "model_version_source",
    "your_plan",
    "your_allowed_model_versions",
    "total_count",
    "unique_speaker_ids",
    "returned_count",
    "offset",
    "limit",
    "has_more",
    "free_plan_speaker_ids",
    "note",
    "speakers"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟢get_plan

Show what this connection may use. With an API key: your current plan, remaining allowance, which model_versions and speaker ids you may render. Without a key: the purchasable plan ladder and how to get a key. Free, no credits consumed. Call this FIRST when any tool returns a 401 or 403 — it is the fastest way to learn what went wrong.

Esquema de entrada

{
  "type": "object",
  "properties": {}
}

Esquema de salida

{
  "type": "object",
  "properties": {
    "authenticated": {
      "type": "boolean"
    },
    "plan": {
      "type": "string"
    },
    "plan_name": {
      "type": "string"
    },
    "status": {
      "type": "string"
    },
    "tts_characters_remaining": {
      "type": [
        "number",
        "null"
      ]
    },
    "tts_characters_quota": {
      "type": [
        "number",
        "null"
      ]
    },
    "tts_characters_used": {
      "type": [
        "number",
        "null"
      ]
    },
    "allowed_model_versions": {
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "cast_and_studio_voices": {
      "type": "boolean"
    },
    "usable_speaker_ids": {
      "type": "object",
      "additionalProperties": {
        "type": "array",
        "items": {
          "type": "string"
        }
      }
    },
    "stt": {},
    "plans_url": {
      "type": "string"
    },
    "keys_url": {
      "type": "string"
    },
    "next_step": {
      "type": "string"
    },
    "catalog": {}
  },
  "required": [
    "authenticated",
    "plans_url",
    "keys_url",
    "next_step"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟡synthesize_speech(text, speaker_id, model_version, format, dialect, ...)

Convert Kurdish text (Sorani or Kurmanji) to speech audio. Requires a TTS API key; characters are billed against your plan. Get speaker_id from list_voices called with the SAME model_version you pass here (default "v3") — ids are not shared between catalogs. Returns one complete clip: MCP cannot stream, so for a live voice agent call POST https://www.kurdishtts.com/api/tts-stream directly instead (SSE, first audio in ~1s). Max 4000 characters per call in the default mp3 container, 600 with format "wav"; free plans are capped at 500 server-side. Note: speed is caller-facing (higher = faster).

Esquema de entrada

{
  "type": "object",
  "properties": {
    "text": {
      "type": "string",
      "minLength": 1,
      "maxLength": 4000,
      "description": "Kurdish text to synthesize. Max 4000 characters (only 600 if you set format \"wav\"); free plans are capped at 500."
    },
    "speaker_id": {
      "type": "string",
      "description": "Voice id from list_voices, e.g. \"sorani_85\" or \"kurmanji_6\" on the default \"v3\" catalog. Dialect is derived from the prefix. Must come from the same model_version you pass below."
    },
    "model_version": {
      "type": "string",
      "enum": [
        "v3",
        "v4",
        "v5"
      ],
      "description": "Voice catalog and plan entitlement. Default \"v3\". Explicit \"v5\" and the Cast/Studio voices require a paid API plan."
    },
    "format": {
      "type": "string",
      "enum": [
        "mp3",
        "opus",
        "wav"
      ],
      "description": "Audio container. Default \"mp3\". Audio is returned base64-encoded inside the tool result, so the container decides how much of your context it costs: for identical speech, mp3 is ~7.5x smaller than wav and opus ~11.5x. Choose \"wav\" only when you need uncompressed audio, and keep the text under 600 characters if you do."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "badini"
      ],
      "description": "Language the text is in. Normally inferred from the speaker_id prefix and safe to omit. REQUIRED to get a correct Kurmanji read from a Cast or Studio voice (cast_*, studio_*): those are tagged sorani after their reference clip but read Sorani and Kurmanji, so without this they pronounce Kurmanji with Sorani phonetics. \"badini\" is valid only with one of the six badini_ voices — pairing it with any other voice is refused with a 400."
    },
    "speed": {
      "type": "number",
      "minimum": 0.25,
      "maximum": 4,
      "description": "Playback speed, higher = faster. Default 1."
    },
    "include_timestamps": {
      "type": "boolean",
      "description": "Return JSON with word-level timestamps instead of an audio block. NOTE: the audio in that JSON is headerless raw PCM16 (24kHz mono), not a WAV file — add a WAV header before saving it, or it will not play."
    }
  },
  "required": [
    "text",
    "speaker_id"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🟡transcribe_audio(audio_base64, dialect, filename, mime_type)

Transcribe Kurdish audio (Sorani or Kurmanji) to text. Requires an STT API key; usage is metered per audio minute against your plan. IMPORTANT: dialect selects the decoder and nothing detects it for you — transcribing Sorani audio as Kurmanji returns fluent, confident, WRONG text with no error. Pass dialect "auto" when you are not certain, and pick the coherent transcript from the two it returns. Max 3MB of decoded audio over MCP (~90s of 16kHz WAV, but ~25 minutes of 64kbps MP3 — send compressed audio to fit more); for larger files call POST https://www.kurdishtts.com/api/stt-proxy directly (multipart).

Esquema de entrada

{
  "type": "object",
  "properties": {
    "audio_base64": {
      "type": "string",
      "minLength": 1,
      "maxLength": 4194308,
      "description": "Base64-encoded audio file (WAV, MP3, FLAC, OGG or M4A). Max 3MB decoded — compressed formats fit far more speech in that budget than WAV does."
    },
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji",
        "auto"
      ],
      "description": "Which decoder to run. \"sorani\" or \"kurmanji\" when you know the dialect. \"auto\" transcribes with BOTH and returns both transcripts so you can choose — it bills the audio twice, so prefer a known dialect when you have one."
    },
    "filename": {
      "type": "string",
      "description": "Original filename, used to infer the format. Default audio.wav."
    },
    "mime_type": {
      "type": "string",
      "description": "Audio MIME type. Default audio/wav."
    }
  },
  "required": [
    "audio_base64",
    "dialect"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}
🔴start_streaming_transcription(dialect)

Open a live speech-to-text session for LISTENING — a microphone or audio stream you transcribe in real time. This does NOT make anything speak; to speak Kurdish use synthesize_speech, or POST https://www.kurdishtts.com/api/tts-stream for progressive audio. WARNING: calling this immediately consumes one streaming session from your STT plan quota — only call it when you are ready to connect. Returns a websocket_url: open it, stream PCM16 mono 16kHz audio chunks, send {"type": "finalize"} to flush and {"type": "done"} to close. Session duration is limited by plan.

Esquema de entrada

{
  "type": "object",
  "properties": {
    "dialect": {
      "type": "string",
      "enum": [
        "sorani",
        "kurmanji"
      ],
      "description": "Dialect that will be spoken."
    }
  },
  "required": [
    "dialect"
  ],
  "additionalProperties": false,
  "$schema": "http://json-schema.org/draft-07/schema#"
}

Comunidad

Califica este servidor

Evidencia

Observaciones recientes

verificadoversión no registrada6 herramientas
verificadoversión no registrada6 herramientas
verificadoversión no registrada6 herramientas