Kenpath Labs

Voices & languages

The voice roster, preview clips, and the languages endpoint.

The voice roster

The library ships 320 human-reviewed voices across 66 native languages and 76 accents — every voice speaks all 82 supported languages; its native one is where its accent lives. Fetch the roster at runtime rather than hardcoding a list; store each voice’s voice_id and display its name.

GET/v1/voices
{
"voices": [
{
"voice_id": "sv_enhdbrj5",
"name": "Aanya",
"category": "premade",
"gender": "female",
"accent_family": "bengali",
"model_id": "svara-tts-turbo",
"description": "A soft female voice with a bengali accent, made for storytelling that leans close — romantic, unhurried, with real warmth in the low notes.",
"labels": {
"native_language": "Bengali",
"native_language_code": "bn",
"accent": "bengali",
"region": "india",
"age": "young",
"quality_band": "A",
"tags": "soft",
"preview_language": "Bengali"
},
"preview_url": "/v1/voices/sv_enhdbrj5/preview"
},
...
]
}
  • voice_id is a stable sv_-prefixed id; it is the only way to address a voice, on both the OpenAI-style and ElevenLabs-style endpoints
  • labels.native_language is where the voice’s accent comes from; it still speaks every supported language
  • names change, ids don’t — store the voice_id and render name from the response
The console Voices page plays every voice and filters by language and gender, if you want to pick one before writing code.

Voice previews

Every voice has a bundled preview clip; use it instead of live-synthesizing sample audio (it’s instant and free):

GET/v1/voices/{voice_id}/preview

The preview_url is included in each voice object from /v1/voices, so you rarely construct this URL by hand.

Languages

The model handles 82 languages and code-switches between them within a single request. Drive language pickers from the languages endpoint:

GET/v1/languages

When calling TTS, the lang hint accepts any ISO form: hi, hin, hindi, hi-IN, ja, zh-CN, ko all resolve. Sending it enables number and unit normalization; omit it to auto-detect from the script. See Text to speech for details.