Announcing our Series A Funding

Announcing our Series A Funding

Female Voices

One hundred and one female voices, spanning six accent groups and all twelve recommended languages. At that scale accent and language usually constrain casting more than gender does.

VOICES FOR THIS USE CASE

Mishkamishka
FemaleYoungIndian
Best for HindiUse voice
Rhearhea
FemaleYoungIndian
Best for HindiUse voice
Autumnautumn
FemaleYoungAmerican
Best for EnglishUse voice
Kaitlynkaitlyn
FemaleYoungAmerican
Best for EnglishUse voice
Ellieellie
FemaleYoungBritish
Best for EnglishUse voice
Bethanybethany
FemaleYoungBritish
Best for EnglishUse voice

One hundred and one voices

Female voices make up 44 percent of the catalog, spanning Indian, American, British, Australian, Canadian and Latin American accents and all twelve recommended languages. At that scale accent and language are usually the real constraint on casting rather than gender.

What can be adjusted

Speed, from 0.5 to 2.0, is the only delivery control. There is no pitch, tone or emotion parameter, so a voice cannot be shifted after generation. Casting is therefore the decision that matters, and it is worth auditioning several rather than choosing from a description.

Setting it in the API

Gender is not a parameter. You select a voice by ID and its gender comes with it, so changing an assistant to a female voice means changing voice_id to one of these. If you are using a third party product rather than the API, the setting lives in that product.

Language coverage

Female voices appear in both language families. Those in the Indic group cover eleven Indic languages plus English; those in the European group cover ten European languages. No voice spans both, so a multilingual product may need one voice per family rather than one overall.

SPECIFICATION

Sample rate

44.1 kHz native, resampled to 8 kHz for telephony

Latency

~200 ms to first byte (p50, warm region)

Output formats

ulaw, alaw, PCM 16-bit, WAV, MP3

Streaming transports

WebSocket, HTTP chunked transfer

Speed range

0.5× – 2.0×, set per request

Explore Voice Similar to

Female Voices

Male Voices

One hundred and twenty male voices across every accent group and all twelve recommended languages. Most carry an Indic recommendation, which is unusual in catalogs that treat Indian languages as an afterthought.

Young Voices

One hundred and forty six voices are tagged young, 63 percent of the catalog, across every accent and all twelve recommended languages. Unlike tone, age is a real tagged field rather than an inference.

Mature Voices

Eighty six voices are tagged mature, and twenty also carry the narrative tag. That overlap is the usual starting point for audiobooks and documentary work, where consistency shows over hours rather than seconds.

Warm Voices

There is no tone field in the catalog, so these were chosen by listening rather than filtered. Warmth is a property of the voice, not a setting, which makes casting the decision that counts.

Conversational Voices

One hundred and ninety five voices carry the conversational tag, 84 percent of the catalog. Streaming is what makes them feel live: audio begins playing before the sentence has finished rendering.

Indian English AI Voices

One hundred and four voices carry an Indian accent, more than twice any other group. They handle English words inside Devanagari sentences without changing register, which is how Indian English is actually spoken.

Serious Voices

For compliance and legal reads the goal is that exact wording lands, not that the voice sounds grave. A speed near 0.9 improves intelligibility more than any casting choice.

Soothing Voices

Only two voices carry the meditative tag, so this set is drawn wider and chosen by ear. Pace matters more than casting: around 0.7 suits most soothing material.

Authoritative Voices

Slowing down reads as more authoritative than speeding up. A setting near 0.9 does more than any voice choice, and short declarative sentences carry more weight than qualified ones.

Cheerful Voices

Only six voices in the catalog show cheerful or positive markers, the thinnest tone signal available. Speed slightly above 1.0 lifts almost any voice, which is the more dependable route.

Male Voices

One hundred and twenty male voices across every accent group and all twelve recommended languages. Most carry an Indic recommendation, which is unusual in catalogs that treat Indian languages as an afterthought.

Young Voices

One hundred and forty six voices are tagged young, 63 percent of the catalog, across every accent and all twelve recommended languages. Unlike tone, age is a real tagged field rather than an inference.

Mature Voices

Eighty six voices are tagged mature, and twenty also carry the narrative tag. That overlap is the usual starting point for audiobooks and documentary work, where consistency shows over hours rather than seconds.

Warm Voices

There is no tone field in the catalog, so these were chosen by listening rather than filtered. Warmth is a property of the voice, not a setting, which makes casting the decision that counts.

Conversational Voices

One hundred and ninety five voices carry the conversational tag, 84 percent of the catalog. Streaming is what makes them feel live: audio begins playing before the sentence has finished rendering.

Indian English AI Voices

One hundred and four voices carry an Indian accent, more than twice any other group. They handle English words inside Devanagari sentences without changing register, which is how Indian English is actually spoken.

Serious Voices

For compliance and legal reads the goal is that exact wording lands, not that the voice sounds grave. A speed near 0.9 improves intelligibility more than any casting choice.

Soothing Voices

Only two voices carry the meditative tag, so this set is drawn wider and chosen by ear. Pace matters more than casting: around 0.7 suits most soothing material.

FAQs

Set the voice_id in your request to a female voice. One hundred and one are available, across Indian, American, British, Australian, Canadian and Latin American accents. No other parameter changes. If you are using a third party assistant rather than the API directly, the voice setting will be in that product rather than here.