Professional Voices
Corporate narration needs consistent terminology more than it needs a particular tone. Pronunciation dictionaries fix brand and product names once, so every module in a library matches the first one.

VOICES FOR THIS USE CASE
Drawn from two working pools
There is no professional tag, so this set comes from the educational and Voice Agent groups, which between them cover the contexts where a composed register matters. Twenty two voices carry one of those tags. Selection within them rests on listening rather than on any documented attribute.
Pace over character
Speed, adjustable from 0.5 to 2.0, is the only delivery control. For corporate narration a setting slightly below 1.0 reads as more measured. There is no tone parameter, so a voice cannot be made to sound more formal after the fact.
Terminology consistency
Corporate material repeats brand names, product names and acronyms constantly, and inconsistency across a library is more noticeable than any individual mispronunciation. Pronunciation dictionaries fix each term once so every generation matches, which matters more for recurring material than for one off scripts.
Consistency across a long script
Requests cap at 250 characters, so a quarterly presentation or training script is assembled from many calls. Keeping voice, speed and sample rate fixed across the whole run removes most variation, but generating a full section and listening across the joins is still worth doing before committing.
SPECIFICATION
Sample rate
44.1 kHz native, resampled to 8 kHz for telephony
Latency
~200 ms to first byte (p50, warm region)
Output formats
ulaw, alaw, PCM 16-bit, WAV, MP3
Streaming transports
WebSocket, HTTP chunked transfer
Speed range
0.5× – 2.0×, set per request
Explore Voice Similar to
Professional Voices
FAQs










