Give your product a voice people remember. Start with a ready-to-use voice or clone one you have permission to use, then bring it to your app with a simple API.
CastReader Voice API is a REST text-to-speech service with a selectable voice library and private, authorized voice cloning. Generate speech in the configured languages for reading, learning, voiceovers and product guidance. It returns MP3 or WAV and provides queued tasks and usage records.
How is the API different from the CastReader reading apps?
The browser extension and mobile apps help people listen to content. Voice API lets developers generate audio inside their own products. You use your CastReader identity, but API access, credits, voices, and usage are managed separately. A Pro reading plan does not include API usage.
How much does text to speech cost?
Successful generation costs $8 per million billable characters, with no monthly API subscription. A paid 100-character request costs $0.0008. Available trial characters are used first. Characters are Unicode code points after NFC normalization, including internal spaces and punctuation; this is not UTF-8 byte billing.
Can I use my own voice or a voice someone shares with me?
You can create a private voice from an authorized recording. The console offers self-recording, a recording invitation, and audio upload. You need the speaker’s permission for cloning and your intended use. Personal voices from the reading apps are not automatically imported into API workspaces.
Is this a real-time or long-form speech API?
The current preview is for short speech and queued workflows. Longer content needs application-side splitting and playback assembly. It does not offer real-time conversations, streaming, or an uptime or latency SLA. Check the published capabilities for current input limits before integrating.
What happens if I retry or close the browser?
A queued task persists when you leave the page. Save its job ID and the original Idempotency-Key. Repeating the same input with that key returns the same task; downloading a retained successful result does not generate or charge again. Waiting and tasks cancelled before execution are free.
Where is API data processed?
One website, account, API key and prepaid wallet. Account, billing and resource metadata use the US control plane. The dispatch API selects China for mainland-China networks and the US elsewhere. Text, reference recordings, inference and private audio stay in the selected region, with no cross-region voice-data fallback. International voice data is separate from the China service, with no automatic cross-region fallback. See the data and voice permissions page for handling and retention rules.
Bring your own voice
Build something worth listening to.
Tell us what you are building, or explore the API before you commit.