Prepared learning clips, product guidance and short voiceovers with authorized cloned voices. Current deployment: 10 languages, 500 characters/job, concurrency 1.
When to choose another service
If streaming conversation, native multi-speaker orchestration, emotion controls or a contracted latency SLA are mandatory, do not select this CastReader offering on price alone. We do not currently provide those capabilities.
What the published offering actually says.
Prices are USD list rates, not a binding quote. Trials, taxes, discounts, top-up minimums and account terms may change the amount you pay. Unknown means verify—not unsupported.
A candidate when faster delivery, more languages or longer individual inputs are required.
Access, limits and source
Pay-as-you-go is available; check free/startup allowances and account terms separately.
Official Flash/Turbo listing: 32 languages, 40,000-character input limit. v2/v3 have different prices and capabilities; not interchangeable benchmarks.
Runs in your browser. No text is uploaded and no audio is generated.
1–1,000,000 whole requests. This models repeating the text, not one oversized request.
All estimates use the same explicitly preprocessed text: CRLF → LF, NFC normalization, outer whitespace trimmed. Internal spaces and punctuation remain.
58Unicode code points / request
58UTF-8 bytes / request
CastReaderclone-v1$0.464
ElevenLabsFlash / Turbo$2.90
Fish Audios2.1-pro (paid)$0.87
DeepgramAura-2$1.74
CartesiaPro planNot comparable from text alone
OpenAIgpt-4o-mini-ttsNot comparable from text alone
Paid-rate illustration, not “lowest price.” Fish also lists s2.1-pro-free at $0; check its access and limits. Trial credits, subscriptions, discounts, taxes and minimum top-ups are not included.
ElevenLabs/Deepgram estimates assume one billed character per Unicode code point; verify their exact metering. Credits and output audio tokens require their own usage data. This is not equal-quality or equal-speed testing.
A lower rate is only one part of the decision.
Connect and recover
Install the exact published client, check credentials and the ready voice, then save the request and idempotency key before submitting. A timeout is not permission to create another billed job.
Our expanded first-party run records 30 attempts, 29 saved outputs and one failure. It is not a competing-model benchmark or a native-speaker verdict. Inspect the original reference, generated audio, input and usage.
total_cost = successful_audio_usage
+ deliberate_regenerations
+ subscription_or_unused_minimums
+ app_hosting_storage_and_delivery
+ integration_and_recovery_work
// Measure complete-job time separately from first-audio time.// Do not infer output audio tokens from text character count.