Google Cloud
converts written text into spoken audio. ElevenLabs, Cartesia.fact-checked, all required fields present, publicly visiblenear-real-time
speech.dev has not published an operational assessment for this entry yet. Facts below are source-linked where available.
Overview
- Publisher
- Google Cloud
- Lab
- google-cloud
- Last updated
- 2026-05-23
- Modality
TTS converts written text into spoken audio. ElevenLabs, Cartesia.
Languages a voice model trained on multiple languages. Quality varies enormously by language — "supports 40 languages" may be great in 3 and mediocre in the other 37.
multilingual
Pricing
Rate tiers
Standard
Per character · $4 per 1M characters
WaveNet / Neural2
Per character · $16 per 1M characters
- Default billing model
- Per character
- Headline rate
- $16 per 1M characters
Last verified May 23, 2026
Benchmarks
Speech Arena quality (AA)
Artificial Analysis TTSListed as: WaveNet
Indexed on Artificial Analysis text-to-speech comparison; quality Elo from Speech Arena.
Verified 2026-05-23
Technical
- Streaming
- Supported
- Hosting
- saas-vendor-cloud
- Self-hostable
- No
API access
HTTPS streaminggRPC
API endpoints
httpsgrpc
SDK languages
python, typescript, go, java
Regions (vendor buckets)
Macro areas as published by the vendor (often broad; not a country list).
Compliance (vendor-published)
GDPR DPAHIPAA BAAISO 27001PCI DSSSOC 2 Type II
Suggest an update
Facts-layer corrections only — source URLs required. Opens a GitHub issue; a maintainer runs the content agent after triage. Not for operational notes.