Fish Audio S2.1 Pro Free
converts written text into spoken audio. ElevenLabs, Cartesia.fact-checked, all required fields present, publicly visiblereal-time
speech.dev has not published an operational assessment for this entry yet. Facts below are source-linked where available.
Overview
- Publisher
- Fish Audio
- Lab
- fish-audio
- Last updated
- 2026-08-25
- Modality
TTS converts written text into spoken audio. ElevenLabs, Cartesia.
Languages a voice model trained on multiple languages. Quality varies enormously by language — "supports 40 languages" may be great in 3 and mediocre in the other 37.
en-us, zh-cn, ja-jp, de-de, fr-fr, es-es, ko-kr, ar-sa, ru-ru, nl-nl, it-it, pl-pl, pt-br
Pricing
- Billing model
- Per character
- Rate
- $0 per 1M UTF-8 bytes
Last verified August 25, 2026
Technical
- Streaming
- Supported
- Hosting
- saas-vendor-cloud
API access
WSSHTTPS streaming
API endpoints
wsshttps
SDK languages
python, typescript
Regions (vendor buckets)
Macro areas as published by the vendor (often broad; not a country list).
Compliance (vendor-published)
GDPR DPA
Suggest an update
Facts-layer corrections only — source URLs required. Opens a GitHub issue; a maintainer runs the content agent after triage. Not for operational notes.
Lab metadata: Update Fish Audio · All request types · Full guide