Skip to navigation

SpeechifyAI Build

Build speech into your product with one API. Generate audio, stream long-form text, clone voices, and control delivery with SSML.

Your first request

POST
/v1/audio/speech
from speechify import Speechify
client = Speechify(
token="YOUR_TOKEN_HERE",
)
client.audio.speech(
audio_format="mp3",
input="Hello! This is the Speechify text-to-speech API.",
model="simba-3.2",
voice_id="geffen_32",
)

Ready to run it end to end? The Quickstart walks you through your first call: get a key, install the SDK, generate speech, and play it.

Grab an API key at platform.speechify.ai/api-keys and set SPEECHIFY_API_KEY so the SDKs authenticate automatically.

Set up

Build With Speech

Integrations

Speechify voices drop into every major voice-agent platform. Native plugins where they exist, an open-source tts-shims proxy where they don’t.

See the Integrations overview for the platform picker, including Puter.

Models and languages

ModelBest forLanguagesHighlights
simba-3.2Recommended for new English integrationsEnglishLowest TTFB, richest expressivity; the recommended Simba 3 model
simba-3.0Streaming-native beyond English6 languages, 7 localesThe API default when model is omitted; English, German, Spanish, French, Italian and Brazilian Portuguese, set language to pick one
simba-english, simba-multilingualRetired (Simba 1.6)A new workspace gets 400 model_retired: use simba-3.2 for English and simba-3.0 for German, Spanish, French, Italian or Portuguese. For any other language, check each model’s languages in GET /v1/audio/models. Only a workspace pinned to an API version before 2026-09-21 can still name them, and from 2026-11-21 they are served by our current models

See Models and Language Support for the full matrix. Still choosing a vendor? Best TTS APIs of 2026 puts Simba 3.2 beside the other models it covers, all evaluated on the Artificial Analysis Speech Arena, with each one’s price per million characters.

Resources