Hinglish and beyond
Real speech mixes languages mid-sentence. Pāṇini detects the switch inline and keeps one continuous voice across scripts, so Hinglish, Tanglish or Arabic-French code mixing never breaks the delivery.
Developers · Pāṇini text to speech
Code switching, emotion, non-verbal expression and custom pronunciation, in 250+ languages from a single endpoint.
250+
Languages, one model
120
Platform voices
< 200ms
Time to first byte
Try it
Capabilities
Each control is a field on the same request body. No separate models, no per-feature endpoints.
Real speech mixes languages mid-sentence. Pāṇini detects the switch inline and keeps one continuous voice across scripts, so Hinglish, Tanglish or Arabic-French code mixing never breaks the delivery.
Set language to auto or pin it per request. The same voice identity carries across every language without re-cloning or a separate model per locale.
Steer delivery with emotion and style controls: empathetic for support, assertive for collections, cheerful for onboarding. Intensity is adjustable per request.
Inline tags render breaths, laughs, sighs, hesitations and gasps as part of the audio stream. No extra calls, no stitched clips.
IFSC, KYC, EMI, GST, SKU and dates, currencies and phone numbers are normalised per language before synthesis, so nothing gets spelled out awkwardly.
Upload a pronunciation dictionary for brand names, product SKUs, drug names or local place names. Entries apply across every language and every voice.
Adjust speed, pitch and pause length without artefacts. Useful for IVR prompts, audiobooks and telephony where pacing matters.
Audio streams as it generates, so playback begins before the full response is synthesised. WAV, MP3, PCM and μ-law for telephony.
POST /v1/speech/stream
Audio is generated sentence by sentence and streamed back as a chunked response, so the first bytes arrive long before the last word is synthesised. Use this for agents, telephony and anything a person is waiting on.
Keys start with vsk_ and are sent as a bearer token. Every voice speaks every language, so the only required fields are text and voice_id.
1curl -N -X POST https://api.sonexlabs.com/v1/speech/stream \2 -H "Authorization: Bearer $SONEX_API_KEY" \3 -H "Content-Type: application/json" \4 -d '{5 "text": "नमस्ते! आपका order कल deliver हो जाएगा.",6 "voice_id": "72ly9crx9v",7 "output_format": "wav"8 }' \9 --output speech.wavVoices
Every voice speaks all 250+ languages while keeping its own timbre. Need something specific? Clone a voice from ten seconds of audio.
Indian · Warm, conversational
Indian · Calm, assured
Indian · Bright, upbeat
Indian · Deep, narrative
Indian · Crisp, professional
American · Neutral, clear
British · Measured, editorial
Latin American · Friendly, natural
Japanese · Soft, polite
West African · Rich, grounded
Levantine · Steady, formal
Global · Even, brandless
Language coverage
Twenty four scripts below, and the full register of 250 languages underneath. Most of them have never had a production voice model before.
Bhojpuri
Marwari
Garhwali
Saraiki
Hindi
Tamil
Telugu
Bengali
Marathi
Gujarati
Punjabi
Kannada
Malayalam
Odia
Assamese
Sanskrit
Urdu
Nepali
Yoruba
Swahili
Zulu
Somali
Dagbani
Arabic
Chinese
The register
220 of 250 languages listed
A
B
C
D
E
F
G
H
I
J
K
L
M
N
O
P
R
S
T
U
V
W
X
Y
Z
Why one model matters
Most multilingual TTS is English first, with other languages bolted on afterwards. Pāṇini was multilingual from the first training run, so latency, price and quality do not change when the language does.
~30
Languages typical multilingual voice AI covers
~4B
People whose first language falls outside that ~30
1,300+
Languages spoken across India alone
250+
Languages Pāṇini covers at one flat rate
Scale
< 200 ms
Time to first byte
250+
Languages, one endpoint
120
Platform voices, plug and play
36+
Indian languages and dialects
8 kHz
Telephony-grade training
10,000
Concurrent streams
Pricing
Bhojpuri costs what English costs. No language tiers, no minimum commitment.
Free tier, no card, 10,000 characters included. Every language available from the first request.