Vakyam AI is a developer platform for natural, human-sounding speech in Indian
languages. Our text-to-speech model, Raaga 1, turns text in English, Hindi,
Tamil, Telugu, Kannada, Marathi, Gujarati, and Bengali into lifelike audio
through a simple HTTP and WebSocket API.
These docs will take you from your first API call to streaming real-time audio
in production.
Eight Indian languages — en-IN, hi-IN, ta-IN, te-IN, kn-IN,
mr-IN, gu-IN, bn-IN.
Multiple voices per language — the same person can speak across languages.
Flexible output — mp3, wav, raw pcm, or mulaw, at sample rates from
8 kHz to 48 kHz.
Fair credit counting — monthly character allotments by plan,
plus optional extra credits at ₹0.75 per 1,000 characters, counted by Unicode
character (aksara), not by byte.
Plan-based limits — Free, Developer, and Growth tiers set your monthly
characters, requests per minute, and concurrency.