Newsletter
Join the Community
Subscribe to our newsletter for the latest news and updates
Ultra-low-latency voice AI APIs for speech generation, transcription, translation, and cloning
Gradium is a voice AI platform that provides ultra-low-latency APIs for text-to-speech, speech-to-text, and voice cloning. It enables developers to generate expressive speech, accurate transcriptions, and high-fidelity voice clones through a single API. The platform is designed for building voice agents and real-time voice applications, with a focus on natural, expressive, and robust performance across multiple languages.
Gradium offers a pricing structure with credits. During the public beta, users who provide feedback receive 1 million Gradium credits. Specific pricing plans are available on the Pricing page.
What is the public beta? Gradium has released a new Text-to-Speech model in public beta. It handles complex cases natively, such as phone numbers, email addresses, IBAN numbers, and time expressions, without requiring pre-processing or text normalization.
How can I test the new TTS model? Anyone can test it through the API. Users who send feedback during the beta receive 1 million Gradium credits.
What languages are supported? Gradium supports English, French, Spanish, Portuguese, and German, with robust performance across every language.
Is there an on-device option? Yes, Gradium offers On-Device TTS for offline real-time natural voice generation.
How do I get started? Sign in to the Gradium platform, choose a model, and integrate the API using the documentation. You can also try the demos and the Agent Demo to see the capabilities.