
What is KugelAudio?
KugelAudio is a state-of-the-art text-to-speech (TTS) platform designed for real-time applications. Whether you’re building voice agents, interactive applications, or content creation tools, KugelAudio provides the speed and quality you need.Quick Start
Get up and running with KugelAudio in under 5 minutes
Generate Speech
Generate high-quality audio from text
Streaming
Real-time audio streaming for low latency
Voices
Browse voices and create custom clones
Key Features
Legacy model compatibility
Legacy model compatibility
Existing model IDs remain accepted for backwards compatibility. See Models for the current and legacy IDs.
WebSocket Streaming
WebSocket Streaming
Stream audio chunks as they’re generated for the lowest possible latency. Perfect for LLM integrations where text arrives token by token.
Voice Cloning
Voice Cloning
Create custom voices from 10–30 seconds of clean reference audio. See Voice cloning for recording requirements.
Multi-Language Support
Multi-Language Support
Multilingual single model — 39 languages including DE, EN, FR, ES, IT, NL, PT, PL, RU, ZH, JA, KO, AR. See the TTS endpoint reference for the full list.
Getting Started
1
Get Your API Key
Sign up at kugelaudio.com and get your API key from the dashboard.
2
Install the SDK
Choose your preferred SDK and install it:
3
Generate Your First Audio
Need Help?
API Reference
Detailed API documentation with examples
