Skip to main content
KugelAudio Hero

What is KugelAudio?

KugelAudio is a state-of-the-art text-to-speech (TTS) platform designed for real-time applications. Whether you’re building voice agents, interactive applications, or content creation tools, KugelAudio provides the speed and quality you need.

Quick Start

Get up and running with KugelAudio in under 5 minutes

Generate Speech

Generate high-quality audio from text

Streaming

Real-time audio streaming for low latency

Voices

Browse voices and create custom clones

Key Features

Use kugel-3 for new integrations. See Models for its supported capabilities.
Existing model IDs remain accepted for backwards compatibility. See Models for the current and legacy IDs.
Stream audio chunks as they’re generated for the lowest possible latency. Perfect for LLM integrations where text arrives token by token.
Create custom voices from 10–30 seconds of clean reference audio. See Voice cloning for recording requirements.
Multilingual single model — 39 languages including DE, EN, FR, ES, IT, NL, PT, PL, RU, ZH, JA, KO, AR. See the TTS endpoint reference for the full list.

Getting Started

1

Get Your API Key

Sign up at kugelaudio.com and get your API key from the dashboard.
2

Install the SDK

Choose your preferred SDK and install it:
3

Generate Your First Audio

Need Help?

API Reference

Detailed API documentation with examples