Skip to main content

Prerequisites

Before you begin, make sure you have:

Installation

Install the Python SDK using pip or uv:
Or with uv (recommended):

Basic Usage

Initialize the Client

Connect once at startup. Call connect() once (JavaScript, Java) or use await KugelAudio.create(...) (async Python) so the first request does not pay the connection handshake. See Latency.

Generate Speech

Examples below use kugel-3, the current model. See Models for details.

Stream Audio

For lower latency, stream audio chunks as they’re generated:
For async applications:

Working with Voices

The examples above use voice 1071. To pick another, list the voices your key can use. The listing returns all voices, one page at a time; pass language to get only the voices that speak it: de matches every German accent (de-DE, de-AT), de-AT only Austrian German. See Voices for pagination, voice details, and cloning.
Pick your voice deliberately. Voices differ in baseline energy, age, and warmth. Listen to several with representative text before choosing one. Building an LLM-driven voice agent? See Voice Agent Prompting.

Next Steps

Generate Speech

All generation options and parameters

Streaming

Real-time audio streaming techniques

Using Voices

Browse voices, paginate, and clone

Text Processing

Normalization and spell tags

Your SDK

Python

Python SDK reference

TypeScript/JavaScript

TypeScript/JavaScript SDK reference

Java

Java SDK reference
Last modified on September 23, 2026