Base URL
Authentication
Include your API key in requests:Sending text safely
Tooling for WebSockets
The streaming endpoints are plain JSON-over-WebSocket. For interactive exploration usewscat
(npm install -g wscat) or websocat:
Endpoints
Generate Speech
REST one-shot — also the canonical request parameter reference
Stream Speech
One request, audio chunks streamed over a WebSocket
Stream Input
Token-by-token text input, turn-based sessions for LLM agents
Multi-Context
Up to 20 independent audio streams over one connection
Audio Formats
PCM, G.711 telephony codecs, chunk fields, AI-generated audio marking
Voices
List, inspect, and clone voices
Models
Inspect accepted model IDs and per-model input limits
Dictionaries
Manage project pronunciation dictionaries and entries