Cartesia Sonic Text-to-Speech API
The Sonic text-to-speech API converts text into ultra-low-latency, emotive speech with sub-100ms time-to-first-byte. It supports REST, server-sent events, and WebSocket streaming for real-time voice agents and applications.
Cartesia Sonic Text-to-Speech API is one of 6 APIs that Cartesia publishes on the APIs.io network, described by an AsyncAPI event-driven specification.
Tagged areas include TTS, Streaming, SSE, WebSocket, and Real-Time. The published artifact set on APIs.io includes API documentation, a getting-started guide, an API reference, an AsyncAPI specification, a GitHub repository, and pricing.
Requests are made against the base URL https://api.cartesia.ai.
Metadata
The identity and technical contract details declared by the specification.
Specification
The full machine-readable OpenAPI contract behind this narrative.
Source
More from Cartesia 5
Other APIs Cartesia publishes across the network.