# Gradium API > Developer documentation for Gradium's low-latency Text-to-Speech and Speech-to-Text APIs, including SDK guides, WebSocket streaming, REST endpoints, voices, and pronunciation tools. - [Introduction](https://docs.gradium.ai/guides/introduction.md): Low-latency, high-quality Text-to-Speech and Speech-to-Text API - [Installation](https://docs.gradium.ai/guides/installation.md): Install the Gradium Python SDK and get started - [Text-to-Speech Overview](https://docs.gradium.ai/guides/text-to-speech-overview.md): When to use the WebSocket API vs the REST endpoint, and how the pieces fit together - [Text-to-Speech (WebSocket)](https://docs.gradium.ai/guides/text-to-speech.md): Convert text to natural-sounding speech via the Gradium SDK and WebSocket - [Text-to-Speech (REST)](https://docs.gradium.ai/guides/text-to-speech-rest.md): One-shot synthesis of a complete text block via HTTP POST - [Voice Settings](https://docs.gradium.ai/guides/voice-settings.md): Fine-tune TTS output: speed, temperature, voice similarity, rewrite rules - [Speech-to-Text Overview](https://docs.gradium.ai/guides/speech-to-text-overview.md): When to use the WebSocket API vs the REST endpoint, and how the pieces fit together - [Speech-to-Text (WebSocket)](https://docs.gradium.ai/guides/speech-to-text.md): Real-time speech-to-text streaming over WebSocket with semantic VAD and flush - [Speech-to-Text (REST)](https://docs.gradium.ai/guides/speech-to-text-rest.md): One-shot transcription of complete audio files via HTTP POST - [Transcription Settings](https://docs.gradium.ai/guides/transcription-settings.md): Fine-tune STT: language, temperature, latency vs. quality tradeoffs - [Speech-to-Speech Overview](https://docs.gradium.ai/guides/speech-to-speech-overview.md): What Speech-to-Speech does and how to set it up - [Speech-to-Speech (WebSocket)](https://docs.gradium.ai/guides/speech-to-speech.md): Real-time speech translation over WebSocket: transcribe, translate, and re-synthesize - [Speech-to-Speech Settings](https://docs.gradium.ai/guides/speech-to-speech-settings.md): Configure S2S: voice, translation target, inner models, and audio formats - [WebSocket Lifecycle](https://docs.gradium.ai/guides/websocket-lifecycle.md): Connection setup, ready messages, input, flush, end-of-stream, multiplexing, and errors - [Browser WebSockets](https://docs.gradium.ai/guides/browser-websockets.md): Use short-lived tokens for browser and mobile WebSocket clients without exposing API keys - [Multiplexing](https://docs.gradium.ai/guides/multiplexing.md): Send multiple independent requests over a single WebSocket connection - [WebSocket Stream Options](https://docs.gradium.ai/guides/websocket-stream-options.md): Connection initialisation flags shared by tts_realtime and stt_realtime - [LLM Tokens to Streaming TTS](https://docs.gradium.ai/guides/recipes/llm-to-tts.md): Stream generated text into TTS while preserving prosody and low latency - [Browser Microphone to STT](https://docs.gradium.ai/guides/recipes/browser-microphone-stt.md): Capture microphone audio in a browser and stream it to Gradium STT - [Telephony Audio Formats](https://docs.gradium.ai/guides/recipes/telephony-audio.md): Use mu-law, A-law, and low-sample-rate PCM with Gradium voice APIs - [Turn-Taking with Semantic VAD](https://docs.gradium.ai/guides/recipes/turn-taking.md): Use STT semantic VAD and adaptive delay to decide when a user has finished speaking - [Keyword Boosting](https://docs.gradium.ai/guides/recipes/keyword-boosting.md): Recognize names, brands, and jargon by biasing STT toward a custom dictionary - [Migrate to Gradium](https://docs.gradium.ai/guides/migration/index.md): Move existing voice API integrations to Gradium with minimal code changes. - [Migrate from ElevenLabs](https://docs.gradium.ai/guides/migration/elevenlabs.md): Replace ElevenLabs TTS calls with Gradium REST and WebSocket endpoints. - [Migrate from Cartesia](https://docs.gradium.ai/guides/migration/cartesia.md): Move Cartesia TTS and STT integrations to Gradium by changing endpoints, auth, and request fields. - [Migrate from Deepgram](https://docs.gradium.ai/guides/migration/deepgram.md): Replace Deepgram speech integrations with Gradium STT and TTS endpoints. - [Migrate from Fish Audio](https://docs.gradium.ai/guides/migration/fish-audio.md): Move Fish Audio TTS, voice cloning, and ASR integrations to Gradium endpoints, x-api-key auth, and JSON WebSocket messages. - [Overview](https://docs.gradium.ai/guides/voices/overview.md): Explore the Gradium voice library - [Flagship Voices](https://docs.gradium.ai/guides/voices/flagship-voices.md): Our top curated voices across multiple languages - [Custom Voices](https://docs.gradium.ai/guides/voices/custom-voices.md): Clone your own voice using a short audio sample - [Voice Design](https://docs.gradium.ai/guides/voices/voice-design.md): Create a voice from a written description, audition it, and convert it into a voice you can use everywhere - [Manage Voices](https://docs.gradium.ai/guides/voices/manage-voices.md): List, update, and delete your custom voices - [Text-to-Speech on Baseten](https://docs.gradium.ai/guides/self-hosted/baseten-tts.md): Stream Gradium TTS from a dedicated deployment hosted on Baseten - [Credits](https://docs.gradium.ai/guides/credits.md): Monitor and manage your API credit balance - [Text Rewriting Rules](https://docs.gradium.ai/guides/text-rewriting.md): Normalize and expand text patterns for better TTS pronunciation - [Limits](https://docs.gradium.ai/guides/limits.md): Session duration, character limits, audio formats, and concurrency - [Data Residency](https://docs.gradium.ai/guides/data-residency.md): Pin your speech sessions to the EU or the US with eu.api.gradium.ai and us.api.gradium.ai - [Errors](https://docs.gradium.ai/guides/errors.md): How errors are reported across REST, WebSocket, and streamed responses - [FAQ](https://docs.gradium.ai/guides/faq.md): Frequently asked questions about the Gradium API - [Integrations](https://docs.gradium.ai/integrations/index.md): Connect Gradium voice applications to tools, data sources, and agent workflows - [Gradbot](https://docs.gradium.ai/integrations/agent-frameworks/gradbot.md): Prototype voice agents quickly with Gradium APIs - [LiveKit](https://docs.gradium.ai/integrations/agent-frameworks/livekit.md): Use Gradium speech models in LiveKit Agents - [OpenClaw](https://docs.gradium.ai/integrations/agent-frameworks/openclaw.md): Use Gradium text-to-speech in OpenClaw agents - [Pipecat](https://docs.gradium.ai/integrations/agent-frameworks/pipecat.md): Use Gradium streaming text-to-speech in Pipecat voice agents - [Vapi](https://docs.gradium.ai/integrations/agent-frameworks/vapi.md): Use Gradium speech models in Vapi voice agents - [Keenable](https://docs.gradium.ai/integrations/web-search/keenable.md): Use Keenable realtime web search with Gradium voice AI agents - [Linkup](https://docs.gradium.ai/integrations/web-search/linkup.md): Use Linkup web search with Gradium voice AI agents - [Tavily](https://docs.gradium.ai/integrations/web-search/tavily.md): Use Tavily web search with Gradium voice AI agents - [API Reference](https://docs.gradium.ai/api-reference/introduction.md): Gradium API endpoints for Text-to-Speech, Speech-to-Text, Voices, Voice Design, Pronunciations, and Metering - [TTS WebSocket Stream](https://docs.gradium.ai/api-reference/endpoint/tts-websocket.md): Stream real-time Gradium text-to-speech audio over WebSocket. - [TTS POST Endpoint](https://docs.gradium.ai/api-reference/endpoint/tts-post.md): Generate text-to-speech audio from a complete text block with HTTP POST. - [STT WebSocket Stream](https://docs.gradium.ai/api-reference/endpoint/stt-websocket.md): Stream audio to Gradium speech-to-text over WebSocket for real-time transcription. - [STT POST Endpoint](https://docs.gradium.ai/api-reference/endpoint/stt-post.md): Transcribe a complete audio file with Gradium speech-to-text over HTTP POST. - [S2S WebSocket Stream](https://docs.gradium.ai/api-reference/endpoint/s2s-websocket.md): Stream audio in and audio out over a single Gradium speech-to-speech WebSocket: transcribe, optionally translate, and re-synthesize in real time. - [Create Voice](https://docs.gradium.ai/api-reference/endpoint/create-voice.md): Create a custom Gradium voice from an uploaded audio sample. - [Get Voices](https://docs.gradium.ai/api-reference/endpoint/get-voices.md): List Gradium voices available to the authenticated organization. - [Get Voice](https://docs.gradium.ai/api-reference/endpoint/get-voice.md): Retrieve metadata for a Gradium voice by voice UID. - [Update Voice](https://docs.gradium.ai/api-reference/endpoint/update-voice.md): Update metadata for an existing custom Gradium voice. - [Delete Voice](https://docs.gradium.ai/api-reference/endpoint/delete-voice.md): Delete a custom Gradium voice by voice UID. - [Generate Voice Candidates](https://docs.gradium.ai/api-reference/endpoint/generate-voice.md): Sample candidate voices from a written description. - [List Voice Candidates](https://docs.gradium.ai/api-reference/endpoint/list-voice-embeddings.md): Check whether a candidate is ready, or list your candidates. - [Create Voice From Candidate](https://docs.gradium.ai/api-reference/endpoint/create-voice-from-embedding.md): Keep a generated candidate as a permanent Gradium voice. - [Delete Voice Candidate](https://docs.gradium.ai/api-reference/endpoint/delete-voice-embedding.md): Remove a generated candidate you are not keeping. - [Create Pronunciation Dictionary](https://docs.gradium.ai/api-reference/endpoint/create-pronunciation.md): Create a Gradium pronunciation dictionary with language-specific rewrite rules. - [List Pronunciation Dictionaries](https://docs.gradium.ai/api-reference/endpoint/list-pronunciations.md): List pronunciation dictionaries for the authenticated Gradium organization. - [Get Pronunciation Dictionary](https://docs.gradium.ai/api-reference/endpoint/get-pronunciation.md): Retrieve a Gradium pronunciation dictionary and its rewrite rules by UID. - [Update Pronunciation Dictionary](https://docs.gradium.ai/api-reference/endpoint/update-pronunciation.md): Update a Gradium pronunciation dictionary and its rewrite rules. - [Delete Pronunciation Dictionary](https://docs.gradium.ai/api-reference/endpoint/delete-pronunciation.md): Delete a Gradium pronunciation dictionary by UID. - [Get Credits](https://docs.gradium.ai/api-reference/endpoint/get-credits.md): Get the current Gradium credit balance for the authenticated subscription. - [September 2026](https://docs.gradium.ai/guides/release-notes/2026-09.md): Gradium releases from September 2026 - [August 2026](https://docs.gradium.ai/guides/release-notes/2026-08.md): Gradium releases from August 2026 - [July 2026](https://docs.gradium.ai/guides/release-notes/2026-07.md): Gradium releases from July 2026 - [June 2026](https://docs.gradium.ai/guides/release-notes/2026-06.md): Gradium releases from June 2026 - [Earlier releases](https://docs.gradium.ai/guides/release-notes/2026-02.md): Gradium releases from early 2026 - [Gradium API Documentation](https://docs.gradium.ai/index.md): Build voice apps with Gradium's low-latency Text-to-Speech and Speech-to-Text APIs, SDK guides, streaming examples, voices, and pronunciation tools. ## OpenAPI Specs - [openapi](/api-reference/openapi.json) ## Optional - [Support](mailto:support@gradium.ai)