How to Add Text-to-Speech to an App with the Fish TTS API
Shipping voice features is often less about picking a model and more about the details around audio format, latency, retries, and voice reuse. This guide walks through a practical Fish TTS API integration using the documented Ace Data Cloud endpoint for text-to-speech, saved voices, and one-time instant voice cloning. What you can do The Fish TTS API exposes one main endpoint: POST https://api.acedata.cloud/fish/tts With that endpoint, you can build several common product flows: Generate an mp3 voiceover from plain text. Return wav or pcm when later processing needs a WAV container. Use a reusable cloned or public voice through reference_id . Use one-time instant voice cloning through references . Adjust speech with prosody.speed and prosody.volume . Move long-running synthesis behind a webhook using callback_url . How it works Authentication uses an authorization header with Bearer {token} , and the request body is JSON. The required body field is text , a non-empt...