Gemini Flash TTS API on KeyoAPI — Low-Latency Speech
Gemini 3.1 Flash TTS Preview for controllable, low-latency speech on the same OpenAI-compatible key as chat.
Overview
gemini-3.1-flash-tts-preview is Gemini Flash text-to-speech on KeyoAPI. Indicative rates about $0.40 / $8.00 per 1M in/out — confirm live on /pricing/gemini-3.1-flash-tts-preview. Call POST /v1/audio/speech at https://www.keyoapi.xyz/v1 and receive audio in the same response (sync). Use IndexTTS-2 or CosyVoice3 when you already run async TTS jobs.
Quick start
Base URL: https://www.keyoapi.xyz/v1
Endpoint: POST /v1/audio/speech
Model: model=gemini-3.1-flash-tts-preview
curl https://www.keyoapi.xyz/v1/audio/speech \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"gemini-3.1-flash-tts-preview","input":"Hello from Gemini Flash TTS.","voice":"Kore"}' \
--output speech.wav
Auth uses the same Bearer API key as chat. Full notes: Keyo docs.
FAQ
Sync or async?
Sync: POST /v1/audio/speech returns audio. Async TTS on Keyo remains Qwen3-TTS / CosyVoice3.
Where is live pricing?
See /pricing/gemini-3.1-flash-tts-preview. Indicative ~$0.40 / $8.00 per 1M in/out.