Audio
Gemini 3.1 Flash Tts
Newest audio model from Google introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.
- Text to speech
Sign in to create. Review the credit cost in the studio before running.
Catalog preview for Gemini 3.1 Flash Tts. Review your own result before publishing.
Use Gemini 3.1 Flash Tts through the API.
Start with this model ID for the endpoint featured on this page:
gemini-3.1-flash-tts
Check the current model schema for supported inputs and options. Other tasks may use different model IDs.
From quote to result.
Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.
Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.