Audio model · Google
Gemini 3.1 Flash TTS
Speech generation with expressive delivery instructions
What is Gemini 3.1 Flash TTS?
Gemini 3.1 Flash TTS generates speech from text with selectable voices. Include inline delivery instructions and expressive tags such as [laughing] and [whispering] to guide the performance.
Capabilities
- Generate speech from text
Gemini 3.1 Flash TTS on Melius
Melius runs Gemini 3.1 Flash TTS alongside every other major image, video, audio, and text model, so you can use it alongside other models in your workflow — all on one infinite canvas. You only pay for what you generate.
Frequently asked questions
What is Gemini 3.1 Flash TTS?
Gemini 3.1 Flash TTS generates speech from text with selectable voices. Include inline delivery instructions and expressive tags such as [laughing] and [whispering] to guide the performance.
How do I use Gemini 3.1 Flash TTS on Melius?
Melius is a node-based canvas: add a audio node, pick Gemini 3.1 Flash TTS from the model picker, and connect it to whatever should feed it or follow it. You only pay for what you generate.