Audio model · Google

Gemini 3.1 Flash TTS

Speech generation with expressive delivery instructions

What is Gemini 3.1 Flash TTS?

Gemini 3.1 Flash TTS generates speech from text with selectable voices. Include inline delivery instructions and expressive tags such as [laughing] and [whispering] to guide the performance.

Capabilities

  • Generate speech from text

Gemini 3.1 Flash TTS on Melius

Melius runs Gemini 3.1 Flash TTS alongside every other major image, video, audio, and text model, so you can use it alongside other models in your workflow — all on one infinite canvas. You only pay for what you generate.

Frequently asked questions

What is Gemini 3.1 Flash TTS?

Gemini 3.1 Flash TTS generates speech from text with selectable voices. Include inline delivery instructions and expressive tags such as [laughing] and [whispering] to guide the performance.

How do I use Gemini 3.1 Flash TTS on Melius?

Melius is a node-based canvas: add a audio node, pick Gemini 3.1 Flash TTS from the model picker, and connect it to whatever should feed it or follow it. You only pay for what you generate.

Start creating with Gemini 3.1 Flash TTS on Melius.