<br />

Gemini 3.1 Flash Live Preview is a legacy preview model for real-time dialogue
and voice-first AI applications.

> [!NOTE]
> **Note:** We recommend updating to [Gemini 3.8 Live](https://ai.google.dev/gemini-api/docs/models/gemini-3.8-live) as the default option for most low-latency voice agent experiences. See the [migration guide](https://ai.google.dev/gemini-api/docs/models/gemini-3.8-live#migrating).

[Try in Google AI Studio](https://aistudio.google.com/live?model=gemini-3.1-flash-live-preview)

## Documentation

Visit the [Live API](https://ai.google.dev/gemini-api/docs/live-api) guide for full coverage
of features and capabilities.

## gemini-3.1-flash-live-preview

| Property | Description |
|---|---|
| Model code | `gemini-3.1-flash-live-preview` |
| Supported data types | **Inputs** Text, images, audio, video **Output** Text and audio |
| Token limits^[\[\*\]](https://ai.google.dev/gemini-api/docs/tokens)^ | **Input token limit** 131,072 **Output token limit** 65,536 |
| Capabilities | **[Audio generation](https://ai.google.dev/gemini-api/docs/speech-generation)** Supported **[Caching](https://ai.google.dev/gemini-api/docs/caching)** Not supported **[Code execution](https://ai.google.dev/gemini-api/docs/code-execution)** Not supported **[File search](https://ai.google.dev/gemini-api/docs/file-search)** Not Supported **[Function calling](https://ai.google.dev/gemini-api/docs/function-calling)** Supported **[Grounding with Google Maps](https://ai.google.dev/gemini-api/docs/maps-grounding)** Not supported **[Image generation](https://ai.google.dev/gemini-api/docs/image-generation)** Not supported **[Live API](https://ai.google.dev/gemini-api/docs/live-api)** Supported **[Search grounding](https://ai.google.dev/gemini-api/docs/google-search)** Supported **[Structured outputs](https://ai.google.dev/gemini-api/docs/structured-output)** Not supported **[Thinking](https://ai.google.dev/gemini-api/docs/thinking)** Supported **[URL context](https://ai.google.dev/gemini-api/docs/url-context)** Not supported |
| Consumption options | **[Batch API](https://ai.google.dev/gemini-api/docs/batch-api)** Not supported |
| Versions | Read the [model version patterns](https://ai.google.dev/gemini-api/docs/models/gemini#model-versions) for more details. - Preview: `gemini-3.1-flash-live-preview` |
| Latest update | March 2026 |
| Model card | [Model card](https://deepmind.google/models/model-cards/gemini-3-1-flash-audio/) |

## Migrating from Gemini 2.5 Flash Live

Gemini 3.1 Flash Live Preview is optimized for low-latency, real-time dialogue.
When migrating from `gemini-2.5-flash-native-audio-preview-12-2025`, consider
the following:

- **Model string** : Update your model string from `gemini-2.5-flash-native-audio-preview-12-2025` to `gemini-3.1-flash-live-preview`.
- **Thinking configuration** : Gemini 3.1 uses `thinkingLevel` (with settings like `minimal`, `low`, `medium`, and `high`) instead of `thinkingBudget`. The default is `minimal` to optimize for lowest latency. See [Thinking levels and budgets](https://ai.google.dev/gemini-api/docs/thinking#levels-budgets).
- **Server events** : A single [`BidiGenerateContentServerContent`](https://ai.google.dev/api/live#bidigeneratecontentservercontent) event can now contain multiple content parts simultaneously (for example, audio chunks and transcript). Update your code to process all parts in each event to avoid missing content.
- **Client content** : `send_client_content` is supported throughout the entire session lifecycle with explicit roles (`user` or `model`). Setting `turn_complete=true` unconditionally interrupts active model generation. See [Incremental content updates](https://ai.google.dev/gemini-api/docs/live-guide#incremental-updates).
- **Turn coverage** : Defaults to [`TURN_INCLUDES_AUDIO_ACTIVITY_AND_ALL_VIDEO`](https://ai.google.dev/api/live#turncoverage) instead of `TURN_INCLUDES_ONLY_ACTIVITY`. The model's turn now includes detected audio activity and all video frames. If your application currently sends a constant stream of video frames, you may want to update your application to only send video frames when there is audio activity to avoid incurring additional costs.
- **Async function calling** : Not yet supported. Function calling is synchronous only. The model will not start responding until you've sent the tool response. See [Async function calling](https://ai.google.dev/gemini-api/docs/live-tools#async-function-calling).
- **Proactive audio and affective dialogue** : These features are not yet supported in Gemini 3.1 Flash Live. Remove any configuration for these features from your code. See [Proactive audio](https://ai.google.dev/gemini-api/docs/live-guide#proactive-audio) and [Affective dialogue](https://ai.google.dev/gemini-api/docs/live-guide#affective-dialog).

For a detailed feature comparison, see the
[Model comparison](https://ai.google.dev/gemini-api/docs/live-guide#model-comparison) table in the
capabilities guide.