Models
Enterprise
Subscribe
Resource
Documentation
Console

google/gemini-3.1-flash-live-preview

Gemini.Color

Google: gemini-3.1-flash-live-preview

Gemini 3.1 Flash Live Preview is Google's low-latency, audio-to-audio model for real-time dialogue and voice-first applications.

Modetext → text
Input price$0.75 per M tokens
Output price$4.5 per M tokens
Context length
Weekly usage0 tokens
Listed at

Google

Latency
Throughput
Upload rate
Context length
Max output
Input price$0.75per 1M tokens
Output price$4.5per 1M tokens
Cache read$0per 1M tokens
Cache write$0per 1M tokens

Performance

Avg TPS
Avg latency
Avg success rate
TPS
TTFT
Latency
Success rate

Success rate

Speed

Code Example

Replace <YOUR_API_KEY> with the API key generated on your token management page.

Authentication

All requests must include the Authorization: Bearer <TOKEN> header.

Endpoint Path

POST

Supported Parameters

ParameterTypeDefaultDescription

Rate Limits

GroupRPMTPMRPD

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.

Pricing

GroupBilling typePrice summary