Gemini 3.6 Flash

google/gemini-3.6-flash
Capabilities
8
VisionVideoAudioToolsSearchThinkJSONCache
Context length
1M
Tokens
Max output tokens
Tokens
Type
Chat
Google

Playground

Up to 50 messages are kept on this page; only the last 10 are sent as context, and refreshing clears them

Pricing

Pricing breakdown
Usage-based pricing, broken down by billing item at your current rates.

Input tokens

Standard rate
0.45 credits / 1M

Output tokens

Standard rate
2.25 credits / 1M

Cache read tokens

Standard rate
0.045 credits / 1M

Billing items are additive. Only one matching price tier applies to each item. Charges are calculated from actual usage per request, with each item rounded up.

API

cURL example
POST/v1beta/models/google/gemini-3.6-flash:generateContent
SynchronousGemini
curl "http://127.0.0.1:11113/v1beta/models/google/gemini-3.6-flash:generateContent" \  -X POST \  -H "x-goog-api-key: $API_KEY" \  -H "Content-Type: application/json" \  -d '{  "contents": [    {      "role": "user",      "parts": [        {          "text": "Explain model gateways in one sentence."        }      ]    }  ]}'

Set your API_KEY environment variable before running. Keep your key server-side.

API documentation

Uptime