Google: google/gemini-3.8-flash
Gemini 3.8 Flash
google/gemini-3.8-flashCapabilities
8
VisionVideoAudioToolsSearchThinkJSONCache
Context length
1M
Tokens
Max output tokens
—
Tokens
Type
Chat
Google
Playground
Pricing
Pricing breakdown
Usage-based pricing, broken down by billing item at your current rates.
Input tokens
Standard rate
0.525 credits / 1M
Output tokens
Standard rate
2.625 credits / 1M
Cache read tokens
Standard rate
0.053 credits / 1M
Billing items are additive. Only one matching price tier applies to each item. Charges are calculated from actual usage per request, with each item rounded up.
API
cURL example
request.sh
POST
/v1beta/models/google/gemini-3.8-flash:generateContentSynchronousGemini
curl "http://127.0.0.1:11113/v1beta/models/google/gemini-3.8-flash:generateContent" \ -X POST \ -H "x-goog-api-key: $API_KEY" \ -H "Content-Type: application/json" \ -d '{ "contents": [ { "role": "user", "parts": [ { "text": "Explain model gateways in one sentence." } ] } ]}'Set your API_KEY environment variable before running. Keep your key server-side.