Official Provider HubMountain View, CA, USA
Google Cloud / Gemini API Pricing & Token Economics
Google's Gemini ecosystem delivers native 1M+ token context windows capable of processing hours of audio, video frames, and massive code repositories. Gemini 3.8 Flash offers free prompt cache reads ($0.00/1M), while Gemini 3.1 Pro provides tiered pricing with long-context multipliers for prompts exceeding 200,000 tokens.
Prompt Caching Policy
Gemini 3.8 Flash offers 100% discount on cache hits ($0.00/1M). Storage incurs $0.50/1M tokens/hour promo fee.
Batch / Asynchronous Rules
Standard API execution rules. Refer to individual endpoint rate limits for asynchronous processing.
Special Billing Conditions
Gemini 3.1 Pro prompts > 200k tokens incur a 2.0x multiplier on input ($2.50/1M) and output ($10.00/1M).
Active Google Cloud / Gemini Models(4 models verified)
All 30 Models →| Model | Context | Input / 1M | Cached / 1M | Output / 1M | Batch Discount | Actions |
|---|---|---|---|---|---|---|
Gemini 3.8 Flash gemini-3.8-flash-001 | 1.05M | $0.75 | $0.075 | $3.75 | 50% off | Specs → |
Gemini 3.1 Pro Preview gemini-3.1-pro-preview | 1.05M | $2.00 | $0.20 | $12.00 | 50% off | Specs → |
Gemini 2.5 Flash-Lite gemini-2.5-flash-lite-001 | 1.05M | $0.10 | $0.025 | $0.40 | 50% off | Specs → |
Gemini 2.5 Flash gemini-2.5-flash | 1.05M | $0.075 | $0.019 | $0.30 | 50% off | Specs → |
Calculate your monthly Google Cloud / Gemini bill
Simulate prompt caching hit rates, batch discounts, and output token costs with zero server logging.