Gemini 3.6 Flash API Pricing
Gemini 3.6 Flash is a high-speed, multimodal language model built for real-time text generation, supporting text and image inputs natively. Token-based pricing: .60/M input tokens and .60/M output tokens. Two endpoints: standard async (/gemini-3-6-flash) and live streaming (/gemini-3-6-flash/stream) via SSE.
About Gemini 3.6 Flash
Gemini 3.6 Flash is Google's next-generation high-speed multimodal language model built for real-time text generation, instruction following, and advanced image understanding. It delivers low-latency performance for conversational AI, real-time customer support, document parsing, and vision reasoning tasks. Token-based pricing is $0.60/M input tokens and $3.60/M output tokens.
Interactive Savings Calculator
Estimate monthly API spend and compare absolute developer savings.
$6000.00
$0.60/M input, $3.60/M output tokens$750.00
$0.075/M input, $0.30/M output tokens (under 128k context)Detailed Pricing Breakdown
| Provider | Estimated Rate | Notes |
|---|---|---|
| muapiapp | $0.60/M input, $3.60/M output tokens | Fast, high-quality, token-based pricing with an upfront minimum of $0.0001. |
| Google (official) | $0.075/M input, $0.30/M output tokens (under 128k context) | Official API pricing. We scale our rates to match standard Gemini rates with preserved developer margin. |
Developer Integration Snippets
Model FAQ
Compare similar models
Ready to scale your production?
Get instant access to developer keys. Integrate high-speed dynamic models in minutes with our robust SDKs.

