Gemini 3.5 Flash (OpenAI) API Pricing
Gemini 3.5 Flash (OpenAI-compatible) is a high-speed, multimodal language model built for real-time text generation, supporting text and image inputs natively. Token-based pricing: $0.60/M input tokens and $3.60/M output tokens. Two endpoints: standard async (/gemini-3-5-flash-openai) and live streaming (/gemini-3-5-flash-openai/stream) via SSE.
About Gemini 3.5 Flash (OpenAI)
Gemini 3.5 Flash (OpenAI-compatible) is a high-speed, multimodal language model optimized for rapid text generation and real-time image understanding, accessed via an OpenAI-compatible API interface. Token-based pricing: $0.60/M input tokens and $3.60/M output tokens.
Interactive Savings Calculator
Estimate monthly API spend and compare absolute developer savings.
$6000.00
$0.60/M input, $3.60/M output tokens$750.00
$0.075/M input, $0.30/M output tokens (under 128k context)Detailed Pricing Breakdown
| Provider | Estimated Rate | Notes |
|---|---|---|
| muapiapp | $0.60/M input, $3.60/M output tokens | Fast, high-quality, token-based pricing with an upfront minimum of $0.0001. |
| Google (official) | $0.075/M input, $0.30/M output tokens (under 128k context) | Official API pricing. We scale our rates to match standard Gemini rates with preserved developer margin. |
Developer Integration Snippets
Model FAQ
Compare similar models
Ready to scale your production?
Get instant access to developer keys. Integrate high-speed dynamic models in minutes with our robust SDKs.

