Gemini 2.5 Pro: AI Large Language Models

Gemini 2.5 Pro is Google's advanced multimodal reasoning model, optimized for complex coding, logical tasks, and deep analysis. Supports text and image inputs. Token-based pricing: $1.25/M input tokens, $10.00/M output tokens. Two endpoints: standard async (/gemini-2-5-pro) and live streaming (/gemini-2-5-pro/stream) via SSE.

📝

Overview

About this model

Gemini 2.5 Pro is Google's advanced multimodal reasoning model, optimized for complex coding, logical tasks, and deep analysis. It delivers high-precision responses and handles extensive context. Pricing: $1.25 per million input tokens and $10.00 per million output tokens.

1Logical Reasoning: Solve complex reasoning tasks, math puzzles, and programming problems.
2Advanced Coding: Refactoring, debugging, code generation, and multi-file code editing.
3Multimodal Processing: Analyze images, diagrams, charts, and text in a single prompt.
4Detailed Synthesis: Generate long-form text, reports, and analytical breakdowns.
💰

Pricing & Value

Cost analysis

muapiapp$1.25/M input tokens, $10.00/M output tokens

Token-based billing. Minimum $0.00025 per call. Extremely competitive pricing.

Fal.ai$1.50/M input tokens, $12.00/M output tokens

muapiapp remains more cost-effective.

Replicate$1.60/M input tokens, $12.50/M output tokens

muapiapp remains more cost-effective.

** Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

Promptstring

The user message or instruction for the model.

Default ValueExplain quantum computing in simple terms.
Image URLstring

Optional image URL to include as multimodal input.

Default Valueundefined
System Promptstring

Optional system-level instruction to guide model behavior.

Default ValueYou are a helpful physics teacher.
📖

Implementation Guide

Developer documentation

Standard (Async)

POST /api/v1/gemini-2-5-pro — returns request_id, poll for result via /api/v1/predictions/{id}/result.

Streaming (SSE)

POST /api/v1/gemini-2-5-pro/stream — returns a live SSE stream. Each chunk: data: {"choices":[{"delta":{"content":"text"}}]}, ending with data: [DONE].

See Streaming Documentation for full code examples.

Common Questions

Frequently asked

How is pricing calculated?

Pricing is token-based: $1.25 per million input tokens and $10.00 per million output tokens. The minimum charge per call is $0.00025. Actual cost is deducted after each call based on token counts returned by the model.

Does Gemini 2.5 Pro support images?

Yes. Pass an image_url in the request body to include an image as part of the user message.

Does Gemini 2.5 Pro support system prompts?

Yes. You can supply an optional system_prompt field to guide the model's behavior and tone.