Grok 4.7 and Grok 4.6 are xAI's next two multimodal reasoning models in the Grok 4 series — adjustable reasoning effort, mixed text/image input, live web search, and both standard async and SSE-streaming endpoints. Both are coming soon on Muapi; build against the identical request shape used by the already-live Grok 4.3 and Grok 4.5 today. This page also covers the full Grok Imagine lineup — text-to-image, image-to-image, text-to-video, image-to-video, and video extension — all live now under one API key.
xAI's next-generation multimodal reasoning model. Adjustable reasoning effort, live web search, and streaming — coming soon.
xAI's upcoming multimodal reasoning model between Grok 4.5 and Grok 4.7. Adjustable reasoning effort and live web search — coming soon.
Live multimodal reasoning model with adjustable reasoning effort, integrated web search, and an SSE streaming endpoint.
Live multimodal reasoning model with mixed text/image input, adjustable reasoning depth, and web search tools.
High-quality text-to-image generation with vivid scenes, characters, and strong lighting. Returns 6 images per call.
High-fidelity text-to-image mode prioritizing accuracy and detail over speed, with sharper lighting and depth.
Edit an existing image with natural-language instructions while preserving scene structure, perspective, and lighting.
Fast, creative text-to-video generation with synchronized ambient audio, smooth motion, and expressive lighting.
Animate a still image into a cinematic video with synchronized ambient audio and fluid, realistic motion.
Preview of Grok Imagine's next video model — multiple aspect ratios and resolutions, durations up to 15 seconds.
Continue and expand an existing Grok Imagine video generation while maintaining visual style and motion.
Grok 4.7 and Grok 4.6 are the next two releases in xAI's Grok 4 reasoning line, following Grok 4.3 and Grok 4.5. Like the rest of the family, both are multimodal reasoning models — they accept mixed text and image input, expose an adjustable reasoning_effort parameter (low, medium, high, xhigh), and can call an integrated web-search tool for real-time information retrieval. Each model ships two endpoints: a standard async route that returns a request_id for polling, and a live /stream route that pushes Server-Sent Events for chat-style interfaces.
Both Grok 4.7 and Grok 4.6 are currently coming soon on Muapi — the request schema is locked in and mirrors the live Grok 4.3/4.5 endpoints exactly, so you can build and test your integration today and cut over the moment each model activates, with no code changes required.
Beyond the reasoning line, this page also covers Grok Imagine — xAI's separate image and video generation model family, spanning text-to-image, image-to-image, text-to-video, image-to-video, and video extension — all live today.
Grok 4.7 and 4.6 expose low/medium/high/xhigh reasoning_effort controls, trading latency for deeper step-by-step problem solving.
Pass an optional image_url alongside your prompt to analyze screenshots, diagrams, and photos in the same request.
Toggle web_search to let Grok 4.6 / 4.7 retrieve real-time information for current-events and fact-checking queries.
Every Grok 4 model ships both an async request/poll endpoint and an SSE /stream endpoint for real-time chat UIs.
Text-to-image and image-to-image with vivid scenes, characters, and strong lighting — 6 images per call from $0.05.
Text-to-video and image-to-video clips from 6–30 seconds with synchronized ambient audio, plus a video-extend endpoint.
| Model | Type | Status | Pricing | Best For |
|---|---|---|---|---|
| Grok 4.7 | Reasoning LLM | Coming Soon | $1.40/M in · $4.20/M out | Latest-generation reasoning, web search, multimodal input |
| Grok 4.6 | Reasoning LLM | Coming Soon | $1.50/M in · $4.50/M out | Mid-tier reasoning upgrade over Grok 4.5 |
| Grok 4.5 | Reasoning LLM | Live | $1.60/M in · $4.80/M out | Production reasoning today at the lowest live price |
| Grok 4.3 | Reasoning LLM | Live | $2.50/M in · $5.00/M out | Original Grok 4 reasoning endpoint |
| Grok Imagine | Image / Video | Live | $0.05–$0.64/call | Text/image-to-image and text/image-to-video generation |
Grok 4.7 is xAI's next-generation multimodal reasoning model, expected to succeed Grok 4.5 in the Grok 4 series. It supports adjustable reasoning effort, mixed text/image input, and integrated web search, with both a standard async endpoint and a live SSE streaming endpoint. It is coming soon on Muapi.
Grok 4.6 is xAI's upcoming multimodal reasoning model positioned between Grok 4.5 and Grok 4.7. It carries the same feature set — adjustable reasoning depth, image + text input, and web search — and is coming soon on Muapi.
Not yet — both are marked coming soon. Grok 4.3 and Grok 4.5 are live today with an identical request shape (prompt, image_url, system_prompt, reasoning_effort, web_search), so you can integrate against those now and switch over to Grok 4.6 or Grok 4.7 with zero code changes once they launch.
Both will be billed token-based, consistent with the rest of the Grok 4 family: Grok 4.6 at $1.50/M input and $4.50/M output tokens, and Grok 4.7 at $1.40/M input and $4.20/M output tokens, with a $0.0001 minimum charge per request.
Grok Imagine is xAI's separate image and video generation model family — not a reasoning LLM. It covers text-to-image, image-to-image, text-to-video, image-to-video, and video extension, and every Grok Imagine endpoint is live on Muapi today.
Sign up at muapi.ai, create an API key from your dashboard, and start calling any live Grok endpoint immediately — no waitlist or approval required. Coming-soon endpoints like Grok 4.6 and Grok 4.7 will activate under the same key the moment they launch.