Grok Imagine Quality is xAI's high-fidelity text-to-image mode that prioritizes accuracy and detail over speed. It produces sharper, more visually accurate images with stronger lighting, depth, and artistic clarity. Get 6 images each time.
About this model
Grok Imagine Quality is xAI's accuracy-first text-to-image mode. It runs the same Grok Imagine engine in pro/quality mode, prioritizing fidelity, sharper detail, and stronger lighting over generation speed. You still receive 6 unique images per generation across your chosen aspect ratio, but each image is rendered with greater visual precision — making it the right pick when the result needs to feel polished rather than fast.
Use Quality mode when you want concept art, marketing visuals, character designs, or detailed scenes that hold up under close inspection. The standard grok-imagine-text-to-image endpoint is the better fit when you want fast iteration; switch to this Quality variant when the output is the final deliverable.
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.05 | muapiapp offers Grok Imagine Quality mode at a competitive flat rate, typically 20–50% cheaper than equivalent quality-tier offerings elsewhere. |
| Fal.ai | Not available | Fal.ai does not currently expose Grok Imagine's quality mode as a managed endpoint. |
| Replicate | Not available | Replicate does not currently expose Grok Imagine's quality mode as a managed endpoint. |
muapiapp offers Grok Imagine Quality mode at a competitive flat rate, typically 20–50% cheaper than equivalent quality-tier offerings elsewhere.
Fal.ai does not currently expose Grok Imagine's quality mode as a managed endpoint.
Replicate does not currently expose Grok Imagine's quality mode as a managed endpoint.
* Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| Prompt | string | Text prompt describing the image. | A futuristic samurai standing under glowing neon lights in a rainy cyberpunk alley, reflections on wet pavement, dramatic rim lighting, highly detailed armor, cinematic atmosphere, ultra-realistic style. |
| Aspect Ratio | Enum (5 options) | Aspect ratio of the output image. Get 6 images each time. | 1:1 |
Text prompt describing the image.
A futuristic samurai standing under glowing neon lights in a rainy cyberpunk alley, reflections on wet pavement, dramatic rim lighting, highly detailed armor, cinematic atmosphere, ultra-realistic style.Aspect ratio of the output image. Get 6 images each time.
1:1Developer documentation
Prepare Your Input:
9:16, 16:9, 2:3, 3:2, or 1:1. Defaults to 1:1 if you omit it.Submit Your Request:
grok-imagine-text-to-image-quality. Quality mode is enabled automatically — you do not need to pass any extra flag.Receive Your Images:
Refine if Needed:
grok-imagine-text-to-image endpoint when you want faster iteration drafts.Frequently asked
This endpoint runs Grok Imagine in quality mode (`enable_pro=true`), which prioritizes accuracy, sharpness, and lighting fidelity over generation speed. The base endpoint runs the speed-optimized variant.
Each request returns 6 unique images, the same as the speed-mode endpoint.
9:16, 16:9, 2:3, 3:2, and 1:1. Defaults to 1:1 if not specified.
Quality mode costs $0.05 per generation, the same as the speed-mode endpoint.
Use Quality for final deliverables, hero shots, and detailed scenes that need to look polished. Use the standard endpoint for faster iteration and brainstorming drafts.