Ext: Grok Imagine 2.0 is a text-to-image model available through APIMart's asynchronous image generation API. Each request can produce 1–12 images in one of seven supported aspect ratios, with billing based on the number of images successfully delivered.

Image link valid for 72 hours
Transparent pricing with no hidden fees. Pay only for what you use.
| Variant | Tier | GoldCurrent price | Platinum | Diamond | Official | Your savings |
|---|---|---|---|---|---|---|
grok-imagine-2.0-ext | Image upload | Free | Free | Free | — | — |
| Standard Generation | 0.15Credits(~$0.015)/pic | 0.1425Credits(~$0.0143)/pic | 0.135Credits(~$0.0135)/pic | 0.6Credits(~$0.06)/pic | 75% | |
| Region edit | 0.15Credits(~$0.015)/gen | 0.1425Credits(~$0.0143)/gen | 0.135Credits(~$0.0135)/gen | 0.1875Credits(~$0.0188)/gen | 20% | |
grok-imagine-image-2.0 | 1K low | 0.32Credits(~$0.032)/pic | 0.304Credits(~$0.0304)/pic | 0.288Credits(~$0.0288)/pic | 0.4Credits(~$0.04)/pic | 20% |
| 1K medium | 0.48Credits(~$0.048)/pic | 0.456Credits(~$0.0456)/pic | 0.432Credits(~$0.0432)/pic | 0.6Credits(~$0.06)/pic | 20% | |
| 2K low | 0.48Credits(~$0.048)/pic | 0.456Credits(~$0.0456)/pic | 0.432Credits(~$0.0432)/pic | 0.6Credits(~$0.06)/pic | 20% | |
| 2K medium | 0.64Credits(~$0.064)/pic | 0.608Credits(~$0.0608)/pic | 0.576Credits(~$0.0576)/pic | 0.8Credits(~$0.08)/pic | 20% | |
| Reference image (each) | 0.08Credits(~$0.008)/pic | 0.076Credits(~$0.0076)/pic | 0.072Credits(~$0.0072)/pic | 0.1Credits(~$0.01)/pic | 20% |
* Actual costs are subject to final output.
Generate and edit images with explicit parameters
Generate visual assets from text for common creative and product workflows
Set up access, submit a text-to-image task, and poll for the result
Sign in to APIMart or create an account to access the dashboard.
Top up your account balance before sending production generation requests.
Generate an API key, submit grok-imagine-2.0-ext with a text prompt, and poll the returned task ID for image URLs.
Answers about inputs, output count, aspect ratios, tasks, quality, URLs, and billing
This page supports Ext text-to-image generation and ratio redraw, free layer retrieval from a completed source task or one uploaded image, and paid region editing. The official grok-imagine-image-2.0 model supports text-to-image generation and editing with up to three reference images.
Ext text-to-image accepts any integer n from 1 to 12, defaulting to 1; the Form offers 1, 4, 8 and 12. The official model supports 1–10 outputs. Ratio redraw and region editing each produce one image per request, regardless of the number of selected regions.
Ext supports 1:1, 2:3, 3:2, 3:4, 4:3, 9:16 and 16:9. The official model uses aspect_ratio and offers auto plus 13 ratios, including 1:2 and 2:1. Choose from the options for the active model.
POST the request to /v1/images/generations, save the returned task ID, and poll /v1/tasks/{task_id}. When the task is completed, read the successfully returned image URLs from the result.
Ext uses the fixed resolution=quality mode. The official model supports 1k/2k resolution and low/medium quality for text-to-image, with medium as the default. With reference images, omit quality and use the default editing tier. Ext result URLs expire after 72 hours; save results promptly.
RUN estimates use the selected model or operation’s original price and actual discount. Output fees scale with the requested image count; official reference images have their own price and discount and are charged once per input per request, without multiplying by output count. Region editing is charged once per completed task, not per selected region. Layer retrieval from a source task or one uploaded image is free. Final charges follow the completed task’s actual billing.
Yes. Prices shown on this page are Gold member prices, already 20% off. Platinum and Diamond members save even more, up to 28%, and you only pay for successful requests.
Sign up for APIMart, open the API Keys page in your dashboard and create a new key. The same key calls the xAI Grok Imagine 2.0 API and every other model on the platform, with pay-as-you-go billing and no subscription.
Explore more models in the same category.

Qwen Image 3.0
Qwen Image 3.0 is Alibaba's image model for text-to-image and editing with 1–3 reference images. It offers stable bilingual text rendering and dense layouts, in standard and Pro versions at 1K or 2K, up to 6 images per request.

Nano banana
Nano Banana (gemini-2.5-flash-image-preview) is Google DeepMind's fast, conversational image generation and editing model. It delivers natural-language text-to-image, precise multi-turn edits, strong character consistency, and multi-image fusion at low latency.

Seedream 5.0 Pro
Seedream 5.0 Pro (seedream-5-0-pro) is quality-first text-to-image model. It produces cinematic 1K and 2K images with best-in-class text rendering, supports up to 10 reference images for style and character consistency, and unifies generation and editing in a single API call.

Midjourney
Midjourney is a leading text-to-image AI model renowned for its painterly aesthetics, strong composition, and cinematic lighting. Generate high-quality images from text prompts or reference images, with fine-grained control over style, aspect ratio, and creative parameters — accessible programmatically through APIMart's unified API, no Discord required.