grok-imagine-2.0-extGrok Imagine 2.0 Ext is a text-to-image model available through APIMart's asynchronous image generation API. Each request can produce 1–12 images in one of seven supported aspect ratios, with billing based on the number of images successfully delivered.

Image link valid for 72 hours
Transparent pricing with no hidden fees. Pay only for what you use.
* Actual costs are subject to final output.
A focused text-to-image workflow with explicit generation parameters
Generate visual assets from text for common creative and product workflows
Set up access, submit a text-to-image task, and poll for the result
Sign in to APIMart or create an account to access the dashboard.
Top up your account balance before sending production generation requests.
Generate an API key, submit grok-imagine-2.0-ext with a text prompt, and poll the returned task ID for image URLs.
Answers about inputs, output count, aspect ratios, tasks, quality, URLs, and billing
It supports text-to-image generation only. Send a non-empty text prompt; image-to-image inputs, image URLs, and image-role inputs are not supported.
The n parameter accepts any integer from 1 to 12 and defaults to 1. A straightforward UI can offer the recommended choices 1, 4, 8, and 12.
The seven recommended ratio values are 1:1, 2:3, 3:2, 3:4, 4:3, 9:16, and 16:9. Unsupported values such as 1:2, 2:1, 4:5, or auto should not be sent.
POST the request to /v1/images/generations, save the returned task ID, and poll /v1/tasks/{task_id}. When the task is completed, read the successfully returned image URLs from the result.
Quality is fixed: use resolution=quality or omit it, rather than selecting 1K, 2K, or 4K tiers. Results are URL-only, and the URLs expire after 72 hours, so save them in time.
Billing is based on the actual number of images delivered successfully. If fewer images succeed than requested, only the successful images are charged; failed tasks are not charged.
Explore more models in the same category.

Nano banana
Nano Banana (gemini-2.5-flash-image-preview) is Google DeepMind's fast, conversational image generation and editing model. It delivers natural-language text-to-image, precise multi-turn edits, strong character consistency, and multi-image fusion at low latency.

Doubao Seedream 5.0 Pro
Seedream 5.0 Pro (doubao-seedream-5-0-pro) is ByteDance's quality-first text-to-image model. It produces cinematic 1K and 2K images with best-in-class text rendering, supports up to 10 reference images for style and character consistency, and unifies generation and editing in a single API call.

Midjourney
Midjourney is a leading text-to-image AI model renowned for its painterly aesthetics, strong composition, and cinematic lighting. Generate high-quality images from text prompts or reference images, with fine-grained control over style, aspect ratio, and creative parameters — accessible programmatically through APIMart's unified API, no Discord required.

Wan 2.7 Image
wan‑2‑7‑image is an advanced image generation model in Alibaba’s Wan 2.7 series designed to create high‑quality visuals from text prompts. It excels at generating detailed, realistic images with accurate object representation and rich composition. The model supports multimodal input (e.g., combining text with reference images) to influence style and structure, and it’s well‑suited for creative workflows such as marketing assets, product visuals, social media graphics, and artistic content production. With strong semantic understanding and prompt adherence, wan‑2‑7‑image delivers both fidelity and expressive visual output.