
Image link valid for 72 hours
숨겨진 수수료 없는 투명한 요금제. 사용한 만큼만 지불하세요.
* 실제 비용은 최종 출력에 따라 달라집니다.
명확한 생성 매개변수로 구성된 텍스트 이미지 전용 워크플로
일반적인 크리에이티브 및 제품 워크플로에 필요한 시각 자료를 텍스트로 생성
접근 권한을 설정하고 텍스트 이미지 작업을 제출한 뒤 결과를 폴링하세요
APIMart에 로그인하거나 계정을 만들어 대시보드에 접속하세요.
실제 생성 요청을 보내기 전에 계정 잔액을 충전하세요.
API 키를 만든 뒤 텍스트 프롬프트와 grok-imagine-2.0-ext를 제출하고, 반환된 작업 ID로 이미지 URL을 폴링하세요.
입력, 생성 수, 화면 비율, 작업, 품질, URL 및 과금에 대한 안내
텍스트 이미지 생성만 지원합니다. 비어 있지 않은 텍스트 프롬프트를 보내야 하며, 이미지 기반 생성, 이미지 URL, 역할이 지정된 이미지 입력은 지원하지 않습니다.
n 매개변수에는 1에서 12 사이의 정수를 사용할 수 있으며 기본값은 1입니다. 간단한 UI에서는 권장 옵션인 1, 4, 8, 12를 제공할 수 있습니다.
권장 비율 값은 1:1, 2:3, 3:2, 3:4, 4:3, 9:16, 16:9입니다. 1:2, 2:1, 4:5, auto처럼 지원하지 않는 값은 보내지 마세요.
/v1/images/generations에 POST 요청을 보내고 반환된 작업 ID를 저장한 뒤 /v1/tasks/{task_id}를 폴링하세요. 작업이 완료되면 결과에서 성공적으로 반환된 이미지 URL을 읽을 수 있습니다.
품질은 고정입니다. resolution=quality를 사용하거나 생략하며, 1K, 2K, 4K 등급은 선택할 수 없습니다. 결과는 URL로만 제공되고 72시간 후 만료되므로 제때 저장하세요.
실제로 전달에 성공한 이미지 수에 따라 과금됩니다. 요청한 수보다 적은 이미지만 성공하면 성공한 이미지에 대해서만 과금되며, 실패한 작업에는 과금되지 않습니다.
같은 카테고리의 다른 모델을 탐색하세요.

Nano banana
Nano Banana (gemini-2.5-flash-image-preview) is Google DeepMind's fast, conversational image generation and editing model. It delivers natural-language text-to-image, precise multi-turn edits, strong character consistency, and multi-image fusion at low latency.

Doubao Seedream 5.0 Pro
Seedream 5.0 Pro (doubao-seedream-5-0-pro) is ByteDance's quality-first text-to-image model. It produces cinematic 1K and 2K images with best-in-class text rendering, supports up to 10 reference images for style and character consistency, and unifies generation and editing in a single API call.

Midjourney
Midjourney is a leading text-to-image AI model renowned for its painterly aesthetics, strong composition, and cinematic lighting. Generate high-quality images from text prompts or reference images, with fine-grained control over style, aspect ratio, and creative parameters — accessible programmatically through APIMart's unified API, no Discord required.

Wan 2.7 Image
wan‑2‑7‑image is an advanced image generation model in Alibaba’s Wan 2.7 series designed to create high‑quality visuals from text prompts. It excels at generating detailed, realistic images with accurate object representation and rich composition. The model supports multimodal input (e.g., combining text with reference images) to influence style and structure, and it’s well‑suited for creative workflows such as marketing assets, product visuals, social media graphics, and artistic content production. With strong semantic understanding and prompt adherence, wan‑2‑7‑image delivers both fidelity and expressive visual output.