Black Forest Labs 的图像生成与编辑模型。可将局部修改和边界框布局控制用于创作,并通过逐步审阅确认结果。
了解图像生成、局部编辑和布局控制,并参考包含人工审阅的创作流程。
APIMart 正在准备接入 FLUX 3 Image API,目前尚未开放该模型的调用。集成就绪后,本页将公布支持的模型 ID、请求格式、限制、价格与访问方式。
图像生成与编辑能力,以及面向创作团队的流程建议
结合人工审阅与批准的创作场景建议
了解图像与视频的区别、编辑方式、布局控制及 APIMart 接入状态
FLUX 3 Image 是 Black Forest Labs 的图像生成与编辑模型。本页介绍的是静态图像能力,并非视频生成服务。
局部编辑可以修改指定部分,同时保留其他区域。实际使用时,建议将结果与原图对照,确认需要保留的细节符合预期。
可以通过边界框指定元素的放置区域,表达布局要求。生成后仍建议人工检查整体构图以及元素之间的位置关系。
可以考虑用于商品视觉构思、广告素材方案、故事板和设计方案比较。对外使用前,应检查内容准确性以及是否适合具体用途。
还不能。APIMart 正在准备集成,目前尚未提供 FLUX 3 Image API 调用。
APIMart 的价格、支持参数、限制和请求示例,请以接入后公布的信息为准。本页暂未公布开放日期。
探索同类型的其他模型。

Nano banana
Nano Banana (gemini-2.5-flash-image-preview) is Google DeepMind's fast, conversational image generation and editing model. It delivers natural-language text-to-image, precise multi-turn edits, strong character consistency, and multi-image fusion at low latency.

Seedream 5.0 Pro
Seedream 5.0 Pro (seedream-5-0-pro) is quality-first text-to-image model. It produces cinematic 1K and 2K images with best-in-class text rendering, supports up to 10 reference images for style and character consistency, and unifies generation and editing in a single API call.

Midjourney
Midjourney is a leading text-to-image AI model renowned for its painterly aesthetics, strong composition, and cinematic lighting. Generate high-quality images from text prompts or reference images, with fine-grained control over style, aspect ratio, and creative parameters — accessible programmatically through APIMart's unified API, no Discord required.

Wan 2.7 Image
wan‑2‑7‑image is an advanced image generation model in Alibaba’s Wan 2.7 series designed to create high‑quality visuals from text prompts. It excels at generating detailed, realistic images with accurate object representation and rich composition. The model supports multimodal input (e.g., combining text with reference images) to influence style and structure, and it’s well‑suited for creative workflows such as marketing assets, product visuals, social media graphics, and artistic content production. With strong semantic understanding and prompt adherence, wan‑2‑7‑image delivers both fidelity and expressive visual output.