Black Forest Labs의 이미지 생성 및 편집 모델입니다. 선택 영역 수정과 바운딩 박스를 활용한 배치 제어를 검토 중심의 제작 흐름에 적용할 수 있습니다.
이미지 생성, 부분 편집 및 배치 제어를 활용하는 제작 흐름을 살펴보세요.
APIMart는 FLUX 3 Image API 연동을 준비하고 있습니다. 현재 APIMart에서는 이 모델을 호출할 수 없습니다. 지원 모델 ID, 요청 형식, 제한, 요금 및 이용 방법은 연동 준비가 완료되면 안내합니다.
이미지 생성 및 편집 기능과 제작팀을 위한 작업 방식
사람의 검토 및 승인을 결합한 제작 활용 예시
이미지 모델의 용도, 편집, 배치 제어 및 APIMart 제공 상태
FLUX 3 Image는 Black Forest Labs의 이미지 생성 및 편집 모델입니다. 이 페이지는 정지 이미지를 다루는 모델을 소개하며, 동영상 생성 서비스를 안내하는 페이지가 아닙니다.
부분 편집으로 다른 영역을 유지하면서 지정한 부분을 수정할 수 있습니다. 실제 작업에서는 원본과 비교해 보존하려는 세부 요소가 유지됐는지 확인하세요.
바운딩 박스로 요소를 배치할 영역을 지정할 수 있습니다. 완성된 이미지의 구성과 요소 간 위치 관계는 사람이 확인하는 것이 좋습니다.
제품 이미지 구상, 광고 시안, 스토리보드, 디자인안 비교 등에 활용을 검토할 수 있습니다. 공개 전에는 내용의 정확성과 용도에 대한 적합성을 확인하세요.
아직 이용할 수 없습니다. APIMart는 연동을 준비 중이며 현재 FLUX 3 Image API 호출을 제공하지 않습니다.
APIMart 요금, 지원 파라미터, 제한 및 요청 예시는 연동 후 공개되는 정보를 기준으로 확인해 주세요. 현재 공개 일정은 안내하지 않습니다.
같은 카테고리의 다른 모델을 탐색하세요.

Nano banana
Nano Banana (gemini-2.5-flash-image-preview) is Google DeepMind's fast, conversational image generation and editing model. It delivers natural-language text-to-image, precise multi-turn edits, strong character consistency, and multi-image fusion at low latency.

Seedream 5.0 Pro
Seedream 5.0 Pro (seedream-5-0-pro) is quality-first text-to-image model. It produces cinematic 1K and 2K images with best-in-class text rendering, supports up to 10 reference images for style and character consistency, and unifies generation and editing in a single API call.

Midjourney
Midjourney is a leading text-to-image AI model renowned for its painterly aesthetics, strong composition, and cinematic lighting. Generate high-quality images from text prompts or reference images, with fine-grained control over style, aspect ratio, and creative parameters — accessible programmatically through APIMart's unified API, no Discord required.

Wan 2.7 Image
wan‑2‑7‑image is an advanced image generation model in Alibaba’s Wan 2.7 series designed to create high‑quality visuals from text prompts. It excels at generating detailed, realistic images with accurate object representation and rich composition. The model supports multimodal input (e.g., combining text with reference images) to influence style and structure, and it’s well‑suited for creative workflows such as marketing assets, product visuals, social media graphics, and artistic content production. With strong semantic understanding and prompt adherence, wan‑2‑7‑image delivers both fidelity and expressive visual output.