Image Generation
FLUX, Stable Diffusion, Seedream 등 최신 이미지 생성 모델로 텍스트-이미지와 이미지 편집을 지원합니다. 배경 제거·업스케일·얼굴 복원은 Image Tools 문서에서 다룹니다.
Quick Start - 30초 만에 이미지 생성
curl -X POST https://api.core.today/v1/predictions \
-H "Content-Type: application/json" \
-H "X-API-Key: $CORE_API_KEY" \
-d '{
"model": "black-forest-labs/flux-schnell",
"input": {
"prompt": "A cute robot painting on a canvas, digital art"
}
}'모델 비교
| 모델 | 크레딧 | 속도 | 특징 |
|---|---|---|---|
| flux-2-klein-4b | 2 | 빠름 | Black Forest Labs, 빠른 생성 |
| sdxl-lightning-4step | 4 | 빠름 | ByteDance, 빠른 생성 |
| pulid | 6 | 빠름 | ByteDance, 빠른 생성 |
| flux-schnell | 7 | 빠름 | Black Forest Labs, 빠른 생성 |
| sdxl | 9 | 보통 | Stability AI, 고품질 |
| photomaker | 10 | 보통 | Tencent ARC, 고품질 |
| p-image | 12 | 빠름 | Pruna AI, 빠른 생성 |
| photomaker-style | 15 | 보통 | Tencent ARC, 고품질 |
| face-to-many | 21 | 빠름 | fofr, 빠른 생성 |
| face-to-sticker | 22 | 빠름 | fofr, 빠른 생성 |
| image-01 | 23 | 빠름 | MiniMax, 빠른 생성 |
| flux-2-dev | 28 | 보통 | Black Forest Labs, 고품질 |
| proteus-v0.2 | 35 | 보통 | Proteus, 고품질 |
| z-image-turbo | 40 | 빠름 | Pruna AI, 빠른 생성 |
| imagen-4-fast | 46 | 빠름 | Google, 빠른 생성 |
| flux-pulid | 47 | 보통 | ByteDance, 고품질 |
| grok-imagine-image | 47 | 빠름 | xAI, 빠른 생성 |
| flux-dev | 58 | 빠름 | Black Forest Labs, 빠른 생성 |
| flux-krea-dev | 58 | 보통 | Krea AI, 고품질 |
| flux-2-pro | 70 | 보통 | Black Forest Labs, 최고 품질 |
| gen4-image-turbo | 70 | 빠름 | Runway, 빠른 생성 |
| ideogram-v4-turbo | 70 | 빠름 | Ideogram, 빠른 생성 |
| ideogram-v3-turbo | 70 | 빠름 | Ideogram, 빠른 생성 |
| krea-2-medium | 70 | 빠름 | Krea AI, 빠른 생성 |
| seedream-4 | 70 | 보통 | ByteDance, 고품질 |
| nano-banana-2-lite | 80 | 빠름 | Google, 빠른 생성 |
| seedream-5-lite | 82 | 빠름 | ByteDance, 빠른 생성 |
| seedream-4.5 | 90 | 보통 | ByteDance, 최고 품질 |
| nano-banana/edit | 91 | 빠름 | Google, 빠른 생성 |
| nano-banana | 91 | 빠름 | Google, 빠른 생성 |
| change-haircut | 93 | 빠름 | Black Forest Labs, 빠른 생성 |
| flux-1.1-pro | 93 | 빠름 | Black Forest Labs, 빠른 생성 |
| flux-kontext-pro | 93 | 보통 | Black Forest Labs, 최고 품질 |
| imagen-4 | 93 | 보통 | Google, 최고 품질 |
| professional-headshot | 93 | 빠름 | Black Forest Labs, 빠른 생성 |
| recraft-v4 | 93 | 보통 | Recraft, 고품질 |
| recraft-v3 | 93 | 보통 | Recraft, 고품질 |
| controlnet-scribble | 120 | 보통 | ControlNet, 고품질 |
| flux-1.1-pro-ultra | 140 | 보통 | Black Forest Labs, 최고 품질 |
| ideogram-v4-balanced | 140 | 보통 | Ideogram, 고품질 |
| krea-2-large | 140 | 보통 | Krea AI, 고품질 |
| flux-2-max | 160 | 빠름 | Black Forest Labs, 빠른 생성 |
| nano-banana-2/edit | 190 | 보통 | Google, 최고 품질 |
| flux-kontext-max | 190 | 보통 | Black Forest Labs, 최고 품질 |
| gen4-image | 190 | 보통 | Runway, 최고 품질 |
| nano-banana-2 | 190 | 빠름 | Google, 빠른 생성 |
| recraft-v4-svg | 190 | 보통 | Recraft, 고품질 |
| seedream-5-pro | 210 | 보통 | ByteDance, 최고 품질 |
| ideogram-v4-quality | 230 | 느림 | Ideogram, 최고 품질 |
| flux-2-flex | 280 | 느림 | Black Forest Labs, 최고 품질 |
| gpt-image-2 | 300 | 보통 | OpenAI, 최고 품질 |
| gpt-image-1.5 | 310 | 보통 | OpenAI, 최고 품질 |
| nano-banana-pro/edit | 350 | 보통 | Google, 최고 품질 |
| ideogram-character | 350 | 보통 | Ideogram, 최고 품질 |
| nano-banana-pro | 350 | 보통 | Google, 최고 품질 |
| mai-image-2.5-pro | 400 | 보통 | Microsoft, 최고 품질 |
모델 상세 정보
각 모델의 상세한 파라미터, 예제 코드, 활용 팁을 확인하세요.
Change Haircut
Black Forest Labs
Change anyone's hairstyle and hair color from a single photo, powered by FLUX.1 Kontext [pro]. Choose from 90+ hairstyles and 30 hair colors — or let 'Random' surprise you — while keeping the face untouched.
ControlNet Scribble
ControlNet
The classic sketch-to-image model (38M+ runs) — turn any scribble or line drawing into a detailed image guided by your prompt. Draw the composition, describe the content.
Nano Banana (Edit)
Dedicated edit endpoint for Nano Banana, Google's Gemini 2.5 Flash-based image model. Pass input image URLs to perform conversational editing with character consistency and multi-image fusion.
Nano Banana 2 (Edit)
Edit endpoint for Nano Banana 2, built on Gemini 3.1 Flash Image. Adds resolution control (1K/2K/4K), Google Search grounding, and thinking mode while preserving conversational editing and multi-image fusion.
Nano Banana Pro (Edit)
Edit endpoint for Nano Banana Pro built on Gemini 3 Pro. Professional-grade controls, legible multilingual typography, real-time grounding via Google Search, and resolution up to 2K for editing.
Face to Many
fofr
Turn a face photo into 6 fun styles (15M+ runs) — 3D, Emoji, Video game, Pixels, Clay, or Toy. The viral avatar generator behind countless profile-picture apps.
Face to Sticker
fofr
Turn any face photo into a fun die-cut sticker (1.6M+ runs). InstantID keeps the likeness while IP-Adapter controls take the artwork from faithful caricature to loose cartoon — with an optional 2x upscale for print quality.
FLUX.2 Klein 4B
Black Forest Labs
Very fast image generation and editing model. 4-step distilled, sub-second inference for production and near real-time applications.
FLUX 2 Flex
Black Forest Labs
Maximum-quality FLUX model supporting up to 10 reference images and advanced typography. The most capable model for complex, multi-reference creative projects.
FLUX 2 Max
Black Forest Labs
The highest fidelity image model from Black Forest Labs. Best-in-class prompt following and the most consistent editing in the FLUX.2 lineup — preserves colors, lighting, faces, text, and objects across edits with up to 8 reference images.
FLUX 2 Pro
Black Forest Labs
Professional-grade FLUX 2 with high-quality editing and up to 8 reference image support. Excellent balance of quality, speed, and creative control.
FLUX.2 Dev
Black Forest Labs
Development version of FLUX.2 with image editing capabilities and reference image support. Ideal for iterative design workflows and experimentation.
FLUX 1.1 Pro
Black Forest Labs
Fast high-quality image generation, an upgrade to FLUX.1 Pro with faster speed and improved quality. Perfect for production workloads requiring both speed and fidelity.
FLUX 1.1 Pro Ultra
Black Forest Labs
FLUX1.1 [pro] in ultra and raw modes. Images are up to 4 megapixels — the highest-resolution tier of the FLUX 1.1 Pro family. Use raw mode for realism.
FLUX Dev
Black Forest Labs
A 12 billion parameter rectified flow transformer capable of generating images from text descriptions, tuned for open, high-quality experimentation.
FLUX Kontext Max
Black Forest Labs
A premium text-based image editing model that delivers maximum performance and improved typography generation for transforming images through natural language prompts.
FLUX Kontext Pro
Black Forest Labs
State-of-the-art text-based image editing model that transforms images through natural language. Excellent for style transfer, object modification, text replacement, background changes, and character consistency.
FLUX.1 Krea [dev]
Krea AI
Photorealistic image generation that specifically avoids the 'AI look', producing natural-looking images indistinguishable from real photographs.
FLUX PuLID
ByteDance
PuLID identity customization on FLUX-dev: generate photorealistic portraits of a specific person from one face photo, with markedly higher fidelity than SDXL-based variants. Tune id_weight and start_step to balance likeness against prompt editability.
FLUX.1 Schnell
Black Forest Labs
Ultra-fast image generation model optimized for speed. Generates high-quality images in just 1-2 seconds, perfect for real-time applications and rapid prototyping.
Runway Gen-4 Image
Runway
Runway's Gen-4 Image model with references: combine up to 3 reference images with @tag mentions in your prompt to keep characters, objects, and locations consistent across every angle and scene, at 720p or 1080p.
Runway Gen-4 Image Turbo
Runway
Gen-4 Image Turbo is 2.5x faster and cheaper than Gen-4 Image, with the same reference-driven API: use 1 to 3 reference images with @tag mentions for consistent characters and objects, at a flat price regardless of resolution.
GPT Image 2
OpenAI
OpenAI's state-of-the-art image generation and editing model with strong instruction following, sharp text rendering, and detailed editing. Quality-based pricing lets you trade off cost vs. fidelity.
GPT Image 1.5
OpenAI
OpenAI's latest image generation model with better instruction following and adherence to prompts, including sharp text rendering and detailed editing.
Grok Imagine Image
xAI
Generate images using xAI's Grok Imagine model. Sibling to Grok Imagine Video, sharing the same underlying Grok Imagine architecture for fast text-to-image generation.
Ideogram V4 Balanced
Ideogram
A middle-ground tier in Ideogram's v4 family, balancing generation speed and output quality. Delivers strong typography and photorealism at a lower cost than the Quality tier.
Ideogram V4 Quality
Ideogram
The highest-fidelity tier of Ideogram's v4 model family, tuned for maximum detail, realism, and typography accuracy. Best suited for final production assets where quality matters more than speed.
Ideogram V4 Turbo
Ideogram
The fastest and cheapest model in Ideogram's v4 family, built for rapid iteration while retaining Ideogram's signature text rendering and style consistency.
Ideogram V3 Turbo
Ideogram
The fastest and cheapest Ideogram v3 tier. V3 creates images with stunning realism, creative designs, and consistent styles.
Ideogram Character
Ideogram
Generate consistent characters from a single reference image. Render the same character in many styles — realistic or fiction — insert them into existing photos with mask inpainting, and rely on Ideogram's signature text rendering for legible signs and typography.
MiniMax Image-01
MiniMax
MiniMax's first image generation model with character reference support: provide a single face photo via subject_reference and generate consistent images of that person across prompts, styles, and aspect ratios — up to 9 images per request.
Imagen 4
Google's Imagen 4 flagship text-to-image model.
Imagen 4 Fast
A fast version of Imagen 4 for when speed and cost are more important than maximum quality.
Krea 2 Large
Krea AI
Krea AI's flagship text-to-image model, focused on photorealistic output with strong prompt adherence. Supports style-reference and moodboard-guided generation for consistent visual direction.
Krea 2 Medium
Krea AI
A lower-cost variant of Krea 2 Large, trading some fidelity for faster and cheaper generation while keeping the same photorealistic style focus.
Microsoft MAI-Image 2.5 Pro
Microsoft
Microsoft's highest-fidelity image model for production-grade text-to-image generation via Fal.AI. Built for hero imagery, detailed compositions, precise text rendering, photorealism, stylized illustration, commercial design, and visually rich concept work.
Nano Banana 2 (Gemini 3.1 Flash Image)
Google's fast image generation model built on Gemini 3.1 Flash Image. The high-efficiency counterpart to Nano Banana Pro — combining Pro-level visual quality with Flash-level speed and pricing. Features conversational editing, multi-image fusion, character consistency, accurate text rendering, and Google Search grounding. Supports up to 14 reference images and resolutions up to 4K.
Nano Banana 2 Lite
Google's lightweight Nano Banana 2 variant built on Gemini 3.1 Flash Image, tuned for faster and cheaper generation. Retains conversational editing, multi-image fusion, and character consistency from the full Nano Banana 2.
Nano Banana
Google Gemini 2.5 Flash-based image generation with multimodal editing capabilities. Fast and versatile for both creation and editing tasks.
Nano Banana Pro (Gemini 3 Pro Image)
Google's state-of-the-art image generation and editing model built on Gemini 3 Pro. Creates detailed visuals with legible text in multiple languages, connects to real-time information from Google Search, and provides professional-grade creative controls. Supports up to 14 reference images and resolutions up to 4K.
Pruna P-Image
Pruna AI
Pruna AI's distilled text-to-image model optimized for extremely low-cost, high-throughput generation. One of the most-run community models on Replicate (15.6M+ runs) thanks to its speed and price.
PhotoMaker
Tencent ARC
Generate stylized photos of a person from 1-4 reference photos (9M+ runs). Ten styles including Cinematic, Disney Character, and Digital Art — keep the identity, change everything else.
PhotoMaker Style
Tencent ARC
The stylization-focused variant of PhotoMaker: turn 1-4 photos of a person into paintings, comics, 3D art, and more with stronger style transfer. Pairs with the base PhotoMaker — use this one when style matters more than photorealism.
Professional Headshot
Black Forest Labs
Turn any single photo into a polished professional business headshot, powered by FLUX.1 Kontext [pro]. Pick a background — white, black, gray, neutral, or office — and get a LinkedIn-ready portrait in one step.
Proteus v0.2
Proteus
Popular anime-focused image model (12M+ runs) — high-quality anime and illustration styles with img2img and inpainting support. The go-to for anime avatars and webtoon-style art.
PuLID
ByteDance
ByteDance PuLID: tuning-free identity customization on SDXL. Give it one face photo and a prompt to generate portraits in any scene or style — no training, 4-step fast sampling, and even two-identity blending. Extremely cost-effective at 2 credits per image.
Recraft V4
Recraft
Recraft's next-generation text-to-image model, improving on V3's typography, style range, and prompt adherence for production-grade brand and design assets.
Recraft V4 SVG
Recraft
Generates vector graphics directly in SVG format using Recraft V4, ideal for logos, icons, and scalable illustrations that need to stay crisp at any size.
Recraft V3
Recraft
Recraft V3 (code-named red_panda) is a text-to-image model with the ability to generate long texts, and images in a wide list of styles. SOTA in image generation per the Artificial Analysis Text-to-Image Benchmark.
Stable Diffusion XL
Stability AI
Stability AI's classic SDXL (85M+ runs) — the battle-tested text-to-image model with img2img, inpainting, refiner, and LoRA support. A dependable workhorse with a huge ecosystem.
SDXL Lightning 4-step
ByteDance
ByteDance's 4-step SDXL Lightning — the most-run model on Replicate (1B+ runs). Near-instant 1024px image generation at one of the lowest prices in the catalog.
Seedream 5 Lite
ByteDance
Seedream 5.0 lite: image generation with built-in reasoning, example-based editing, and deep domain knowledge. Supports multi-reference generation with up to 14 images and sequential batch generation.
Seedream 5 Pro
ByteDance
ByteDance's flagship Seedream 5.0 Pro image generation model, with built-in reasoning, precise instruction following, and multi-reference support for up to 10 images. Offers 1K and 2K resolution output with higher fidelity than Seedream 5 Lite.
Seedream 4.5
ByteDance
Upgraded ByteDance image model with stronger spatial understanding and world knowledge. Supports single/multi-reference image-to-image editing and sequential (multi-image) generation.
Seedream 4.0
ByteDance
ByteDance's latest image generation model with exceptional prompt understanding and creative capabilities.
Z-Image Turbo
Pruna AI
Pruna's ultra-fast Z-Image Turbo (48M+ runs). Megapixel-priced image generation up to 2048x2048 — pay exactly for the resolution you generate, with excellent price/performance for high-volume use.
FLUX Schnell
가장 빠른 이미지 생성 모델. 실시간 애플리케이션에 최적.
파라미터
| Parameter | Type | Required | Description |
|---|---|---|---|
| prompt | string | Yes | 이미지 설명 (영어 권장) |
| aspect_ratio | string | No | 1:1, 16:9, 9:16, 4:3, 3:4 등 |
| num_outputs | integer | No | 생성할 이미지 수 (1-4) |
| seed | integer | No | 결과 재현용 시드값 |
예제 코드
curl -X POST https://api.core.today/v1/predictions \
-H "Content-Type: application/json" \
-H "X-API-Key: $CORE_API_KEY" \
-d '{
"model": "black-forest-labs/flux-schnell",
"input": {
"prompt": "A beautiful sunset over mountains, cinematic lighting, 8K resolution",
"aspect_ratio": "16:9",
"num_outputs": 2
}
}'프롬프트 작성 팁
- 영어 프롬프트가 가장 좋은 결과를 냅니다
- 스타일 키워드 추가: "digital art", "photography", "oil painting"
- 품질 키워드 추가: "8K", "detailed", "professional"
- 조명 설명: "cinematic lighting", "golden hour", "studio lighting"
Seedream 4
ByteDance의 최신 모델. 최고 품질과 텍스트 렌더링 특화.
특징
텍스트 렌더링
간판, 로고, 포스터 등 텍스트가 포함된 이미지 생성에 탁월
다국어 지원
한국어, 중국어, 일본어 프롬프트도 잘 이해
텍스트 포함 이미지 예제
{
"model": "bytedance/seedream-4",
"input": {
"prompt": "A coffee shop sign that says \"CAFE MOCHA\" in elegant gold typography, warm lighting, cozy atmosphere",
"aspect_ratio": "1:1",
"guidance_scale": 7.5
}
}텍스트 렌더링 팁
"HELLO WORLD"Use Cases
E-commerce 상품 이미지
제품 컨셉 이미지, 배경 생성, 상품 목업 자동 생성
마케팅 배너 & 광고
SNS 배너, 광고 이미지, 프로모션 그래픽 제작
게임 & 엔터테인먼트
게임 아트, 캐릭터 컨셉, 배경 일러스트
실시간 생성 앱
사용자 입력에 즉시 반응하는 이미지 생성 서비스
모델 선택 가이드
빠른 생성
flux-schnell
7 credits, 1-2초
최고 품질
seedream-4
70 credits, 10-15초
텍스트 포함
seedream-4
로고, 간판, 포스터
대량 생성
nano-banana
91 credits
응답 형식
{
"job_id": "pred_abc123",
"status": "completed",
"model": "black-forest-labs/flux-schnell",
"result": [
"https://files.core.today/aiapi/9f3a/pred_abc123/o/0.png?Expires=...&Signature=...",
"https://files.core.today/aiapi/9f3a/pred_abc123/o/1.png?Expires=...&Signature=..."
],
"public_url": null,
"output_files": [
{
"object_key": "aiapi/9f3a/pred_abc123/o/0.png",
"url": "https://files.core.today/aiapi/9f3a/pred_abc123/o/0.png?Expires=...&Signature=..."
},
{
"object_key": "aiapi/9f3a/pred_abc123/o/1.png",
"url": "https://files.core.today/aiapi/9f3a/pred_abc123/o/1.png?Expires=...&Signature=..."
}
]
}이미지 URL
is_public: true 옵션을 사용하세요.