Video Generation
텍스트나 이미지로 고품질 비디오를 생성하는 최신 AI 모델들입니다. Kling, MiniMax Hailuo, Veo, Sora 등 다양한 모델을 지원합니다.
모델 비교표
| 모델 | 크레딧 | 유형 | 특징 |
|---|---|---|---|
| wan-2.2-i2v-fast | 120 | I2V | Alibaba, 빠른 생성 |
| wan-2.2-t2v-fast | 120 | T2V | Alibaba, 빠른 생성 |
| p-video | 240 | T2V/I2V | Pruna AI, 빠른 생성 |
| seedance-1-lite | 420 | T2V/I2V | ByteDance, 빠른 생성 |
| hailuo-2.3-fast | 450 | T2V/I2V | MiniMax, 고품질 |
| kling-v2.1 | 580 | T2V/I2V | Kuaishou, 최고 품질 |
| hailuo-2.3 | 650 | T2V/I2V | MiniMax, 최고 품질 |
| seedance-1-pro-fast | 690 | T2V/I2V | ByteDance, 빠른 생성 |
| pixverse-v5 | 700 | T2V/I2V | PixVerse, 최고 품질 |
| ray-flash-2-720p | 700 | T2V/I2V | Luma, 빠른 생성 |
| wan-2.5-i2v-fast | 790 | I2V | Alibaba, 고품질 |
| wan-2.5-t2v-fast | 790 | T2V | Alibaba, 고품질 |
| kling-v2.5-turbo-pro | 810 | T2V/I2V | Kuaishou, 최고 품질 |
| pixverse-v6 | 810 | T2V/I2V | PixVerse, 고품질 |
| p-video-avatar | 870 | T2V/I2V | Pruna AI, 고품질 |
| seedance-2.0-mini | 890 | T2V/I2V | ByteDance, 빠른 생성 |
| sora-2 | 930 | T2V | OpenAI, 최고 품질 |
| p-video-animate | 1,050 | T2V/I2V | Pruna AI, 빠른 생성 |
| wan-2.5-i2v | 1,160 | I2V | Alibaba, 최고 품질 |
| wan-2.5-t2v | 1,160 | T2V | Alibaba, 최고 품질 |
| grok-imagine-video | 1,300 | T2V/I2V | xAI, 빠른 생성 |
| gen-4.5 | 1,400 | T2V/I2V | Runway, 최고 품질 |
| seedance-2.0-fast | 1,430 | T2V/I2V | ByteDance, 빠른 생성 |
| h3 | 1,510 | T2V/I2V | MiniMax, 최고 품질 |
| image-to-video | 1,510 | T2V/I2V | MiniMax, 최고 품질 |
| reference-to-video | 1,510 | T2V | MiniMax, 최고 품질 |
| text-to-video | 1,510 | T2V | MiniMax, 최고 품질 |
| kling-v2.6 | 1,630 | T2V/I2V | Kuaishou, 최고 품질 |
| green-screen-despill | 1,740 | T2V | Bria, 빠른 생성 |
| ray-3.2 | 1,740 | T2V/I2V | Luma, 최고 품질 |
| seedance-1-pro | 1,740 | T2V/I2V | ByteDance, 최고 품질 |
| wan-2.7-i2v | 1,740 | I2V | Alibaba, 고품질 |
| wan-2.7-t2v | 1,740 | T2V | Alibaba, 고품질 |
| seedance-2.0 | 1,780 | T2V/I2V | ByteDance, 최고 품질 |
| image-to-video | 1,980 | T2V/I2V | Black Forest Labs, 최고 품질 |
| happyhorse-1.1 | 2,090 | T2V | Alibaba, 고품질 |
| gemini-omni-flash | 2,320 | T2V | Google, 빠른 생성 |
| wan-3 | 2,330 | T2V/I2V | Alibaba, 고품질 |
| kling-v3-omni-video | 2,610 | T2V/I2V | Kuaishou, 최고 품질 |
| kling-v3-video | 2,610 | T2V/I2V | Kuaishou, 최고 품질 |
| edit | 2,790 | T2V | xAI, 고품질 |
| sora-2-pro | 2,790 | T2V | OpenAI, 최고 품질 |
| veo-3.1-fast | 2,790 | T2V/I2V | Google, 고품질 |
| gemini-omni-1.1 | 3,490 | T2V/I2V | Google, 빠른 생성 |
| grok-imagine-video-1.5 | 4,650 | T2V/I2V | xAI, 최고 품질 |
| image-to-video | 5,500 | T2V/I2V | ByteDance, 최고 품질 |
| reference-to-video | 5,500 | T2V/I2V | ByteDance, 최고 품질 |
| seedance-2.5 | 5,500 | T2V/I2V | ByteDance, 최고 품질 |
| text-to-video | 5,500 | T2V | ByteDance, 최고 품질 |
| veo-3.1 | 7,440 | T2V/I2V | Google, 최고 품질 |
모델 상세 정보
각 모델의 상세한 파라미터, 예제 코드, 활용 팁을 확인하세요.
Grok Imagine Video Edit
xAI
Edit an existing video with xAI's Grok Imagine Video model. The output inherits the source video's duration, resolution (720p cap), and aspect ratio.
Gemini Omni 1.1 Flash
Google's fast multimodal video generation and editing model with native audio, using the Interactions API. Text-to-video, image-to-video, keyframe interpolation, reference-to-video, and video editing at 360p-4K.
Google Gemini Omni Flash
Google Gemini Omni Flash text-to-video via Fal.AI. Generates a video directly from a descriptive text prompt — pacing and audio (dialogue, background music) are controlled in the prompt itself, in 16:9 or 9:16 at 3-10 second durations.
Gen-4.5
Runway
Runway's Gen-4.5 model, offering state-of-the-art video motion quality, prompt adherence, and visual fidelity for text-to-video and image-to-video generation.
Bria Green Screen Background Removal
Bria
Bria's chroma-key background removal via Fal.AI — removes the green-screen background from a video with automatic green-spill suppression, outputting on a transparent background (alpha-capable codecs) with clean, professional edges.
Grok Imagine Video 1.5
xAI
xAI's Grok Imagine Video 1.5 over the direct xAI API: multi-mode video creation - text-to-video and image-to-video with synchronized audio, up to native 1080p.
Grok Imagine Video
xAI
Generate videos with xAI's Grok Imagine Video model over the direct xAI API. Text-to-video and image-to-video with synchronized audio. For editing an existing clip, use xai/grok-imagine-video/edit.
MiniMax H3
MiniMax
MiniMax H3 multimodal video generation via Replicate — one endpoint for text-to-video, first/last-frame image-to-video, and reference-based generation with images, videos, and audio cited in the prompt. 4-15 second duration at 768P or 2K.
MiniMax Hailuo 2.3
MiniMax
Realistic human motion video generation with advanced character consistency and natural movement.
MiniMax Hailuo 2.3 Fast
MiniMax
Lower-latency version of Hailuo 2.3 optimized for faster generation while maintaining good quality for human motion videos.
Happy Horse 1.1
Alibaba
Alibaba's Happy Horse 1.1 generates videos from text, animates a single image, or builds a video from multiple reference images. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.
Flux 3 Image to Video
Black Forest Labs
FLUX.3 is Black Forest Labs' frontier video model. This endpoint animates a single still image into coherent, natural motion at 720p or 1080p for 5-20 seconds, with synchronized audio generated alongside the video and a tunable safety tolerance.
MiniMax Hailuo-03 (H3) Image to Video
MiniMax
MiniMax Hailuo-03 (H3) image-to-video via Fal.AI. 2K video generation from a first-frame image, 5-15 second duration, native audio, with optional first-to-last keyframe control via end_image_url.
Seedance 2.5 Image to Video
ByteDance
ByteDance Seedance 2.5 image-to-video via Fal.AI. Animates a single still into a native clip of up to 30 seconds at 480p/720p without the drift or stitching of multi-clip workflows, with optional end-frame control and synchronized audio generated in the same pass.
Kling v3 Omni Video
Kuaishou
Kling Video 3.0 Omni: a unified multimodal video model that generates and edits video from text, images, reference images, and existing video. Combines text-to-video, image-to-video, reference-based generation, and video editing with native audio and multi-shot control.
Kling v3 Video
Kuaishou
Kling Video 3.0: Kuaishou's flagship text/image-to-video model generating cinematic videos up to 15 seconds with multi-shot control, native audio, start/end frame images, and a dedicated 4K mode.
Kling v2.6
Kuaishou
Kling 2.6 Pro: top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation. Audio generation is enabled by default.
Kling 2.5 Turbo Pro
Kuaishou
Cinematic-grade video generation with enhanced motion and scene coherence. Top-tier Kling model for professional output.
Kling v2.1
Kuaishou
Kling v2.1 with 720p/1080p support and frame transition capabilities for smooth, high-quality video generation.
Pruna P-Video
Pruna AI
PrunaAI's fast video generator with a built-in draft mode for rapid creative iteration. Text-to-video, image-to-video, and audio-conditioned generation in a single endpoint, with clips up to 20 seconds — one of the longest durations on the platform.
Pruna P-Video Animate
Pruna AI
Pruna AI's motion-transfer video model — animates a reference image with the motion and audio of a source video. Optimized for speed and cost (about 5.24s of generation per 1s of output video).
Pruna P-Video Avatar
Pruna AI
Pruna AI's talking-avatar video model — animates a single input image to speak provided text (voice_script) or an uploaded audio track, with selectable Gemini-family preset voices, language, and visual delivery prompt.
PixVerse V6
PixVerse
PixVerse's flagship video generation model. Generates cinematic videos with synchronized audio, multi-shot sequences with scene transitions, astonishing physics, and precise camera control at up to 1080p.
PixVerse V5
PixVerse
Advanced video generation with special effects capabilities and anime-optimized output, supporting multiple visual styles.
Luma Ray 3.2
Luma
Luma's flagship Ray video model. Text-to-video and keyframe (start/end image) generation with optional HDR-encoded output and professional EXR export.
Luma Ray Flash 2 720p
Luma
Luma's Ray Flash 2 generates 5 or 9 second 720p videos faster and cheaper than Ray 2. Supports keyframe control via start and end images, seamless loops, and a rich library of camera motion concepts — the first Luma model on the platform.
MiniMax Hailuo-03 (H3) Reference to Video
MiniMax
MiniMax Hailuo-03 (H3) reference-to-video via Fal.AI. Generates video from multimodal references — subject/style images, motion video clips, and audio clips, each cited in the prompt by order — keeping subjects consistent while following the referenced motion and audio.
Seedance 2.5 Reference to Video
ByteDance
ByteDance Seedance 2.5 reference-to-video via Fal.AI. Generates video from up to 50 multimodal references — up to 30 images, 10 videos, and 10 audio files — locking a character, set, and palette across a full 30-second take. Reference inputs are addressed from the prompt as @Image1, @Video1, @Audio1, and drive motion transfer, editing, extension, and lip-sync.
Seedance 2.5
ByteDance
ByteDance's flagship multimodal video model with native audio, native 30-second single-pass generation, and large multimodal reference sets (up to 30 images, 10 videos, 10 audios). Supports text-to-video, image-to-video, first/last-frame control, video editing, extension, and lip-sync.
Seedance 2.0
ByteDance
ByteDance's next-generation multimodal video model with native synchronized audio. Combines up to 9 reference images, 3 videos, and 3 audio files in a single generation for character-consistent, lip-synced video creation, editing, and extension.
Seedance 2.0 Fast
ByteDance
A faster, cheaper variant of Seedance 2.0 for quicker video generation with multimodal reference inputs (up to 9 images, 3 videos, 3 audios) and native audio, at 480p or 720p.
Seedance 2.0 Mini
ByteDance
Lighter, cheaper variant of ByteDance's Seedance 2.0. Native audio, multimodal reference inputs (images/videos/audio), text-to-video and image-to-video, capped at 720p (no 1080p/4K tier).
Seedance 1 Lite
ByteDance
A lightweight ByteDance video generation model offering text-to-video and image-to-video support for 4-12 second videos at 480p, 720p, or 1080p resolution.
Seedance 1 Pro
ByteDance
A pro version of Seedance that offers text-to-video and image-to-video support for 2-12 second videos, at 480p, 720p, and 1080p resolution.
Seedance 1 Pro Fast
ByteDance
ByteDance's cinematic video generation model with fast generation speed and professional output quality.
OpenAI Sora 2
OpenAI
OpenAI's video generation model with realistic physics simulation and audio generation capabilities, producing highly coherent videos.
OpenAI Sora 2 Pro
OpenAI
OpenAI's most advanced synced-audio video generation model. The premium tier of Sora 2 with higher fidelity, up to 1024p resolution, and image-to-video via an input reference frame.
MiniMax Hailuo-03 (H3) Text to Video
MiniMax
MiniMax Hailuo-03 (H3) text-to-video via Fal.AI. State-of-the-art 2K video generation from a text prompt, 5-15 second duration, native audio, and wide aspect-ratio support (21:9 through 9:16).
Seedance 2.5 Text to Video
ByteDance
ByteDance Seedance 2.5 text-to-video via Fal.AI. Generates a native single-shot clip of up to 30 seconds at 480p/720p from a text prompt alone, reasoning about the whole shot at once so motion, lighting, and subject identity stay coherent from first frame to last. Synchronized audio (dialogue, sound effects, music) is generated in the same pass.
Google Veo 3.1
Google's state-of-the-art video generation model with built-in audio generation, producing cinematic-quality videos with synchronized sound.
Google Veo 3.1 Fast
Fast version of Veo 3.1 with audio generation, optimized for speed while maintaining high quality output.
Wan 3.0
Alibaba
Alibaba's Wan 3.0 generates video from a text prompt or a starting image, with cinematic motion and support for 480p, 720p, and 1080p output up to 30 seconds.
Wan 2.7 I2V
Alibaba
Alibaba Wan 2.7 image-to-video model. Animates a first frame (and optional last frame or continuation clip) with audio synchronization, up to 15 seconds, at 720p or 1080p.
Wan 2.7 T2V
Alibaba
Alibaba Wan 2.7 text-to-video model. Supports up to 15 seconds, audio synchronization for voice/music, multilingual prompts, and prompt expansion, at 720p or 1080p.
Wan 2.5 I2V
Alibaba
Image-to-video model with lip sync support, animating still images into realistic videos with natural motion.
Wan 2.5 I2V Fast
Alibaba
Fast image-to-video variant of Wan 2.5, optimized for rapid generation of animated videos from still images.
Wan 2.5 T2V
Alibaba
Text-to-video model with audio synchronization support, producing high-quality videos from text prompts with natural motion.
Wan 2.5 T2V Fast
Alibaba
Fast text-to-video generation variant of Wan 2.5, optimized for speed with good quality output.
Wan 2.2 I2V Fast
Alibaba
A very fast and cheap PrunaAI-optimized version of Alibaba's Wan 2.2 A14B image-to-video model. Turns a single still image plus a prompt into a short animated clip at 480p or 720p, with an optional frame-interpolation pass for smoother motion.
Wan 2.2 T2V Fast
Alibaba
A very fast and cheap PrunaAI-optimized version of Alibaba's Wan 2.2 A14B text-to-video model. Generates short clips at 480p or 720p directly from a text prompt, with 30 FPS frame interpolation enabled by default — the text-to-video sibling of Wan 2.2 I2V Fast.
Hailuo 2.3 (MiniMax)
MiniMax의 고품질 비디오 생성 모델.
curl -X POST https://api.core.today/v1/predictions \
-H "Content-Type: application/json" \
-H "X-API-Key: cdt_your_api_key" \
-d '{
"model": "minimax/hailuo-2.3",
"input": {
"prompt": "A golden retriever running on the beach at sunset, slow motion, cinematic",
"duration": 6,
"aspect_ratio": "16:9"
}
}'이미지 → 비디오 (Image-to-Video)
첫 프레임 이미지를 제공하여 비디오를 생성합니다.
1. 이미지 업로드
curl -X POST https://api.core.today/v1/files/upload-url \
-H "X-API-Key: cdt_your_api_key" \
-H "Content-Type: application/json" \
-d '{"filename": "first_frame.jpg", "content_type": "image/jpeg"}'2. 비디오 생성
curl -X POST https://api.core.today/v1/predictions \
-H "Content-Type: application/json" \
-H "X-API-Key: cdt_your_api_key" \
-d '{
"model": "kwaivgi/kling-v2.5-turbo-pro",
"input": {
"prompt": "Camera slowly zooms in, gentle wind blowing",
"image": "https://files.core.today/.../first_frame.jpg",
"duration": 5
}
}'카메라 워크 프롬프트
더 나은 비디오를 위해 카메라 움직임을 프롬프트에 포함하세요:
기본 움직임
zoom in- 줌 인zoom out- 줌 아웃pan left/right- 패닝tilt up/down- 틸트
고급 움직임
tracking shot- 트래킹drone shot- 드론 샷slow motion- 슬로우 모션time lapse- 타임랩스
모델 선택 가이드
저비용 빠른 생성
seedance-1-pro-fast (190크레딧)
고품질 T2V
hailuo-2.3 (420크레딧)
I2V (이미지→비디오)
kling-v2.5-turbo-pro (525크레딧)
최고 품질
veo-3.1 (4,800크레딧)