Skip to main content
Core.Today
Model APIs

Video Generation

텍스트나 이미지로 고품질 비디오를 생성하는 최신 AI 모델들입니다. Kling, MiniMax Hailuo, Veo, Sora 등 다양한 모델을 지원합니다.

Give this page to your AI — all model specs as LLM-friendly text
llms.txt ↗

모델 비교표

모델크레딧유형특징
wan-2.2-i2v-fast120I2VAlibaba, 빠른 생성
wan-2.2-t2v-fast120T2VAlibaba, 빠른 생성
p-video240T2V/I2VPruna AI, 빠른 생성
seedance-1-lite420T2V/I2VByteDance, 빠른 생성
hailuo-2.3-fast450T2V/I2VMiniMax, 고품질
kling-v2.1580T2V/I2VKuaishou, 최고 품질
hailuo-2.3650T2V/I2VMiniMax, 최고 품질
seedance-1-pro-fast690T2V/I2VByteDance, 빠른 생성
pixverse-v5700T2V/I2VPixVerse, 최고 품질
ray-flash-2-720p700T2V/I2VLuma, 빠른 생성
wan-2.5-i2v-fast790I2VAlibaba, 고품질
wan-2.5-t2v-fast790T2VAlibaba, 고품질
kling-v2.5-turbo-pro810T2V/I2VKuaishou, 최고 품질
pixverse-v6810T2V/I2VPixVerse, 고품질
p-video-avatar870T2V/I2VPruna AI, 고품질
seedance-2.0-mini890T2V/I2VByteDance, 빠른 생성
sora-2930T2VOpenAI, 최고 품질
p-video-animate1,050T2V/I2VPruna AI, 빠른 생성
wan-2.5-i2v1,160I2VAlibaba, 최고 품질
wan-2.5-t2v1,160T2VAlibaba, 최고 품질
grok-imagine-video1,300T2V/I2VxAI, 빠른 생성
gen-4.51,400T2V/I2VRunway, 최고 품질
seedance-2.0-fast1,430T2V/I2VByteDance, 빠른 생성
h31,510T2V/I2VMiniMax, 최고 품질
image-to-video1,510T2V/I2VMiniMax, 최고 품질
reference-to-video1,510T2VMiniMax, 최고 품질
text-to-video1,510T2VMiniMax, 최고 품질
kling-v2.61,630T2V/I2VKuaishou, 최고 품질
green-screen-despill1,740T2VBria, 빠른 생성
ray-3.21,740T2V/I2VLuma, 최고 품질
seedance-1-pro1,740T2V/I2VByteDance, 최고 품질
wan-2.7-i2v1,740I2VAlibaba, 고품질
wan-2.7-t2v1,740T2VAlibaba, 고품질
seedance-2.01,780T2V/I2VByteDance, 최고 품질
image-to-video1,980T2V/I2VBlack Forest Labs, 최고 품질
happyhorse-1.12,090T2VAlibaba, 고품질
gemini-omni-flash2,320T2VGoogle, 빠른 생성
wan-32,330T2V/I2VAlibaba, 고품질
kling-v3-omni-video2,610T2V/I2VKuaishou, 최고 품질
kling-v3-video2,610T2V/I2VKuaishou, 최고 품질
edit2,790T2VxAI, 고품질
sora-2-pro2,790T2VOpenAI, 최고 품질
veo-3.1-fast2,790T2V/I2VGoogle, 고품질
gemini-omni-1.13,490T2V/I2VGoogle, 빠른 생성
grok-imagine-video-1.54,650T2V/I2VxAI, 최고 품질
image-to-video5,500T2V/I2VByteDance, 최고 품질
reference-to-video5,500T2V/I2VByteDance, 최고 품질
seedance-2.55,500T2V/I2VByteDance, 최고 품질
text-to-video5,500T2VByteDance, 최고 품질
veo-3.17,440T2V/I2VGoogle, 최고 품질

모델 상세 정보

각 모델의 상세한 파라미터, 예제 코드, 활용 팁을 확인하세요.

50 models

Grok Imagine Video Edit

xAI

2,790 credits

Edit an existing video with xAI's Grok Imagine Video model. The output inherits the source video's duration, resolution (720p cap), and aspect ratio.

MediumHigh
View details

Gemini Omni 1.1 Flash

Google

3,490 credits

Google's fast multimodal video generation and editing model with native audio, using the Interactions API. Text-to-video, image-to-video, keyframe interpolation, reference-to-video, and video editing at 360p-4K.

FastHigh
View details

Google Gemini Omni Flash

Google

2,320 credits

Google Gemini Omni Flash text-to-video via Fal.AI. Generates a video directly from a descriptive text prompt — pacing and audio (dialogue, background music) are controlled in the prompt itself, in 16:9 or 9:16 at 3-10 second durations.

FastHigh
View details

Gen-4.5

Runway

1,400 credits

Runway's Gen-4.5 model, offering state-of-the-art video motion quality, prompt adherence, and visual fidelity for text-to-video and image-to-video generation.

MediumUltra
View details

Bria Green Screen Background Removal

Bria

1,740 credits

Bria's chroma-key background removal via Fal.AI — removes the green-screen background from a video with automatic green-spill suppression, outputting on a transparent background (alpha-capable codecs) with clean, professional edges.

FastHigh
View details

Grok Imagine Video 1.5

xAI

4,650 credits

xAI's Grok Imagine Video 1.5 over the direct xAI API: multi-mode video creation - text-to-video and image-to-video with synchronized audio, up to native 1080p.

MediumUltra
View details

Grok Imagine Video

xAI

1,300 credits

Generate videos with xAI's Grok Imagine Video model over the direct xAI API. Text-to-video and image-to-video with synchronized audio. For editing an existing clip, use xai/grok-imagine-video/edit.

FastHigh
View details

MiniMax H3

MiniMax

1,510 credits

MiniMax H3 multimodal video generation via Replicate — one endpoint for text-to-video, first/last-frame image-to-video, and reference-based generation with images, videos, and audio cited in the prompt. 4-15 second duration at 768P or 2K.

MediumUltra
View details

MiniMax Hailuo 2.3

MiniMax

650 credits

Realistic human motion video generation with advanced character consistency and natural movement.

SlowUltra
View details

MiniMax Hailuo 2.3 Fast

MiniMax

450 credits

Lower-latency version of Hailuo 2.3 optimized for faster generation while maintaining good quality for human motion videos.

MediumHigh
View details

Happy Horse 1.1

Alibaba

2,090 credits

Alibaba's Happy Horse 1.1 generates videos from text, animates a single image, or builds a video from multiple reference images. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.

MediumHigh
View details

Flux 3 Image to Video

Black Forest Labs

1,980 credits

FLUX.3 is Black Forest Labs' frontier video model. This endpoint animates a single still image into coherent, natural motion at 720p or 1080p for 5-20 seconds, with synchronized audio generated alongside the video and a tunable safety tolerance.

MediumUltra
View details

MiniMax Hailuo-03 (H3) Image to Video

MiniMax

1,510 credits

MiniMax Hailuo-03 (H3) image-to-video via Fal.AI. 2K video generation from a first-frame image, 5-15 second duration, native audio, with optional first-to-last keyframe control via end_image_url.

MediumUltra
View details

Seedance 2.5 Image to Video

ByteDance

5,500 credits

ByteDance Seedance 2.5 image-to-video via Fal.AI. Animates a single still into a native clip of up to 30 seconds at 480p/720p without the drift or stitching of multi-clip workflows, with optional end-frame control and synchronized audio generated in the same pass.

MediumUltra
View details

Kling v3 Omni Video

Kuaishou

2,610 credits

Kling Video 3.0 Omni: a unified multimodal video model that generates and edits video from text, images, reference images, and existing video. Combines text-to-video, image-to-video, reference-based generation, and video editing with native audio and multi-shot control.

SlowUltra
View details

Kling v3 Video

Kuaishou

2,610 credits

Kling Video 3.0: Kuaishou's flagship text/image-to-video model generating cinematic videos up to 15 seconds with multi-shot control, native audio, start/end frame images, and a dedicated 4K mode.

SlowUltra
View details

Kling v2.6

Kuaishou

1,630 credits

Kling 2.6 Pro: top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation. Audio generation is enabled by default.

MediumUltra
View details

Kling 2.5 Turbo Pro

Kuaishou

810 credits

Cinematic-grade video generation with enhanced motion and scene coherence. Top-tier Kling model for professional output.

MediumUltra
View details

Kling v2.1

Kuaishou

580 credits

Kling v2.1 with 720p/1080p support and frame transition capabilities for smooth, high-quality video generation.

SlowUltra
View details

Pruna P-Video

Pruna AI

240 credits

PrunaAI's fast video generator with a built-in draft mode for rapid creative iteration. Text-to-video, image-to-video, and audio-conditioned generation in a single endpoint, with clips up to 20 seconds — one of the longest durations on the platform.

FastHigh
View details

Pruna P-Video Animate

Pruna AI

1,050 credits

Pruna AI's motion-transfer video model — animates a reference image with the motion and audio of a source video. Optimized for speed and cost (about 5.24s of generation per 1s of output video).

FastHigh
View details

Pruna P-Video Avatar

Pruna AI

870 credits

Pruna AI's talking-avatar video model — animates a single input image to speak provided text (voice_script) or an uploaded audio track, with selectable Gemini-family preset voices, language, and visual delivery prompt.

MediumHigh
View details

PixVerse V6

PixVerse

810 credits

PixVerse's flagship video generation model. Generates cinematic videos with synchronized audio, multi-shot sequences with scene transitions, astonishing physics, and precise camera control at up to 1080p.

MediumHigh
View details

PixVerse V5

PixVerse

700 credits

Advanced video generation with special effects capabilities and anime-optimized output, supporting multiple visual styles.

SlowUltra
View details

Luma Ray 3.2

Luma

1,740 credits

Luma's flagship Ray video model. Text-to-video and keyframe (start/end image) generation with optional HDR-encoded output and professional EXR export.

MediumUltra
View details

Luma Ray Flash 2 720p

Luma

700 credits

Luma's Ray Flash 2 generates 5 or 9 second 720p videos faster and cheaper than Ray 2. Supports keyframe control via start and end images, seamless loops, and a rich library of camera motion concepts — the first Luma model on the platform.

FastHigh
View details

MiniMax Hailuo-03 (H3) Reference to Video

MiniMax

1,510 credits

MiniMax Hailuo-03 (H3) reference-to-video via Fal.AI. Generates video from multimodal references — subject/style images, motion video clips, and audio clips, each cited in the prompt by order — keeping subjects consistent while following the referenced motion and audio.

MediumUltra
View details

Seedance 2.5 Reference to Video

ByteDance

5,500 credits

ByteDance Seedance 2.5 reference-to-video via Fal.AI. Generates video from up to 50 multimodal references — up to 30 images, 10 videos, and 10 audio files — locking a character, set, and palette across a full 30-second take. Reference inputs are addressed from the prompt as @Image1, @Video1, @Audio1, and drive motion transfer, editing, extension, and lip-sync.

MediumUltra
View details

Seedance 2.5

ByteDance

5,500 credits

ByteDance's flagship multimodal video model with native audio, native 30-second single-pass generation, and large multimodal reference sets (up to 30 images, 10 videos, 10 audios). Supports text-to-video, image-to-video, first/last-frame control, video editing, extension, and lip-sync.

MediumUltra
View details

Seedance 2.0

ByteDance

1,780 credits

ByteDance's next-generation multimodal video model with native synchronized audio. Combines up to 9 reference images, 3 videos, and 3 audio files in a single generation for character-consistent, lip-synced video creation, editing, and extension.

MediumUltra
View details

Seedance 2.0 Fast

ByteDance

1,430 credits

A faster, cheaper variant of Seedance 2.0 for quicker video generation with multimodal reference inputs (up to 9 images, 3 videos, 3 audios) and native audio, at 480p or 720p.

FastHigh
View details

Seedance 2.0 Mini

ByteDance

890 credits

Lighter, cheaper variant of ByteDance's Seedance 2.0. Native audio, multimodal reference inputs (images/videos/audio), text-to-video and image-to-video, capped at 720p (no 1080p/4K tier).

FastHigh
View details

Seedance 1 Lite

ByteDance

420 credits

A lightweight ByteDance video generation model offering text-to-video and image-to-video support for 4-12 second videos at 480p, 720p, or 1080p resolution.

FastHigh
View details

Seedance 1 Pro

ByteDance

1,740 credits

A pro version of Seedance that offers text-to-video and image-to-video support for 2-12 second videos, at 480p, 720p, and 1080p resolution.

MediumUltra
View details

Seedance 1 Pro Fast

ByteDance

690 credits

ByteDance's cinematic video generation model with fast generation speed and professional output quality.

FastHigh
View details

OpenAI Sora 2

OpenAI

930 credits

OpenAI's video generation model with realistic physics simulation and audio generation capabilities, producing highly coherent videos.

SlowUltra
View details

OpenAI Sora 2 Pro

OpenAI

2,790 credits

OpenAI's most advanced synced-audio video generation model. The premium tier of Sora 2 with higher fidelity, up to 1024p resolution, and image-to-video via an input reference frame.

SlowUltra
View details

MiniMax Hailuo-03 (H3) Text to Video

MiniMax

1,510 credits

MiniMax Hailuo-03 (H3) text-to-video via Fal.AI. State-of-the-art 2K video generation from a text prompt, 5-15 second duration, native audio, and wide aspect-ratio support (21:9 through 9:16).

MediumUltra
View details

Seedance 2.5 Text to Video

ByteDance

5,500 credits

ByteDance Seedance 2.5 text-to-video via Fal.AI. Generates a native single-shot clip of up to 30 seconds at 480p/720p from a text prompt alone, reasoning about the whole shot at once so motion, lighting, and subject identity stay coherent from first frame to last. Synchronized audio (dialogue, sound effects, music) is generated in the same pass.

MediumUltra
View details

Google Veo 3.1

Google

7,440 credits

Google's state-of-the-art video generation model with built-in audio generation, producing cinematic-quality videos with synchronized sound.

SlowUltra
View details

Google Veo 3.1 Fast

Google

2,790 credits

Fast version of Veo 3.1 with audio generation, optimized for speed while maintaining high quality output.

MediumHigh
View details

Wan 3.0

Alibaba

2,330 credits

Alibaba's Wan 3.0 generates video from a text prompt or a starting image, with cinematic motion and support for 480p, 720p, and 1080p output up to 30 seconds.

MediumHigh
View details

Wan 2.7 I2V

Alibaba

1,740 credits

Alibaba Wan 2.7 image-to-video model. Animates a first frame (and optional last frame or continuation clip) with audio synchronization, up to 15 seconds, at 720p or 1080p.

MediumHigh
View details

Wan 2.7 T2V

Alibaba

1,740 credits

Alibaba Wan 2.7 text-to-video model. Supports up to 15 seconds, audio synchronization for voice/music, multilingual prompts, and prompt expansion, at 720p or 1080p.

MediumHigh
View details

Wan 2.5 I2V

Alibaba

1,160 credits

Image-to-video model with lip sync support, animating still images into realistic videos with natural motion.

SlowUltra
View details

Wan 2.5 I2V Fast

Alibaba

790 credits

Fast image-to-video variant of Wan 2.5, optimized for rapid generation of animated videos from still images.

MediumHigh
View details

Wan 2.5 T2V

Alibaba

1,160 credits

Text-to-video model with audio synchronization support, producing high-quality videos from text prompts with natural motion.

SlowUltra
View details

Wan 2.5 T2V Fast

Alibaba

790 credits

Fast text-to-video generation variant of Wan 2.5, optimized for speed with good quality output.

MediumHigh
View details

Wan 2.2 I2V Fast

Alibaba

120 credits

A very fast and cheap PrunaAI-optimized version of Alibaba's Wan 2.2 A14B image-to-video model. Turns a single still image plus a prompt into a short animated clip at 480p or 720p, with an optional frame-interpolation pass for smoother motion.

FastHigh
View details

Wan 2.2 T2V Fast

Alibaba

120 credits

A very fast and cheap PrunaAI-optimized version of Alibaba's Wan 2.2 A14B text-to-video model. Generates short clips at 480p or 720p directly from a text prompt, with 30 FPS frame interpolation enabled by default — the text-to-video sibling of Wan 2.2 I2V Fast.

FastHigh
View details
Text-to-Video

Hailuo 2.3 (MiniMax)

MiniMax의 고품질 비디오 생성 모델.

curl -X POST https://api.core.today/v1/predictions \
  -H "Content-Type: application/json" \
  -H "X-API-Key: cdt_your_api_key" \
  -d '{
    "model": "minimax/hailuo-2.3",
    "input": {
      "prompt": "A golden retriever running on the beach at sunset, slow motion, cinematic",
      "duration": 6,
      "aspect_ratio": "16:9"
    }
  }'

이미지 → 비디오 (Image-to-Video)

첫 프레임 이미지를 제공하여 비디오를 생성합니다.

1. 이미지 업로드

curl -X POST https://api.core.today/v1/files/upload-url \
  -H "X-API-Key: cdt_your_api_key" \
  -H "Content-Type: application/json" \
  -d '{"filename": "first_frame.jpg", "content_type": "image/jpeg"}'

2. 비디오 생성

curl -X POST https://api.core.today/v1/predictions \
  -H "Content-Type: application/json" \
  -H "X-API-Key: cdt_your_api_key" \
  -d '{
    "model": "kwaivgi/kling-v2.5-turbo-pro",
    "input": {
      "prompt": "Camera slowly zooms in, gentle wind blowing",
      "image": "https://files.core.today/.../first_frame.jpg",
      "duration": 5
    }
  }'

카메라 워크 프롬프트

더 나은 비디오를 위해 카메라 움직임을 프롬프트에 포함하세요:

기본 움직임

  • zoom in - 줌 인
  • zoom out - 줌 아웃
  • pan left/right - 패닝
  • tilt up/down - 틸트

고급 움직임

  • tracking shot - 트래킹
  • drone shot - 드론 샷
  • slow motion - 슬로우 모션
  • time lapse - 타임랩스

모델 선택 가이드

저비용 빠른 생성

seedance-1-pro-fast (190크레딧)

고품질 T2V

hailuo-2.3 (420크레딧)

I2V (이미지→비디오)

kling-v2.5-turbo-pro (525크레딧)

최고 품질

veo-3.1 (4,800크레딧)