팁: Start with a storyboard creation, then use 'Create Video Project' to generate clips for each scene. You can reorder, regenerate, or skip individual clips before finalizing.
동영상 AI 모델
각 동영상 클립을 생성하는 데 사용되는 AI 모델입니다. 표시된 비용은 동영상 초당 비용입니다.
Hailuo 23 -16 /s
Advanced physics simulation for realistic complex movements. Supports end frame for loop creation.
Seedance 1.0 Pro Fast3 -12 /s
Cheaper, faster variant of Seedance 1.0 Pro via BytePlus. ~50% cost savings vs the standard 1.0 Pro at the same resolutions.
Seedance 1.5 Pro3 -12 /s
ByteDance Seedance 1.5 Pro via BytePlus. High-precision audio-visual sync, cinematic motion, and emotional expression. Supports first/last frames.
Seedance 1.0 Lite I2V4 -14 /s
ByteDance Seedance 1.0 Lite Image-to-Video — efficient, cost-effective animation of source images. Direct via BytePlus.
Dreamina Seedance 2.0 Fast5 -20 /s
Faster, cheaper sibling of Seedance 2.0 via BytePlus. Same multimodal capability set (text + image + video + audio references, native sync audio, multi-shot narration) at roughly 40% lower per-second cost. Trades a bit of fidelity for speed — ideal for iteration loops.
Seedance 1.0 Pro6 -30 /s
ByteDance Seedance 1.0 Pro via BytePlus. Comprehensive and powerful video generation with strong motion control.
P-Video8 /s
Pruna P-Video — fast, affordable text-to-video at $0.02/sec. Standard aspect ratios, 3-15 second clips. Strong bang-for-buck for prototypes and short-form content.
Grok Imagine Video10 /s
xAI Imagine API video — fast text-to-video and image-to-video at a flat $0.05/sec. Async polling pattern; supports 720p, 1-15 second durations. Native audio is included on every generation (cannot be turned off).
Wan 2.6 I2V Flash10 -15 /s
Fast image-to-video with optional audio sync. Faster inference than standard Wan 2.6 I2V. Up to 15 seconds.
Dreamina Seedance 2.014 -69 /s
ByteDance flagship multimodal video model via BytePlus — accepts text + reference images + reference video + reference audio. Native synchronized audio, pro camera controls, multi-shot narration. Cheaper than the Replicate route and exposes capabilities Replicate hides.
Happy Horse 1.0 I2V14 /s
Direct-to-DashScope Happy Horse 1.0 image-to-video. Strict consistency with the source image, fluent natural motion, native audio. 3-15 seconds.
Happy Horse 1.0 T2V14 /s
Direct-to-DashScope Happy Horse 1.0 text-to-video. Cheaper than the Replicate path. 3-15 second durations, five aspect ratios, native audio always on.
Happy Horse 1.0 R2V20 /s
Reference-to-video — combines up to 9 reference images for strong subject + scene consistency. Direct-to-DashScope only (Replicate proxy hides this).
Veo 3.1 Fast20 /s
Google's Veo 3.1 Fast with native audio and frame-to-frame generation. Supports start and end frames for seamless transitions.
Wan 2.6 I2V20 -30 /s
Alibaba Wan 2.6 image-to-video with multi-shot storytelling, native audio, and precise lip-sync. Up to 15 seconds.
Wan 2.6 T2V20 -30 /s
Alibaba's latest text-to-video model with multi-shot storytelling, native audio, and precise lip-sync. Up to 15 seconds.
Dreamina Seedance 2.521 -46 /s
The newest ByteDance multimodal video model via BytePlus. Up to 30-second multi-shot narratives, precision in-video editing, and up to 50 reference assets (images, video, audio) per prompt. Native synchronized audio in 10+ languages. 480p and 720p only — no 1080p tier.
Happy Horse 1.0 Video Edit24 /s
Local or global edits to an existing video using natural-language instructions and up to 5 reference images. Preserves original motion. Direct-to-DashScope only.
Kling v334 -45 /s
Kuaishou Kling v3 — cinematic text-to-video and image-to-video up to 15 seconds with native audio and lip-synced dialogue. Supports start and end frames. Standard mode = 720p, Pro mode = 1080p.
Veo 3.140 /s
Google's flagship video model with strongest prompt adherence and cinematic motion. Synchronized native audio, reference images, and start+end frame control.
Kling v2.1 Master56 /s
Premium Kling model with enhanced quality and longer durations.
품질
클립당 길이
5 seconds per clip0 /clip
가로세로 비율
스토리보드에서 상속됨. 이를 변경하면 시각적 일관성에 영향을 줄 수 있습니다.
unspecifiedUnspecified
16:9Wide
9:16Tall
개인정보 모드
비공개
이 창작물은 나만 볼 수 있습니다. 기본 설정입니다.
일부 공개
링크를 보낸 사람은 계정이 없어도 누구나 열어 볼 수 있습니다. 검색 엔진에는 노출되지 않습니다.
팀
이 창작물을 팀원에게만 공유합니다. 팀 외부의 다른 사람은 접근할 수 없습니다.
출력 언어
자동 (입력 언어)
English (United States)
Spanish (Spain)
French (France)
German (Germany)
Italian (Italy)
Portuguese (Brazil)
Arabic (Generic)
Bengali (India)
Bulgarian (Bulgaria)
Croatian (Croatia)
Czech (Czech Republic)
Danish (Denmark)
Dutch (Belgium)
Dutch (Netherlands)
Estonian (Estonia)
Finnish (Finland)
Greek (Greece)
Gujarati (India)
Hebrew (Israel)
Hindi (India)
Hungarian (Hungary)
Indonesian (Indonesia)
Japanese (Japan)
Kannada (India)
Korean (South Korea)
Latvian (Latvia)
Lithuanian (Lithuania)
Malayalam (India)
Mandarin Chinese (China)
Marathi (India)
Norwegian Bokmål (Norway)
Polish (Poland)
Romanian (Romania)
Russian (Russia)
Serbian (Cyrillic)
Slovak (Slovakia)
Slovenian (Slovenia)
Swahili (Kenya)
Swedish (Sweden)
Tamil (India)
Telugu (India)
Thai (Thailand)
Turkish (Turkey)
Ukrainian (Ukraine)
Urdu (India)
Vietnamese (Vietnam)
1.0
LLM 모델의 창의성 대 일관성을 제어합니다. 기본값: 1.0. 낮을수록 = 집중적/결정적, 높을수록 = 창의적/무작위적.
생각 중
추론 모델은 답변하기 전에 생각합니다. 끄기 = 1×, 켜기 = 2×, 높음 = 모델 기본 호출당 비용의 3×.
끄기1
켜기2
높음3
무료 AI Video Project Generator
Transform your storyboard scenes into a cohesive multi-clip video project. Generate video clips from each scene with visual coherence between clips, then combine them into a single final video.
Perfect For
Filmmakers creating short films from storyboards, content creators producing multi-scene videos, marketers building product demonstration videos, educators creating instructional content with multiple steps, game developers prototyping cutscenes, and storytellers bringing their visual narratives to life.
Key Features
Automatic visual coherence between clips by using last frame of each clip as reference for the next, drag-and-drop clip reordering, per-clip regeneration and prompt editing, skip clips or mark them for exclusion, progress tracking with clip status indicators, FFmpeg-powered video concatenation for the final output, and batch pricing with discounts for multiple clips.