
Vidu Q3 Pro / テキストから動画
Flagship video model supporting text-to-video, image-to-video, and start/end-frame workflows. Generates up to 16-second clips with native audio sync and storyboard capabilities. Available in 540p–1080p resolutions with premium motion dynamics
Upload Wm Url
JPG, JPEG, PNG (Max 10MB)
Video Playground Ready
左側のパラメータパネルでプロンプトを入力し、設定を行って Generate をクリックしてください。
料金詳細
このモデルの実際の課金は、API リクエストで渡される特定のパラメータに基づいて動的に計算されます。以下は具体的な組み合わせとそれに対応する料金です:
| モダリティ | クレジット | 料金 (USD) |
|---|---|---|
| Text to Video/540P/Peak Shifting | 24/ Second | $0.024 |
| Text to Video/540P | 43/ Second | $0.043 |
| Text to Video/720P/Peak Shifting | 48/ Second | $0.048 |
| Text to Video/720P | 95/ Second | $0.095 |
| Text to Video/1080P/Peak Shifting | 57/ Second | $0.057 |
| Text to Video/1080P | 114/ Second | $0.114 |
| Image to Video/540P/Peak Shifting | 24/ Second | $0.024 |
| Image to Video/540P | 43/ Second | $0.043 |
| Image to Video/720P/Peak Shifting | 48/ Second | $0.048 |
| Image to Video/720P | 95/ Second | $0.095 |
| Image to Video/1080P/Peak Shifting | 57/ Second | $0.057 |
| Image to Video/1080P | 114/ Second | $0.114 |
| Start&End Frame to Video/540P/Peak Shifting | 24/ Second | $0.024 |
| Start&End Frame to Video/540P | 43/ Second | $0.043 |
| Start&End Frame to Video/720P/Peak Shifting | 48/ Second | $0.048 |
| Start&End Frame to Video/720P | 95/ Second | $0.095 |
| Start&End Frame to Video/1080P/Peak Shifting | 57/ Second | $0.057 |
| Start&End Frame to Video/1080P | 114/ Second | $0.114 |
Models from the Same Channel
Explore complementary models and alternative versions from the same provider channel.


Vidu Q3 Turbo
viduq3-turbo
Optimized for speed, generating 1–16 second video clips from text or images with faster inference than the Pro variant. Supports 540p–1080p output, balancing generation quality with reduced latency for rapid iteration


Vidu Q3 Pro Fast
viduq3-pro-fast
High-efficiency variant optimized for cost and generation speed, ideal for high-volume creative pipelines. Supports 1–16 second image-to-video at 720p or 1080p, offering the most affordable per-second pricing in the Q3 series


Vidu Q3
viduq3
Base reference-to-video model built for narrative video creation. Supports native audio-video generation and multi-character dialogue. Delivers robust character consistency and scene coherence across up to 16-second clips
Recommended Related Models
Explore complementary video and multimodal models with your unified API key.


Seedance 2.5
dreamina-seedance-2-5-260628
🔥 LIMITED-TIME OFFER: 1080P at 20% OFF! 🔥 ByteDance's latest flagship video generation model, built for longer-form storytelling and production-ready output. Generates up to 30 seconds of continuous, cinematic video with native audio sync in a single pass . Accepts up to 50 multimodal references (images, videos, audio, character sheets, storyboards) for precise scene, character, and motion consistency . Features localized region editing to fix specific areas without full regeneration……


Wan 3.0
wan3.0-video
Wan3.0-Video is an all-in-one video generation model unified support for multiple creative capabilities, including reference, editing, replication, and driving. It generates videos up to 30 seconds with omni-modal reference, and can parse files, web pages and complex images. With production-grade character consistency and lifelike visuals and sound, it delivers an immersive audiovisual experience.


Seedance 2.0 Mini
dreamina-seedance-2-0-mini-260615
Lightweight, cost-efficient video model from ByteDance, optimized for speed and high-volume content creation. Supports text-to-video, image-to-video, and reference-based generation with up to 12 references (6 images, 3 audio, 3 video). Delivers faster generation and lower credit consumption than Seedance 2.0, with strong motion quality and character consistency. Ideal for social media content, product videos, AI short dramas, and rapid creative iteration


Seedance 2.0
dreamina-seedance-2-0-260128
Generate videos from reference images, videos, and audio; edit videos; extend videos; generate videos from start and end frames