Get 100 free credits on sign up to explore and build your AI applicationsClaim Free
viduq3

Vidu Q3 / Reference to Video

Commercial
ID: viduq3

Base reference-to-video model built for narrative video creation. Supports native audio-video generation and multi-character dialogue. Delivers robust character consistency and scene coherence across up to 16-second clips

Price$0.019/ Second(19 Credits)

No subjects added yet.

0 / 4 subjects

Input Prompt
0 / 5000
5

Upload Wm Url

JPG, JPEG, PNG (Max 10MB)

Video Generation

Video Playground Ready

Enter prompts in the left parameter panel, configure options, and click Generate.

Vidu Q3 • Reference-to-Video

Vidu Q3Subject-Reference Video with Audio

Vidu Q3 generates reference-conditioned video at 540p–1080p with optional native audio, audio type selection, and watermark position control.

Subject Reference Conditioning
540p / 720p / 1080p
Optional Native Audio
Audio Type Select
ByteDance Seed Foundation Architecture
Commercial License & Enterprise SLA
Vidu Q3 Hero

At a Glance

R2V
Mode
Subject reference
1080p
Max Resolution
Also 540p / 720p
Audio
Native Optional
3 audio types
Seed
Reproducible
Optional pin
Capability Highlights

Vidu Q3 Architecture

Subject-aware multimodal video for production.

Subject Reference Conditioning

Feed subjects so identity stays locked across generated motion variants.

subjectsIdentity Lock
Reference
Audio

Optional Native Audio

Enable audio generation with type selection — all, speech only, or sound effects only.

audiospeech_onlysound_effect_only

540p to 1080p

Draft at 540p, ship social at 720p, and promote heroes to 1080p.

540p720p1080p
540p to 1080p
Watermark Position Control

Watermark Position Control

Optional watermark with selectable corner positions (1–4) and custom watermark URL.

wm_positionwm_url
How It Works

How It Works

From subject reference to finished clip.

01

Prepare Subjects

Collect reference subjects that define identity.

02

Write the Motion Brief

Describe the action and camera language.

03

Set Resolution & Audio

540p/720p/1080p; optional audio type.

04

Watermark & Deliver

Optionally place a watermark and pull the finished clip.

Vidu Q3 Domains

Where subject identity and sound ship together.

Content

Character Motion

Reference-conditioned character clips.

R2V
Growth

Social with Sound

Native audio for Reels and TikTok.

Audio
Commerce

Product Film

Reference-conditioned product motion.

Subjects
Platform

Branded Watermarks

Corner position and custom watermark URL.

wm_position
Best Practices

Prompt Tips

Cleaner Vidu Q3 output.

One subject per reference

Subject references work best when each image has a single clear job.

Mention acoustic mood

If audio is on, name environmental cues so sound matches picture.

Draft at 540p

Explore motion cheaply, then promote to 1080p for finals.

Developer Quickstart

Vidu Q3 Quickstart

Subject-reference video generation.

curl -X POST "https://api.powertokens.ai/v1/videos" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "viduq3",
  "prompt": "Cinematic tracking shot through a neon cyberpunk city in heavy rain.",
  "seconds": "5",
  "size": "1080p",
  "ratio": "16:9",
  "resolution": "1080p",
  "duration": 5,
  "aspect_ratio": "16:9",
  "audio": true
}'

Technical Specifications

Confirmed parameters and runtime execution protocols.

Provider & Model ID
Vidu • viduq3
Mode
Reference-to-video
Resolutions
540p, 720p, 1080p (default 540p)
Duration
Configurable via duration slider (default 5)
Aspect Ratios
16:9, 9:16, 1:1 (default 16:9)
Audio
Optional; all / speech_only / sound_effect_only
Seed
Optional reproducible seed
Watermark
Optional with wm_position 1–4 and wm_url
Billing
See pricing matrix
API Endpoint
POST /v1/videos

Pricing Details

The actual billing for this model is dynamically calculated based on the specific parameters passed in your API request. Below are the specific combinations and their corresponding pricing:

Reference to Video540PPeak Shifting
Credits19/ Second
Price (USD)$0.019
Reference to Video540P
Credits34/ Second
Price (USD)$0.034
Reference to Video720PPeak Shifting
Credits29/ Second
Price (USD)$0.029
Reference to Video720P
Credits57/ Second
Price (USD)$0.057
Reference to Video1080PPeak Shifting
Credits34/ Second
Price (USD)$0.034
Reference to Video1080P
Credits72/ Second
Price (USD)$0.072
Ecosystem Models

Recommended Related Models

Explore complementary video and multimodal models with your unified API key.

Browse All Models
Seedance 2.5
videoCommercial
Seedance 2.5

Seedance 2.5

dreamina-seedance-2-5-260628

ByteDance's latest flagship video generation model, built for longer-form storytelling and production-ready output. Generates up to 30 seconds of continuous, cinematic video with native audio sync in a single pass . Accepts up to 50 multimodal references (images, videos, audio, character sheets, storyboards) for precise scene, character, and motion consistency . Features localized region editing to fix specific areas without full regeneration……

Text to VideoImage to Video
Wan 3.0
videoCommercial
Wan 3.0

Wan 3.0

wan3.0-video

Wan3.0-Video is an all-in-one video generation model unified support for multiple creative capabilities, including reference, editing, replication, and driving. It generates videos up to 30 seconds with omni-modal reference, and can parse files, web pages and complex images. With production-grade character consistency and lifelike visuals and sound, it delivers an immersive audiovisual experience.

Text to VideoImage to Video
Seedance 2.0 Mini
videoCommercial
Seedance 2.0 Mini

Seedance 2.0 Mini

dreamina-seedance-2-0-mini-260615

Lightweight, cost-efficient video model from ByteDance, optimized for speed and high-volume content creation. Supports text-to-video, image-to-video, and reference-based generation with up to 12 references (6 images, 3 audio, 3 video). Delivers faster generation and lower credit consumption than Seedance 2.0, with strong motion quality and character consistency. Ideal for social media content, product videos, AI short dramas, and rapid creative iteration

Text to VideoImage to Video
Seedance 2.0
videoCommercial
Seedance 2.0

Seedance 2.0

dreamina-seedance-2-0-260128

Generate videos from reference images, videos, and audio; edit videos; extend videos; generate videos from start and end frames

Text to VideoImage to Video

Frequently Asked Questions

Everything you need to know before integrating this model.

Feed subjects so identity stays locked across generated motion variants.

Start Building with Vidu Q3 Today

Create an account in seconds to receive 100 free credits and start generating immediately. No credit card or upfront contract required.