
MiniMax H3 / 텍스트로 비디오
Open-weight general-purpose multimodal video model. Unifies text, image, video, and audio understanding in a single context window, generating up to 2K resolution, 15-second clips with native stereo audio at 24fps. Supports multimodal reference inputs: up to 9 images, 3 videos, and 3 audio clips (12 total references) per generation. Features first-frame, last-frame, and full reference modes with conversational editing capabilities.
Video Playground Ready
왼쪽 매개변수 패널에서 프롬프트를 입력하고 옵션을 설정한 후 Generate를 클릭하세요.

MiniMax H35-Mode Cinematic Video with 2K Output
MiniMax H3 is a multimodal video foundation model with five generation modes — text, image, end-frame, start-and-end-frame, and reference conditioning — plus native 2K resolution and up to 15-second clips.
Multimodal Video Architecture
One model for every conditioning path — from pure text to multi-reference cinematography.
Five Conditioning Modes
Text-to-video, image-to-video, end-frame, start-and-end-frame, and multimodal reference-to-video — switch modes without changing providers.


Native 2K Resolution
Ship 768P for drafts and social, or promote hero placements to 2K without a separate upscaler pass.
Start & End Frame Bridges
Land exactly on a target end frame for match cuts, product morphs, and seamless scene transitions.


Cinematic Ratio Suite
Text-to-video supports 21:9, 16:9, 4:3, 1:1, 3:4, and 9:16; reference mode also allows adaptive framing.
How It Works
From brief to finished multimodal clip.
Pick a Mode
Start from text, a still, an end frame, both frames, or multimodal references.
Set Canvas & Length
Choose 768P or 2K, aspect ratio, and 4–15 second duration for the placement.
Condition with References
In reference mode, feed stills or motion cues that lock identity and style.
Deliver & Chain
Pull the finished clip and bridge into the next shot with end-frame continuity.
H3 Production Domains
Where multimodal control replaces stitched pipelines.
Brand Film Previs
Reference-conditioned previs before full live-action shoots.
Product Morph Ads
End-frame and dual-frame transitions for packshot reveals.
Ultra-Wide Storytelling
Native 21:9 cinematic boards for hero brand placements.
Social Vertical Cuts
9:16 clips up to 15 seconds for Reels and TikTok.
Prompt & Usage Tips
Get cleaner first-pass H3 generations.
If you need a bridge, mention the end composition in the prompt so the model aims for that landing frame.
Draft at 768P, promote only final hero placements to 2K to control spend.
Reference mode is strongest when each image has a single clear job — identity, style, or scene.
H3 Multimodal Quickstart
Submit text or reference-driven clips with native 2K support.
curl -X POST "https://api.powertokens.ai/v1/videos" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "MiniMax-H3",
"prompt": "Cinematic tracking shot through a neon cyberpunk city in heavy rain.",
"seconds": "5",
"size": "1080p",
"ratio": "16:9",
"resolution": "2K",
"duration": 8
}'Technical Specifications
Confirmed parameters and runtime execution protocols.
가격 상세
이 모델의 실제 요금은 API 요청에서 전달된 특정 매개변수를 기반으로 동적으로 계산됩니다. 아래는 구체적인 조합과 해당 가격입니다:
참고: 각 요청의 처음 5개 입력 이미지는 무료입니다. 이후 입력 이미지는 표에 표시된 입력 가격에 따라 청구됩니다.
요금 청구 규칙: 입력 비디오와 출력 비디오 모두 비디오 초당 요금이 청구되며, 청구 시간 = 입력 비디오 시간 + 출력 비디오 시간입니다.
| 모달리티 | 입력 | 출력 |
|---|---|---|
| 768P | $0.040 40/ Image | $0.080 80/ Second |
| 2K | $0.040 40/ Image | $0.130 130/ Second |
Models from the Same Channel
Explore complementary models and alternative versions from the same provider channel.


MiniMax Hailuo 2.3 Fast
MiniMax-Hailuo-2.3-Fast
A streamlined AI video model prioritizing speed and cost-efficiency. Generates 6-second 768p videos rapidly with 50% lower batch costs. Maintains solid motion physics and stylization for quick iterations, drafts, and short-form content.


MiniMax Hailuo 2.3
MiniMax-Hailuo-2.3
A flagship AI video model delivering breathtaking motion and lifelike emotion. Excels in fluid character movements, cinematic lighting, and multi-style support (anime, ink wash). Produces 1080p, 6/10-second videos with natural micro-expressions for high-fidelity creative work.


MiniMax Hailuo 02
MiniMax-Hailuo-02
A foundational AI video model with top-tier physics simulation and temporal consistency. Excels at dynamic scenes and clear subject rendering, supporting diverse art styles. Ideal for reliable, high-quality video generation across creative and commercial projects.
Recommended Related Models
Explore complementary video and multimodal models with your unified API key.


Seedance 2.5
dreamina-seedance-2-5-260628
🔥 LIMITED-TIME OFFER: 1080P at 20% OFF! 🔥 ByteDance's latest flagship video generation model, built for longer-form storytelling and production-ready output. Generates up to 30 seconds of continuous, cinematic video with native audio sync in a single pass . Accepts up to 50 multimodal references (images, videos, audio, character sheets, storyboards) for precise scene, character, and motion consistency . Features localized region editing to fix specific areas without full regeneration……


Wan 3.0
wan3.0-video
Wan3.0-Video is an all-in-one video generation model unified support for multiple creative capabilities, including reference, editing, replication, and driving. It generates videos up to 30 seconds with omni-modal reference, and can parse files, web pages and complex images. With production-grade character consistency and lifelike visuals and sound, it delivers an immersive audiovisual experience.


Seedance 2.0 Mini
dreamina-seedance-2-0-mini-260615
Lightweight, cost-efficient video model from ByteDance, optimized for speed and high-volume content creation. Supports text-to-video, image-to-video, and reference-based generation with up to 12 references (6 images, 3 audio, 3 video). Delivers faster generation and lower credit consumption than Seedance 2.0, with strong motion quality and character consistency. Ideal for social media content, product videos, AI short dramas, and rapid creative iteration


Seedance 2.0
dreamina-seedance-2-0-260128
Generate videos from reference images, videos, and audio; edit videos; extend videos; generate videos from start and end frames
Frequently Asked Questions
Everything you need to know before integrating this model.
Start Building with MiniMax H3 Today
Create an account in seconds to receive 100 free credits and start generating immediately. No credit card or upfront contract required.