Three control sources
Start with text, endpoint frames, or a pack led by at least one image or video.
Create AI video drafts directly in Img2Vid. Choose a compatible model, review price and plan eligibility, then generate in the same workspace.
Use the MiniMax H3 video model to direct a native 2K, 4-15 second shot from a brief, endpoint frames, or a visual-led reference pack. Audio can support a visual source, but cannot be used alone. An active subscription is required.
Model
Prompt
These MiniMax H3 video model studies define one subject, camera route, and visual constraint. They are planning patterns, not verified H3 benchmark results.
Prompt direction
Yellow raincoat, train entering mist. Lateral tracking, warm carriage light, rain, no text.
Inputs and Controls at a Glance
The active MiniMax H3 video model form is the source of truth. It exposes three control surfaces, not one catch-all upload.
Start with text, endpoint frames, or a pack led by at least one image or video.
Use the duration control for one short beat. Native 2K is configured output, not a separate selector.
Text uses six fixed ratios. Reference mode is adaptive or fixed; endpoint mode has no separate selector.
The form exposes watermark choice and a live quote. Reference-image count can affect it.
Input Decision Guide
A MiniMax H3 video model request starts by choosing what controls the shot: written direction, visible endpoints, or a visual-led reference pack.
| Decision | Text Brief | Endpoint Frames | Reference Pack |
|---|---|---|---|
| Start with | A required shot brief | A first frame, last frame, or both | At least one image or video; optional audio |
| Best for | A new scene and camera | A launch, landing, or transition | Identity, motion, scene, or ambience |
| Tell the MiniMax H3 Video Model | Subject, action, camera, light, pace | How endpoints should connect | What each cue should guide |
| Current capacity | Required prompt; max 7,000 characters | 1-2 images; max 30 MB each | Up to 9 images, 3 videos, and 3 audio files |
| Do not assume | A vague mood list defines a shot | Frames guarantee a transition | Audio replaces visual reference |
Multimodal Reference Boundary
The MiniMax H3 video model supports mixed references, but visual material is required. Audio supports visual context, not a standalone soundtrack or dialogue path.
Four-Step H3 Workflow
A MiniMax H3 video model run is a controlled test: choose the control source, prepare useful assets, write one shot, then change one variable.
Use text for exploration, frames for visible states, and a pack when visual context matters.
Keep images clear, trim clips to the useful moment, and remove conflicting assets.
Name subject, action, environment, camera, light, pace, and the detail to inspect after the MiniMax H3 video model generation.
Check against the control source, then alter one prompt detail, asset, duration, or ratio.
Best-Fit Production Tasks
Use the MiniMax H3 video model when a 2K short clip needs a specific control source, not a generic motion layer.
Use MiniMax H3 video model endpoint frames to test a product launch, landing state, reveal, or closing composition.
E-commerce · Launches
Assign image, video, and optional audio references product, setting, motion, or ambience roles.
Campaigns · Moodboards
Use a visual pack when character, costume, creature, or location guides a directed shot.
Storyboards · Concept art
Test camera travel, blocking, atmosphere, or scale before committing a larger budget.
Pitching · Previsualization
Production Boundary
A valid MiniMax H3 video model request is not production approval. The form defines input validity; creators decide whether the result is usable and publishable.
A MiniMax H3 video model first or last frame gives a visible target, but identity, geometry, motion, and transition need inspection.
Reference audio must accompany an image or video. It does not prove dialogue, music generation, or lip-sync.
The live quote changes with the request. Additional reference images above the included count can add cost.
Review source permissions, logos, likenesses, facts, disclosures, and platform rules before publishing.
FAQ
MiniMax H3 video model answers about 2K output, frames, mixed references, naming, pricing, and review.
Create on Img2Vid
Open the MiniMax H3 video model, choose text, frames, or a visual-led pack, then direct one coherent 2K shot.
Explore: Image to Video · AI Video Generator · Hailuo 2.3