Put the subject and action first
Search-like keyword lists do not direct motion well. Start with who or what is on screen and what changes during the shot.
Turn a written shot brief into a 2K AI video. Direct the subject, action, setting, camera, light, pacing, and sound in one focused prompt—no source image required.
Prompt anatomy
Give each phrase a job. The example below moves from subject and action to camera, light, pacing, and sound.
Prompt formula
Subject + action + setting + camera + lighting + pacing + sound
Example prompt
A lone cyclist crosses a rain-darkened bridge at blue hour. Slow side-tracking camera, wet reflections, restrained pace, soft tire spray, distant city ambience, cinematic natural light.
Open the generator and keep MiniMax H3 selected. The text workflow is preselected on this page.
Name the subject and action first, then add setting, camera movement, lighting, pacing, and sound.
Choose an available aspect ratio, duration, and resolution for the destination where the clip will be used.
Review motion and framing. Change one important instruction at a time so you can identify what improved the result.
Start without source media
Text-to-video is the right starting mode when the scene exists as an idea rather than an asset. MiniMax H3 interprets the prompt as a shot plan, so concrete visual instructions usually outperform abstract adjectives. Use it for concept exploration first; move to image or reference mode when an exact composition or recognizable subject matters more than variation.
Search-like keyword lists do not direct motion well. Start with who or what is on screen and what changes during the shot.
Use directions such as slow dolly in, locked tripod, overhead drift, or side tracking instead of only saying cinematic.
A short clip has limited time. One readable action and one camera idea usually produce a more coherent result than several cuts.
One-time MiniMax H3 credits
Choose a one-time credit pack for MiniMax H3 text-to-video generation. Every paid pack supports the same 2K prompt-to-video workflow without a recurring subscription.
Clear answers
It is a workflow that creates a video from a written scene description without requiring an uploaded source image.
Describe the subject, action, setting, camera movement, lighting, pacing, and sound in that order. Keep the brief focused on one coherent shot.
MiniMax H3 prompts can describe both the picture and the audio direction. Include voice, ambience, music, or sound-effect cues when sound is part of the intended shot.
MiniMax H3 includes 2K output options. The generator displays the resolution, duration, and credit cost available for the selected setup before generation.
Use image-to-video when you already have a composition, product still, illustration, or keyframe that should anchor the generated motion.