MiniMax H3 Prompt Guide: Formula, Examples, and Fixes

MiniMax H3
|
Published on Aug 4, 2026

Quick answer

A practical MiniMax H3 prompt names six things: the subject, the action, the environment, the camera, the lighting or visual style, and any constraints. Put the action in time order. If the shot must begin or end on a specific image, use the frame workflow and describe the path between those frames. If images, video, or audio should control identity, motion, camera behavior, or sound, use the reference workflow and assign each file a clear job.

Start with one main action and one camera move. Generate, watch the entire clip, then change one variable. MiniMax does not promise that every negative phrase, seed, or camera term will work exactly as written. Treat prompt language as direction, not a guarantee.

Last verified: August 4, 2026. minimaxh3.tv is an independent service and is not owned, operated, authorized, or endorsed by MiniMax.

Six parts of a MiniMax H3 video prompt

A practical prompt moves from subject and action to setting, camera, visual treatment, and constraints.

The MiniMax H3 prompt formula

Use this as a working formula:

Subject + Action + Environment + Camera + Lighting/Style + Constraints

Part What to write Example
Subject The main person, animal, object, or scene anchor A red paper boat
Action What changes, in the order it happens Drifts forward, catches a current, turns toward the bridge
Environment Place, weather, background activity, and important objects A shallow city canal after rain
Camera Shot size, direction, movement, and speed Low tracking shot, slow push in
Lighting or style Visible treatment rather than vague quality words Overcast daylight, muted documentary color
Constraints Details that should remain stable or actions to avoid Keep the boat red; no people enter the water

The official H3 prompt writing guide also documents a more structured form for advanced timelines. It separates the main audiovisual timeline into integrated_multimodal_description, overall_soundscape, and non_diegetic_music. Use that structure when you need timed shots, dialogue, diegetic sound, or a score. A short browser prompt does not need all three fields.

Text to video prompt template

Text mode has no source image, so the prompt must establish the whole shot.

[Visual style and shot size]. [Subject] begins in [environment].
[Action 1], then [Action 2]. The camera [one camera movement] at [speed].
[Lighting and visible atmosphere]. End with [clear final state].
Keep [identity, object, color, or composition constraint] stable.

Example:

Live-action macro film. A red paper boat floats through a shallow city canal after rain. It drifts forward, catches a current, and turns toward a stone bridge. The camera tracks beside it at slow speed, close to the water. Overcast daylight, muted color, small ripples and wet reflections. End with the boat entering the bridge shadow. Keep the boat red and intact.

Image to video motion template

In the Start and End Frames workflow, the image already defines the subject, layout, color, and opening composition. Spend the prompt on motion and preservation.

Use the uploaded start frame as the opening composition.
[Subject or object] begins by [small first motion], then [continuous action].
The camera [one movement] while [background or lighting response].
Preserve [identity, clothing, shape, layout, or product details].
If an end frame is supplied, reach its pose and composition naturally at the end.

Avoid repeating every visible detail. The official guide recommends a continuous path from the first frame toward the last frame, especially when both are supplied.

Reference workflow template

Reference mode accepts ordered image, video, and audio inputs. Assign each one a role instead of writing "use all references."

Use Image 1 for the subject's identity and clothing.
Use Video 2 only for the camera path and movement pace.
Use Audio 3 for rhythm and timing.
Place the subject in [environment]. The subject [ordered action].
The camera follows the movement from [opening state] to [ending state].
Preserve [identity and object constraints]. Do not copy unrelated details from the references.

Text, frame, and reference prompt workflows

Text mode builds the shot from words. Frame mode describes a path from a fixed image. Reference mode assigns a job to each source asset.

Ten verifiable prompt examples

The examples below have a public source. Only the alpine lake example is a verified minimaxh3.tv prompt and output pair. The other nine are adapted from current MiniMax documentation. They are starting points, not claims about guaranteed results on this site.

# Mode and prompt What it demonstrates Evidence
1 Text: A cinematic sunrise over a calm alpine lake, soft mist drifting above the water, slow camera push forward. Subject, atmosphere, visible motion, and one camera move Verified site prompt and output below
2 Text: A space-opera captain watches the last fleet jump away through an observation window as the bridge shakes. Slow push toward her reflection. End on the empty stars. Event order and ending state Adapted from the official V2 API example
3 First frame: Preserve the dancer and studio from the uploaded image. She begins a contemporary turn, crosses the floor, and settles into a balanced pose. Slow lateral tracking shot. Animate a fixed opening without restating it Adapted from the official video guide
4 First and last frame: Begin from the child's pose in the first image. Show a continuous passage of time through posture, clothing, and lighting, then arrive at the adult pose in the last image. A motion path between two anchors Adapted from the official video guide
5 Reference: Use Image 1 for the model's face and beret. Keep the overcast cobbled alley. She walks forward, adjusts the beret, and smiles as the camera tracks backward. Clear reference roles Adapted from the official reference example
6 Text with sound: Before sunrise, a baker opens wooden shutters and places a hot loaf on the counter. Slow push in. Street ambience, tray clinks, and one doorbell. Sparse acoustic guitar. Visual timeline, soundscape, and score Adapted from the official prompt writing guide
7 First frame with dialogue: Preserve the woman, train seat, folded letter, and rainy window. The camera trucks right slowly as she looks toward the city lights and says one short line. Identity preservation plus a controlled speaking beat Adapted from the official prompt writing guide
8 First and last frame: A cyclist releases the bicycle handle, raises a closed umbrella, opens it, and settles into the exact end-frame pose. Use one slow pull out. Observable intermediate changes Adapted from the official prompt writing guide
9 Last frame: Begin with an intact glass near the table edge. A hand knocks it down. Track the fall and impact, then let the fragments settle into the supplied final composition. Work backward toward a fixed ending Adapted from the official prompt writing guide
10 Text: A runner leaves a tunnel into daylight. Hold a static wide shot until she reaches the exit, then pan right once to reveal the open track. A single camera change tied to new information Adapted from official camera-motion guidance

This existing minimaxh3.tv clip used example 1:

The verified clip shows drifting mist and a slow push forward. It is one prompt and one output, not a success-rate claim or evidence for frame and reference modes.

More verified clips are available in the MiniMax H3 showcase. Use only examples that expose both the prompt and the result when you are studying prompt behavior.

Camera and motion vocabulary

The official guide defines camera motion by type, amplitude, and speed. Add amplitude or speed only when it changes the shot meaningfully.

Term What the camera does Plain prompt phrase
Push in / pull out Moves physically toward or away from the subject The camera pushes in slowly toward the letter
Zoom in / zoom out Changes focal length from a fixed position The lens zooms out to reveal the room
Pan left / right Rotates horizontally from one position The camera pans right to the doorway
Truck left / right Moves sideways The camera trucks left beside the cyclist
Tilt up / down Rotates vertically The camera tilts up from the shoes to the face
Arc shot Moves around the subject The camera arcs halfway around the sculpture
Tracking shot Follows a moving subject A low tracking shot follows the paper boat
Static shot Does not move Hold a static wide shot as the runner exits
POV Uses the subject's viewpoint POV from the passenger seat

Do not stack camera words at the end like tags. Tie the move to an action or reveal. One deliberate move is easier to evaluate than four competing moves in a short clip.

Common prompt failures and single-variable fixes

Symptom Likely cause Change one thing Evidence or test status
The shot looks static The prompt describes appearance but no change over time Add one visible action with a start and end Supported by official timeline guidance and the verified lake example
The subject changes identity Too many new appearance details conflict with a frame or reference Remove new appearance details and state what must remain stable Official keyframe and reference guidance
The camera feels random Several moves compete in one short shot Keep one camera move tied to one reveal Official camera-motion guidance
The ending is abrupt The prompt has actions but no final state Add one clear ending pose, location, or composition Practical revision; not a measured success-rate claim
The frame transition jumps The prompt repeats two static images without a path Describe the intermediate changes between them Official first-and-last-frame guidance
References blend together Each file has no assigned role Name each numbered asset and its job Official reference workflow
Dialogue changes The line is paraphrased inside a long scene description Put the exact line in a short, isolated speaking instruction Official dialogue guidance
A negative phrase has no effect The model may not treat natural-language negatives as hard controls Rewrite the desired positive state and remove one conflict Prompt revision to test; no guaranteed negative-prompt control

Prompt checklist

  • Name one main subject and one main action.
  • Put actions in chronological order.
  • State the environment only where it affects the shot.
  • Use one primary camera move for the first attempt.
  • Describe visible lighting or style instead of saying "high quality."
  • Preserve identity, color, layout, or product details explicitly when needed.
  • Assign a job to every reference file.
  • Add a clear ending state.
  • Check the full prompt for conflicting directions.
  • After a result, change one variable before generating again.

FAQ

How long should a MiniMax H3 prompt be?

Long enough to define the shot, but short enough that the action order stays clear. A simple shot may need three or four sentences. Timed dialogue, several cuts, or multimodal references need more structure. The official API accepts up to 7,000 characters, while the current minimaxh3.tv prompt box has a smaller limit. Follow the limit shown in the interface you use.

Does MiniMax H3 support negative prompts?

The current public site uses natural-language constraints inside the main prompt. Do not assume a separate negative-prompt field or guaranteed negative behavior. Describe the desired positive state and remove conflicting instructions.

Can I specify a seed?

The current minimaxh3.tv generator does not expose a seed control. Do not promise reproducibility from a seed that the interface does not provide.

Should I use camera terms in brackets?

MiniMax documentation includes camera instructions such as pan, zoom, and static. Plain natural-language camera actions also appear in the official H3 prompt guide. Use the form accepted by your current interface, and do not assume a term forces an exact physical result.

What should an image-to-video prompt describe?

Describe motion, camera travel, the ending state, and the details that must remain unchanged. The source image already supplies the opening composition.

How do I keep references from mixing?

Number the files and give each one a job. For example, use one image for identity, one video for motion, and one audio file for timing. Remove any reference that does not contribute to the intended shot.

Should I copy a long prompt exactly?

Copy the structure, then replace the subject, action, environment, and constraints. A prompt tied to another image or reference set may fail when copied without those inputs.

Sources and methodology

This guide was checked on August 4, 2026 against the current MiniMax Video Generation guide, the official Create Video Generation Task reference, the official MiniMax H3 video prompt writing guide, the public MiniMax H3 text-to-video generator, the showcase, and one existing public minimaxh3.tv prompt/output pair. No new paid video generation was used. No external video or community claim is included.

Ready to try a prompt? Open MiniMax H3 Text to Video. For the full browser workflow, read How to Use MiniMax H3.

#MiniMax H3#prompt guide#AI video#text to video
Related Posts
View all articles
How to Use MiniMax H3: Text, Images, and Video References

How to Use MiniMax H3: Text, Images, and Video References

Learn how to use MiniMax H3 with text, start and end frames, or image, video, and audio references using current 2K settings and limits.

MiniMax H3 API Guide: Create, Poll, and Download Video

MiniMax H3 API Guide: Create, Poll, and Download Video

Use the official MiniMax H3 V2 API to create text, frame, or reference video tasks, poll status, download results, and estimate current pricing.

MiniMax H3 Open Source Status: Hugging Face, GitHub, Weights, and License

MiniMax H3 Open Source Status: Hugging Face, GitHub, Weights, and License

Check MiniMax H3 Hugging Face, GitHub, ModelScope, open weights, download, and license status, verified August 1, 2026.

MiniMax H3 Release Date: What Launched and Where to Use It

MiniMax H3 Release Date: What Launched and Where to Use It

MiniMax H3 release date, confirmed launch capabilities, official API limits, and where to use Hailuo H3 online.