Start a render
MiniMax H3
Audio enabled
0:00 / 0:00

MiniMax H3 lighthouse keeper rescue video with native audio

This MiniMax H3 text-to-video example follows Elara through a storm as she relights a lighthouse, speaks in the lantern room and guides a lifeboat home across four connected shots.

MiniMax H3Text to video15s159:91Enabled$2.41
MiniMax H3Text to video15s159:91Audio

Prompt breakdown

Prompt used to generate this render.

Create a 15-second cinematic 16:9 film with native multi-shot continuity and native stereo sound. Original fictional character: ELARA, a 34-year-old lighthouse keeper with wind-burned olive skin, dark wavy hair tied bac…Show full prompt

Create a 15-second cinematic 16:9 film with native multi-shot continuity and native stereo sound. Original fictional character: ELARA, a 34-year-old lighthouse keeper with wind-burned olive skin, dark wavy hair tied back, a weathered mustard raincoat, charcoal sweater, and practical sea boots. Keep her face, clothing, and proportions consistent in every shot. 0.0-3.5 seconds: wide exterior at blue hour, a stone lighthouse on a violent Atlantic cliff, rain sweeping sideways, Elara runs toward the tower carrying a small weathered signal case; handheld tracking camera, realistic wet fabric and foot contact. 3.5-7.5 seconds: interior spiral stairwell, medium close tracking shot as she climbs fast, breath visible, one hand gripping the rail, practical amber lamps, no cuts that break identity. 7.5-11.5 seconds: lantern room, she reaches the controls and throws the emergency lever; orbit from her determined face to the huge Fresnel lens as it powers on. At 10.5 seconds Elara says clearly, with natural lip sync: “Harbor light is alive. Bring them home.” 11.5-15.0 seconds: exterior aerial pullback as the rescue beam sweeps across black waves and reveals a distant lifeboat turning toward safety. Audio: native stereo only—storm wind outside, rain on glass, boots on iron stairs, strained breathing, lever clunk, electrical rise, deep lighthouse mechanism, distant surf, and Elara's exact spoken line. No music. No subtitles, captions, on-screen text, logos, watermarks, beauty-ad framing, or product pack shot. Photorealistic skin and motion, coherent hands, restrained cinematic contrast.

Workflow

Text to video

Camera

Tracking, Drone

Output

15s · 159:91 · 2K

Recorded render cost

$2.41

Audio

Enabled

Constraints

Text To Video, Multi Shot, Native Audio

Prompt improvement notes

Note 1

Keep the subject, camera move, lighting, duration, aspect ratio and audio requirement grouped so the render has one clear production brief.

Note 2

Change one variable at a time when cloning this prompt: model, duration, camera motion or reference input. That makes quality and price differences easier to compare.

Note 3

Add a short negative prompt if you need to block text overlays, logos, distorted hands, face warping or unwanted camera shake.

Note 4

For multi-shot prompts, keep each beat short and give every cut a clear start state, camera direction and landing frame.

Compare this model

Review this example beside nearby engines before choosing a render path.

Why MiniMax H3 fits this shot

Create 5–15 second character-led videos from text, images, video, and audio references with MiniMax H3 in MaxVideoAI.

Image input

Audio option

15s max

Key frames

Opening frame
Motion beat
Final shot

Related examples

View all examples