Skip to content
Get started
GENERATION API CALLS
Video Generation

MiniMax

This page is auto-generated from model configurations. Last updated: 2026-09-03.

This reference lists all available MiniMax video generation models and their parameters. Use these parameter names when calling the Generation API.


A cinematography model offering advanced camera control for cinematic storytelling and professional framing.

Model ID: model_minimax-video-01-director

Capabilities: txt2video, img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-video-01-director/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe your video. Try Prompt Spark or check our Help Center for detailed prompt tips for this model.
firstFrameImage file No - - - - Image used as the first frame of the video.
promptOptimizer boolean No true - - - Use prompt optimizer

MiniMax H3 2K / 24 fps video generation with native stereo audio: text, first/last frame, or multimodal references (images, videos, optional audio).

Model ID: model_minimax-h3

Capabilities: txt2video, img2video, video2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-h3/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe the video you want.
firstFrameImage file No - - - - An image to use as the video’s opening frame. The video’s shape follows this image. Can’t be combined with reference images, videos, or audio.
lastFrameImage file No - - - - An image to use as the video’s closing frame. Only works when a first frame is also set. Can’t be combined with reference images, videos, or audio.
referenceImages file_array No - - - - Images to guide the subjects or style (up to 9, max 30 MB each). Can’t be combined with a first or last frame. First 5 are free; additional images are billed.
referenceVideos file_array No - - - - Videos to guide the motion (up to 3, each 2–15s and 2–15s in total, max 50 MB each). Can’t be combined with a first or last frame. Billed per second of uploaded duration.
referenceAudio file_array No - - - - Audio to guide the video (up to 3, each 2–15s and 2–15s in total, max 15 MB each). Requires at least one reference image or video.
duration number No 5 5 15 - How long the video lasts, in seconds (5-15). Longer videos cost more.
resolution string No 2K - - 768P, 2K The output quality. Higher resolutions cost more.
aspectRatio string No adaptive - - 21:9, 16:9, 4:3, 1:1, 3:4, 9:16, adaptive The shape of the video. Auto picks a fitting shape. Ignored when a first frame is set, since the shape follows that image.

Fal’s post-trained MiniMax H3 Max image-to-video: animate a first frame (optional last frame), 5–15s at 480p or 768p.

Model ID: model_minimax-h3-max-i2v

Capabilities: img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-h3-max-i2v/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe the video you want.
firstFrameImage file Yes - - - - An image to use as the video’s opening frame. The video’s shape follows this image.
lastFrameImage file No - - - - An image to use as the video’s closing frame. Only works when a first frame is also set.
duration number No 5 5 15 - How long the video lasts, in seconds (5-15). Longer videos cost more.
resolution string No 768P - - 480P, 768P The output quality. Higher resolutions cost more.
promptExpansionMode string No balanced - - disabled, balanced, quality How much to rewrite the prompt before generation. Disabled skips it. Balanced takes about a second. Quality spends up to ~30s on a richer prompt.
seed number No - - - - Use a seed for reproducible results. Leave blank for a random seed.

MiniMax H3 Max reference-to-video via Fal: stronger prompt adherence and aesthetics from images, videos, and optional audio, with native stereo sound at 480P or 768P.

Model ID: model_minimax-h3-max-reference-to-video

Capabilities: img2video, video2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-h3-max-reference-to-video/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe the video. Refer to uploads by order: Image 1, Image 2, Video 1, Audio 1, and so on.
referenceImages file_array No - - - - Images to guide subjects or style (up to 9, max 30 MB each). Refer to them as Image 1, Image 2, and so on. Combined images, videos, and audio must not exceed 12 files. First 4 (1024×1024) are included; additional images are billed.
referenceVideos file_array No - - - - Videos to guide motion (up to 3, each 2–15s and 2–15s in total, max 50 MB each). Refer to them as Video 1, Video 2, and so on. Billed per second of uploaded duration at the output resolution.
referenceAudio file_array No - - - - Audio to guide the video (up to 3, each 2–15s and 2–15s in total, max 15 MB each). Requires at least one reference image or video.
duration number No 5 5 15 - How long the video lasts, in seconds (5–15). Longer videos cost more.
resolution string No 768P - - 480P, 768P The native generation resolution. Reference video billing depends on this value.
aspectRatio string No adaptive - - 21:9, 16:9, 4:3, 1:1, 3:4, 9:16, adaptive The shape of the video. Auto picks a fitting shape.
promptExpansionMode string No balanced - - balanced, quality How much to rewrite the prompt before generation. Balanced returns in about a second; quality spends up to ~30s on a richer prompt.
seed number No - 0 2147483647 - Optional seed for reproducible results. A random seed is used when omitted.

Fal’s post-trained MiniMax H3 Max text-to-video: stronger prompt adherence and aesthetics, 5–15s at 480p or 768p.

Model ID: model_minimax-h3-max-t2v

Capabilities: txt2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-h3-max-t2v/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe the video you want.
duration number No 5 5 15 - How long the video lasts, in seconds (5-15). Longer videos cost more.
resolution string No 768P - - 480P, 768P The output quality. Higher resolutions cost more.
aspectRatio string No 16:9 - - 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 The shape of the video.
promptExpansionMode string No balanced - - disabled, balanced, quality How much to rewrite the prompt before generation. Disabled skips it. Balanced takes about a second. Quality spends up to ~30s on a richer prompt.
seed number No - - - - Use a seed for reproducible results. Leave blank for a random seed.

Fal’s post-trained MiniMax H3 Max Turbo image-to-video: faster inference with stronger prompt adherence, 5–15s at 480p or 768p.

Model ID: model_minimax-h3-max-turbo-i2v

Capabilities: img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-h3-max-turbo-i2v/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe the video you want.
firstFrameImage file Yes - - - - An image to use as the video’s opening frame. The video’s shape follows this image.
lastFrameImage file No - - - - An image to use as the video’s closing frame. Only works when a first frame is also set.
duration number No 5 5 15 - How long the video lasts, in seconds (5-15). Longer videos cost more.
resolution string No 768P - - 480P, 768P The output quality. Higher resolutions cost more.
promptExpansionMode string No balanced - - disabled, balanced, quality How much to rewrite the prompt before generation. Disabled skips it. Balanced takes about a second. Quality spends up to ~30s on a richer prompt.
seed number No - - - - Use a seed for reproducible results. Leave blank for a random seed.

Fal’s post-trained MiniMax H3 Max Turbo text-to-video: faster inference with stronger prompt adherence, 5–15s at 480p or 768p.

Model ID: model_minimax-h3-max-turbo-t2v

Capabilities: txt2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-h3-max-turbo-t2v/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe the video you want.
duration number No 5 5 15 - How long the video lasts, in seconds (5-15). Longer videos cost more.
resolution string No 768P - - 480P, 768P The output quality. Higher resolutions cost more.
aspectRatio string No 16:9 - - 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 The shape of the video.
promptExpansionMode string No balanced - - disabled, balanced, quality How much to rewrite the prompt before generation. Disabled skips it. Balanced takes about a second. Quality spends up to ~30s on a richer prompt.
seed number No - - - - Use a seed for reproducible results. Leave blank for a random seed.

Model ID: model_minimax-hailuo-02

Capabilities: txt2video, img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-hailuo-02/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe your video
firstFrameImage file No - - - - Image used as the first frame of the video. The output video will have the same aspect ratio as this image.
lastFrameImage file No - - - - Used to generate a video that transitions from the first frame to this image. Requires a first frame image.
duration number No 6 - - 6, 10 Duration of the video in seconds. 10 seconds is only available for 768p resolution.
resolution string No 1080p - - 768p, 1080p Pick between standard 768p, or pro 1080p resolution. The pro model is not just high resolution, it is also higher quality.
promptOptimizer boolean No true - - - Use prompt optimizer

A high-fidelity video generation model optimized for realistic human motion, cinematic VFX, expressive characters, and strong prompt and style adherence across both text-to-video and image-to-video workflows

Model ID: model_minimax-hailuo-2-3

Capabilities: txt2video, img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-hailuo-2-3/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Text prompt for generation
firstFrameImage file No - - - - First frame image for video generation. The output video will have the same aspect ratio as this image.
duration number No 6 - - 6, 10 Duration of the video in seconds. 10 seconds is only available for 768p resolution.
resolution string No 768p - - 768p, 1080p Pick between 768p or 1080p resolution. 1080p supports only 6-second duration.
promptOptimizer boolean No true - - - Use prompt optimizer

A lower-latency image-to-video version of Hailuo 2.3 that preserves core motion quality, visual consistency, and stylization performance while enabling faster iteration cycles.

Model ID: model_minimax-hailuo-2-3-fast

Capabilities: txt2video, img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-hailuo-2-3-fast/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Text prompt for generation
firstFrameImage file No - - - - First frame image for video generation. The output video will have the same aspect ratio as this image.
duration number No 6 - - 6, 10 Duration of the video in seconds. 10 seconds is only available for 768p resolution.
resolution string No 768p - - 768p, 1080p Pick between 768p or 1080p resolution. 1080p supports only 6-second duration.
promptOptimizer boolean No true - - - Use prompt optimizer

Minimax Video-01 is a versatile model generating high-quality videos with 720p resolution and smooth motion.

Model ID: model_minimax-video-01

Capabilities: txt2video, img2video

LLM Markdown: https://app.scenario.com/api/models/model_minimax-video-01/markdown

Parameter Type Required Default Min Max Allowed Values Description
prompt string Yes - - - - Describe your video
firstFrameImage file No - - - - Image used as the first frame of the video.
subjectReference file No - - - - An optional character reference image to use as the subject in the generated video
promptOptimizer boolean No true - - - Use prompt optimizer