You're looking at a specific version of this model. Jump to the model overview.

bghira /minimax-h3-lora-test:9b8e03fb

Input schema

The fields you can use to run this model with an API. If you don’t give a value for a field its default value will be used.

Field Type Default value Description
prompt
string
Text prompt for MiniMax H3 video and audio generation
mode
None
Auto
Task family. Auto uses Ref2VA when reference inputs are provided, otherwise FL2VA/T2VA.
cache_mode
None
Fast
Transformer block-cache strength
image
string
Optional first keyframe for FL2VA generation
last_image
string
Optional last keyframe for FL2VA generation
reference_image
string
Optional Ref2VA image reference
reference_video
string
Optional Ref2VA video reference. A soundtrack in the file is used as that reference's audio.
reference_audio
string
Optional Ref2VA audio reference. Must be paired with an image or video reference.
reference_manifest
string
Optional JSON list of Ref2VA references for URL/path inputs. Entries may be strings or objects with image, video, audio, fps, and sample_rate fields.
reference_order
string
image,video,audio
Order for the simple Ref2VA reference inputs
lora
string
Optional LoRA path, URL, Hugging Face repo id, comma-separated list, or JSON list
lora_scale
number
1

Min: -4

Max: 4

Default LoRA scale
lora_scales
string
Optional comma-separated or JSON list of scales matching lora
aspect_ratio
None
16:9
Output aspect ratio
resolution
None
544p
Output resolution tier
num_frames
integer
124

Min: 120

Max: 345

Frame count. Rounded up to MiniMax H3's 17n+5 frame grid.
num_inference_steps
integer
0

Max: 80

Actual transformer evaluations. Use 0 for the base model default of 20.
seed
integer
0
Random seed. Use 0 for a random seed.
save_compile_cache
boolean
False
Write torch compiler cache artifacts after this run if available

Output schema

The shape of the response you’ll get when you run this model with an API.

Schema
{'format': 'uri', 'title': 'Output', 'type': 'string'}