--- title: MiniMax-H3 Prompt Rewriter (Omni) emoji: ๐ŸŽฌ colorFrom: red colorTo: pink sdk: gradio sdk_version: 6.26.0 app_file: app.py python_version: "3.12" startup_duration_timeout: 1h pinned: false short_description: Structured MiniMax-H3 audio-video prompts from short ideas --- # MiniMax-H3 Prompt Rewriter ยท Qwen2.5-Omni LoRA Turns a short request โ€” plus optional image, video or audio references โ€” into a structured, production-ready **MiniMax-H3** audio-video prompt. - LoRA adapter: [`lightx2v/MiniMax-H3-Prompt-Rewriter-LoRA-Omni`](https://huggingface.co/lightx2v/MiniMax-H3-Prompt-Rewriter-LoRA-Omni) - Base model: [`Qwen/Qwen2.5-Omni-7B`](https://huggingface.co/Qwen/Qwen2.5-Omni-7B) (Thinker, bf16) - Target video model: [`MiniMaxAI/MiniMax-H3`](https://huggingface.co/MiniMaxAI/MiniMax-H3) **Output is text only.** This is a prompt rewriter, not a video generator โ€” feed the rewritten prompt and the same reference assets into a MiniMax-H3 pipeline to render. ## Tasks | Task | Inputs | Output schema | | --- | --- | --- | | `T2AV` | text only | `integrated_multimodal_description`, `overall_soundscape`, `non_diegetic_music` | | `I2AV` | text + 1 image (exact **first** frame) | same 3 fields | | `L2AV` | text + 1 image (exact **last** frame) | same 3 fields | | `FL2AV` | text + 2 ordered images (first & last frames) | same 3 fields | | `Ref2AV` | text + ordered images / video / audio references | `subject_definitions`, `summary`, `retention_analysis`, `detailed_description`, `overall_soundscape`, `non_diegetic_music` | Duration is requested in whole seconds (4โ€“15) and snapped to MiniMax-H3's legal `17 * n + 5` frame grid at 24 fps. Base tasks accept any aspect preset; `Ref2AV` supports only `16:9` and `9:16`. `Ref2AV` prompts must mention every supplied reference label (``, `