KennethFal commited on
Commit
78a594b
Β·
verified Β·
1 Parent(s): b65a99e

add ComfyUI section: base checkpoint requirement (int8_convrot incompatible), setup, strengths

Browse files
Files changed (1) hide show
  1. README.md +38 -0
README.md CHANGED
@@ -100,6 +100,43 @@ image-to-video LoRA siblings:
100
  480P / 4:3 is the period-correct sweet spot; the style carries fine to higher
101
  resolutions if you want clean pixels of a dirty tape.
102
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
103
  ## Training
104
 
105
  Trained with fal's MiniMax H3 t2v trainer (rank 32, 5,000 steps, 4:3, joint
@@ -114,6 +151,7 @@ A/B testing; more steps past 5k made the damage *tamer*, not heavier.
114
  | file | what |
115
  |---|---|
116
  | [`vh5tape.safetensors`](./vh5tape.safetensors) | the LoRA, rank 32, 5,000 steps |
 
117
  | [`PROMPTS.md`](./PROMPTS.md) | the exact prompts + seeds behind every sample above |
118
 
119
  Made for **H3-TV** β€” a 1970s portable CRT being rebuilt to play an endless,
 
100
  480P / 4:3 is the period-correct sweet spot; the style carries fine to higher
101
  resolutions if you want clean pixels of a dirty tape.
102
 
103
+ ## ComfyUI
104
+
105
+ `vh5tape.safetensors` loads in ComfyUI as-is β€” no conversion needed. It uses the
106
+ ComfyUI-native key layout (`diffusion_model.blocks.N.attn.{qkv_proj,out_proj}`,
107
+ fused QKV, alpha = rank), so all 104 modules load with no warnings.
108
+
109
+ **Base checkpoint β€” this is the thing that matters.** Use a non-pruned,
110
+ non-rotated base:
111
+
112
+ - βœ… `minimax_h3_fl2va_bf16.safetensors` (recommended)
113
+ - βœ… `minimax_h3_fl2va_pruned_fp8_scaled.safetensors`
114
+ - ❌ `*_int8_convrot`, `nvfp4`, `w4a8` β€” these store weights in a rotated /
115
+ quantized basis. The LoRA loads without any error and then produces warped
116
+ faces, melting limbs and disappearing objects. **The default checkpoint in the
117
+ official ComfyUI H3 tutorial is `minimax_h3_fl2va_pruned_int8_convrot` β€” you
118
+ must change it.** (Not specific to this LoRA β€” MiniMax-H3 LoRAs in general are
119
+ made for fp16/fp8 bases.)
120
+
121
+ **Setup**
122
+
123
+ 1. Put `vh5tape.safetensors` in `ComfyUI/models/loras/`.
124
+ 2. `UNETLoader` (bf16 or fp8_scaled fl2va) β†’ `LoraLoaderModelOnly` β†’
125
+ `MiniMaxH3ImageToVideo` / your sampler.
126
+ 3. Strength **1.0** (useful range 0.7–1.2).
127
+
128
+ **Stacking with a turbo LoRA:** drop this one to **0.4–0.6** β€” the two
129
+ perturbations add, and 4-step turbo has little headroom. (The `turbo_mode`
130
+ checkbox occupies its own LoRA slot; add a second `LoraLoaderModelOnly` for
131
+ this one.) 4-step turbo also softens the tape grain β€” ~20–25 full-precision
132
+ steps show the style best. Stay near 1 MP (very low resolutions are unstable)
133
+ and keep prompts to a single shot.
134
+
135
+ `vh5tape-comfyui.safetensors` is an optional convenience build: bit-identical
136
+ weights plus explicit `.alpha` tensors (= rank, scale 1.0) and the base-model
137
+ requirement in its metadata β€” same look, just harder to mis-load in
138
+ third-party loaders.
139
+
140
  ## Training
141
 
142
  Trained with fal's MiniMax H3 t2v trainer (rank 32, 5,000 steps, 4:3, joint
 
151
  | file | what |
152
  |---|---|
153
  | [`vh5tape.safetensors`](./vh5tape.safetensors) | the LoRA, rank 32, 5,000 steps |
154
+ | [`vh5tape-comfyui.safetensors`](./vh5tape-comfyui.safetensors) | same weights bit-exact + explicit alpha tensors, for ComfyUI (see the ComfyUI section) |
155
  | [`PROMPTS.md`](./PROMPTS.md) | the exact prompts + seeds behind every sample above |
156
 
157
  Made for **H3-TV** β€” a 1970s portable CRT being rebuilt to play an endless,