realrebelai commited on
Commit
1628bfb
Β·
verified Β·
1 Parent(s): f23c2e3

Upload README.md

Browse files
Files changed (1) hide show
  1. README.md +221 -0
README.md ADDED
@@ -0,0 +1,221 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: inclusionAI/LLaDA-Image
3
+ base_model_relation: quantized
4
+ library_name: diffusers
5
+ pipeline_tag: text-to-image
6
+ tags:
7
+ - llada-image
8
+ - image-generation
9
+ - image-editing
10
+ - comfyui
11
+ - int8
12
+ - gguf
13
+ - quantized
14
+ - rebelai
15
+ ---
16
+
17
+ # LLaDA-Image Base β€” ComfyUI INT8 + GGUF
18
+
19
+ Quantized **LLaDA-Image Base** weights for the **RealRebelAI LLaDA-Image ComfyUI** custom nodes.
20
+
21
+ Based on **inclusionAI/LLaDA-Image**, this release is intended for lower-memory ComfyUI inference while retaining the Base model workflow.
22
+
23
+ ## ComfyUI Custom Nodes
24
+
25
+ https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
26
+
27
+ Install into:
28
+
29
+ ```text
30
+ ComfyUI/custom_nodes/ComfyUI-LLaDA-Image/
31
+ ```
32
+
33
+ Restart ComfyUI after installation.
34
+
35
+ ## Upstream Model
36
+
37
+ Official Base model:
38
+
39
+ https://huggingface.co/inclusionAI/LLaDA-Image
40
+
41
+ Official project:
42
+
43
+ https://github.com/inclusionAI/LLaDA-Image
44
+
45
+ This is an unofficial quantized derivative and is not affiliated with inclusionAI.
46
+
47
+ ## Included Weights
48
+
49
+ ### Base INT8 Transformer
50
+
51
+ ```text
52
+ LLaDA-Image-Base-transformer-INT8.safetensors
53
+ ```
54
+
55
+ Place in:
56
+
57
+ ```text
58
+ ComfyUI/models/diffusion_models/
59
+ ```
60
+
61
+ This is a native INT8 **Safetensors transformer**, not a transformer GGUF.
62
+
63
+ ### Base Q4_K_M Text Encoder
64
+
65
+ ```text
66
+ LLaDA-Image-Base-text_encoder-Q4_K_M.gguf
67
+ ```
68
+
69
+ Place in:
70
+
71
+ ```text
72
+ ComfyUI/models/text_encoders/
73
+ ```
74
+
75
+ The custom runtime uses City96 ComfyUI-GGUF support for the quantized LLaDA2 MoE text encoder.
76
+
77
+ ## Requirement
78
+
79
+ Install City96 ComfyUI-GGUF:
80
+
81
+ https://github.com/city96/ComfyUI-GGUF
82
+
83
+ ## VAE
84
+
85
+ Place the compatible LLaDA VAE in:
86
+
87
+ ```text
88
+ ComfyUI/models/vae/
89
+ ```
90
+
91
+ Then select it in **LLaDA Image Loader**.
92
+
93
+ The current custom nodes expose native VAE tiled decoding:
94
+
95
+ ```text
96
+ vae_tiling:
97
+ On
98
+ Auto
99
+ Off
100
+ ```
101
+
102
+ For low-VRAM GPUs, **On** is a good starting point.
103
+
104
+ ## Base Recommended Settings
105
+
106
+ ```text
107
+ Steps: 50
108
+ CFG: 5.0
109
+ ```
110
+
111
+ LLaDA-Image Base is the full model rather than the distilled few-step Turbo variant.
112
+
113
+ ## Text-to-Image
114
+
115
+ ```text
116
+ LLaDA Image Loader
117
+ |
118
+ v
119
+ LLaDA Image Text to Image
120
+ |
121
+ v
122
+ Save Image
123
+ ```
124
+
125
+ Loader example:
126
+
127
+ ```text
128
+ diffusion_model: LLaDA-Image-Base-transformer-INT8.safetensors
129
+ text_encoder: LLaDA-Image-Base-text_encoder-Q4_K_M.gguf
130
+ vae: your LLaDA VAE
131
+ dtype: bfloat16
132
+ vae_tiling: On
133
+ ```
134
+
135
+ Start with **50 steps / CFG 5.0**.
136
+
137
+ ## Native Image Editing
138
+
139
+ The custom nodes support LLaDA-Image's **native image-editing mode**.
140
+
141
+ This is not conventional img2img or denoise-strength emulation. The source image is passed through LLaDA-Image's native:
142
+
143
+ ```text
144
+ generation_mode="editing"
145
+ ```
146
+
147
+ using its image-conditioning/SigVQ path.
148
+
149
+ ```text
150
+ LLaDA Image Loader -----------+
151
+ |
152
+ Load Image -------------------+--> LLaDA Image Edit --> Save Image
153
+ ```
154
+
155
+ Example:
156
+
157
+ ```text
158
+ Turn the fox into a white arctic fox while preserving the forest composition and realistic photography.
159
+ ```
160
+
161
+ For Base editing, start with:
162
+
163
+ ```text
164
+ Steps: 50
165
+ CFG: 5.0
166
+ ```
167
+
168
+ Editing width and height must be divisible by **32**.
169
+
170
+ The same Base transformer and text encoder are used for generation and editing. No separate editing checkpoint is required.
171
+
172
+ ## Base vs Turbo
173
+
174
+ ```text
175
+ Base:
176
+ Steps: ~50
177
+ CFG: ~5.0
178
+
179
+ Turbo:
180
+ Steps: ~4
181
+ CFG: ~1.0
182
+ ```
183
+
184
+ ## Low-VRAM Notes
185
+
186
+ - Use the INT8 transformer.
187
+ - Use the Q4_K_M GGUF text encoder.
188
+ - Enable VAE tiling.
189
+ - Use CPU offload when necessary.
190
+ - Do not load the original full text encoder alongside the quantized encoder.
191
+
192
+ ## Model Placement
193
+
194
+ ```text
195
+ ComfyUI/
196
+ └── models/
197
+ β”œβ”€β”€ diffusion_models/
198
+ β”‚ └── LLaDA-Image-Base-transformer-INT8.safetensors
199
+ β”œβ”€β”€ text_encoders/
200
+ β”‚ └── LLaDA-Image-Base-text_encoder-Q4_K_M.gguf
201
+ └── vae/
202
+ └── <LLaDA VAE>.safetensors
203
+ ```
204
+
205
+ ## Important
206
+
207
+ Do **not** load the INT8 transformer through a GGUF diffusion loader. It is native INT8 Safetensors.
208
+
209
+ The `.gguf` file in this release is the **text encoder**, not the diffusion transformer.
210
+
211
+ ## Credits
212
+
213
+ - **inclusionAI** β€” LLaDA-Image architecture and official model
214
+ - **Hugging Face Diffusers** β€” pipeline/component infrastructure
215
+ - **City96** β€” ComfyUI-GGUF / GGML support
216
+ - **ComfyUI** β€” node and inference ecosystem
217
+ - **RealRebelAI** β€” custom ComfyUI integration and quantized release
218
+
219
+ ## License
220
+
221
+ The upstream LLaDA-Image model and components remain subject to their original licenses and terms. Review the upstream license before redistribution or commercial use.