realrebelai commited on
Commit
b911227
Β·
verified Β·
1 Parent(s): 948242f

Upload README.md

Browse files
Files changed (1) hide show
  1. README.md +158 -0
README.md ADDED
@@ -0,0 +1,158 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: inclusionAI/LLaDA-Image-Turbo
3
+ base_model_relation: quantized
4
+ library_name: diffusers
5
+ pipeline_tag: text-to-image
6
+ tags:
7
+ - llada-image
8
+ - llada-image-turbo
9
+ - image-generation
10
+ - image-editing
11
+ - comfyui
12
+ - int8
13
+ - gguf
14
+ - quantized
15
+ - rebelai
16
+ ---
17
+
18
+ # LLaDA-Image-Turbo β€” ComfyUI / RebelAI Quantization
19
+
20
+ Optimized weights for **LLaDA-Image-Turbo** intended for the RebelAI ComfyUI integration.
21
+
22
+ This repository supports both **text-to-image generation** and **native LLaDA image editing**.
23
+
24
+ ## Repositories
25
+
26
+ - **ComfyUI nodes / GitHub:** https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
27
+ - **Model weights / this Hugging Face repository:** https://huggingface.co/realrebelai/LLaDa-Image-Turbo_ComfyUI
28
+ - **Official base model:** https://huggingface.co/inclusionAI/LLaDA-Image-Turbo
29
+ - **Official LLaDA-Image source:** https://github.com/inclusionAI/LLaDA-Image
30
+
31
+ ## Base Model / Quantization Tree
32
+
33
+ This repository is a quantized derivative of:
34
+
35
+ **`inclusionAI/LLaDA-Image-Turbo`**
36
+
37
+ The model-card metadata intentionally contains:
38
+
39
+ ```yaml
40
+ base_model: inclusionAI/LLaDA-Image-Turbo
41
+ base_model_relation: quantized
42
+ ```
43
+
44
+ This associates the repository with the main LLaDA-Image-Turbo model as a quantized derivative in the Hugging Face model relationship/quantization tree.
45
+
46
+ ## Available Weights
47
+
48
+ | File | Format | Description |
49
+ |---|---|---|
50
+ | `LLaDA-Image-Turbo-transformer-BF16.safetensors` | BF16 Safetensors | Full transformer weights |
51
+ | `LLaDA-Image-Turbo-transformer-INT8.safetensors` | INT8 Safetensors | Native INT8 transformer |
52
+ | `LLaDA-Image-Turbo-text_encoder-Q4_K_M-v3.gguf` | GGUF Q4_K_M | Quantized LLaDA2-MoE text encoder |
53
+
54
+ ### Native INT8 Transformer
55
+
56
+ The INT8 transformer is stored as native Safetensors. It is **not a transformer GGUF**.
57
+
58
+ The custom ComfyUI runtime loads the quantized transformer while preserving the LLaDA-Image transformer architecture.
59
+
60
+ ### Q4_K_M Text Encoder
61
+
62
+ The LLaDA2-MoE text encoder is provided as Q4_K_M GGUF for substantially lower storage/runtime memory requirements than the original full text encoder.
63
+
64
+ ## Text-to-Image
65
+
66
+ LLaDA-Image-Turbo is designed for fast generation.
67
+
68
+ Recommended starting settings in the ComfyUI integration:
69
+
70
+ - Steps: **4**
71
+ - CFG / guidance scale: **1.0**
72
+
73
+ The ComfyUI nodes are available here:
74
+
75
+ https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
76
+
77
+ ## Native Image Editing
78
+
79
+ The same LLaDA-Image-Turbo model also supports **native image editing**. No separate editing checkpoint is required.
80
+
81
+ The RebelAI ComfyUI node uses the model's native:
82
+
83
+ `generation_mode="editing"`
84
+
85
+ The source image is processed through LLaDA's image-conditioning/SigVQ path. This is **not** conventional img2img implemented with a denoise-strength slider.
86
+
87
+ Typical ComfyUI graph:
88
+
89
+ ```text
90
+ LLaDA Image Loader
91
+ |
92
+ v
93
+ LLaDA Image Edit <---- Load Image
94
+ |
95
+ v
96
+ Save Image
97
+ ```
98
+
99
+ Editing inputs include:
100
+
101
+ - Source image
102
+ - Edit instruction
103
+ - Width / height
104
+ - Steps
105
+ - Guidance scale
106
+ - Seed
107
+ - Optional negative prompt
108
+
109
+ Editing dimensions should be divisible by 32.
110
+
111
+ Example instruction:
112
+
113
+ > Turn the fox into a white arctic fox while preserving the forest composition and realistic photography.
114
+
115
+ ## ComfyUI Installation
116
+
117
+ Get the custom nodes here:
118
+
119
+ https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
120
+
121
+ Typical model placement:
122
+
123
+ ```text
124
+ ComfyUI/
125
+ └── models/
126
+ β”œβ”€β”€ diffusion_models/
127
+ β”‚ β”œβ”€β”€ LLaDA-Image-Turbo-transformer-INT8.safetensors
128
+ β”‚ └── LLaDA-Image-Turbo-transformer-BF16.safetensors
129
+ β”œβ”€β”€ text_encoders/
130
+ β”‚ └── LLaDA-Image-Turbo-text_encoder-Q4_K_M-v3.gguf
131
+ └── vae/
132
+ └── LLaDa_VAE.safetensors
133
+ ```
134
+
135
+ The ComfyUI integration also loads the supporting LLaDA pipeline components required by the official architecture.
136
+
137
+ ## Upstream
138
+
139
+ All model architecture and original model weights originate from inclusionAI's LLaDA-Image project.
140
+
141
+ Official LLaDA-Image-Turbo:
142
+
143
+ https://huggingface.co/inclusionAI/LLaDA-Image-Turbo
144
+
145
+ Official source:
146
+
147
+ https://github.com/inclusionAI/LLaDA-Image
148
+
149
+ Please refer to the upstream project for the original model documentation, research information, and applicable licensing terms.
150
+
151
+ ## Credits
152
+
153
+ - **inclusionAI** β€” LLaDA-Image / LLaDA-Image-Turbo
154
+ - **RealRebelAI** β€” ComfyUI integration, INT8 runtime/weights, and GGUF text-encoder integration
155
+
156
+ ## License
157
+
158
+ These files are derivatives of LLaDA-Image-Turbo. The upstream model's applicable license and usage terms continue to apply. The ComfyUI integration code is maintained separately in the GitHub repository linked above.