Text-to-Image
Diffusers
GGUF
llada-image
llada-image-turbo
image-generation
image-editing
comfyui
int8
quantized
rebelai
Instructions to use realrebelai/LLaDa-Image-Turbo_ComfyUI with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusers
How to use realrebelai/LLaDa-Image-Turbo_ComfyUI with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("realrebelai/LLaDa-Image-Turbo_ComfyUI", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- Draw Things
- DiffusionBee
Upload README.md
Browse files
README.md
ADDED
|
@@ -0,0 +1,158 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
---
|
| 2 |
+
base_model: inclusionAI/LLaDA-Image-Turbo
|
| 3 |
+
base_model_relation: quantized
|
| 4 |
+
library_name: diffusers
|
| 5 |
+
pipeline_tag: text-to-image
|
| 6 |
+
tags:
|
| 7 |
+
- llada-image
|
| 8 |
+
- llada-image-turbo
|
| 9 |
+
- image-generation
|
| 10 |
+
- image-editing
|
| 11 |
+
- comfyui
|
| 12 |
+
- int8
|
| 13 |
+
- gguf
|
| 14 |
+
- quantized
|
| 15 |
+
- rebelai
|
| 16 |
+
---
|
| 17 |
+
|
| 18 |
+
# LLaDA-Image-Turbo β ComfyUI / RebelAI Quantization
|
| 19 |
+
|
| 20 |
+
Optimized weights for **LLaDA-Image-Turbo** intended for the RebelAI ComfyUI integration.
|
| 21 |
+
|
| 22 |
+
This repository supports both **text-to-image generation** and **native LLaDA image editing**.
|
| 23 |
+
|
| 24 |
+
## Repositories
|
| 25 |
+
|
| 26 |
+
- **ComfyUI nodes / GitHub:** https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
|
| 27 |
+
- **Model weights / this Hugging Face repository:** https://huggingface.co/realrebelai/LLaDa-Image-Turbo_ComfyUI
|
| 28 |
+
- **Official base model:** https://huggingface.co/inclusionAI/LLaDA-Image-Turbo
|
| 29 |
+
- **Official LLaDA-Image source:** https://github.com/inclusionAI/LLaDA-Image
|
| 30 |
+
|
| 31 |
+
## Base Model / Quantization Tree
|
| 32 |
+
|
| 33 |
+
This repository is a quantized derivative of:
|
| 34 |
+
|
| 35 |
+
**`inclusionAI/LLaDA-Image-Turbo`**
|
| 36 |
+
|
| 37 |
+
The model-card metadata intentionally contains:
|
| 38 |
+
|
| 39 |
+
```yaml
|
| 40 |
+
base_model: inclusionAI/LLaDA-Image-Turbo
|
| 41 |
+
base_model_relation: quantized
|
| 42 |
+
```
|
| 43 |
+
|
| 44 |
+
This associates the repository with the main LLaDA-Image-Turbo model as a quantized derivative in the Hugging Face model relationship/quantization tree.
|
| 45 |
+
|
| 46 |
+
## Available Weights
|
| 47 |
+
|
| 48 |
+
| File | Format | Description |
|
| 49 |
+
|---|---|---|
|
| 50 |
+
| `LLaDA-Image-Turbo-transformer-BF16.safetensors` | BF16 Safetensors | Full transformer weights |
|
| 51 |
+
| `LLaDA-Image-Turbo-transformer-INT8.safetensors` | INT8 Safetensors | Native INT8 transformer |
|
| 52 |
+
| `LLaDA-Image-Turbo-text_encoder-Q4_K_M-v3.gguf` | GGUF Q4_K_M | Quantized LLaDA2-MoE text encoder |
|
| 53 |
+
|
| 54 |
+
### Native INT8 Transformer
|
| 55 |
+
|
| 56 |
+
The INT8 transformer is stored as native Safetensors. It is **not a transformer GGUF**.
|
| 57 |
+
|
| 58 |
+
The custom ComfyUI runtime loads the quantized transformer while preserving the LLaDA-Image transformer architecture.
|
| 59 |
+
|
| 60 |
+
### Q4_K_M Text Encoder
|
| 61 |
+
|
| 62 |
+
The LLaDA2-MoE text encoder is provided as Q4_K_M GGUF for substantially lower storage/runtime memory requirements than the original full text encoder.
|
| 63 |
+
|
| 64 |
+
## Text-to-Image
|
| 65 |
+
|
| 66 |
+
LLaDA-Image-Turbo is designed for fast generation.
|
| 67 |
+
|
| 68 |
+
Recommended starting settings in the ComfyUI integration:
|
| 69 |
+
|
| 70 |
+
- Steps: **4**
|
| 71 |
+
- CFG / guidance scale: **1.0**
|
| 72 |
+
|
| 73 |
+
The ComfyUI nodes are available here:
|
| 74 |
+
|
| 75 |
+
https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
|
| 76 |
+
|
| 77 |
+
## Native Image Editing
|
| 78 |
+
|
| 79 |
+
The same LLaDA-Image-Turbo model also supports **native image editing**. No separate editing checkpoint is required.
|
| 80 |
+
|
| 81 |
+
The RebelAI ComfyUI node uses the model's native:
|
| 82 |
+
|
| 83 |
+
`generation_mode="editing"`
|
| 84 |
+
|
| 85 |
+
The source image is processed through LLaDA's image-conditioning/SigVQ path. This is **not** conventional img2img implemented with a denoise-strength slider.
|
| 86 |
+
|
| 87 |
+
Typical ComfyUI graph:
|
| 88 |
+
|
| 89 |
+
```text
|
| 90 |
+
LLaDA Image Loader
|
| 91 |
+
|
|
| 92 |
+
v
|
| 93 |
+
LLaDA Image Edit <---- Load Image
|
| 94 |
+
|
|
| 95 |
+
v
|
| 96 |
+
Save Image
|
| 97 |
+
```
|
| 98 |
+
|
| 99 |
+
Editing inputs include:
|
| 100 |
+
|
| 101 |
+
- Source image
|
| 102 |
+
- Edit instruction
|
| 103 |
+
- Width / height
|
| 104 |
+
- Steps
|
| 105 |
+
- Guidance scale
|
| 106 |
+
- Seed
|
| 107 |
+
- Optional negative prompt
|
| 108 |
+
|
| 109 |
+
Editing dimensions should be divisible by 32.
|
| 110 |
+
|
| 111 |
+
Example instruction:
|
| 112 |
+
|
| 113 |
+
> Turn the fox into a white arctic fox while preserving the forest composition and realistic photography.
|
| 114 |
+
|
| 115 |
+
## ComfyUI Installation
|
| 116 |
+
|
| 117 |
+
Get the custom nodes here:
|
| 118 |
+
|
| 119 |
+
https://github.com/RealRebelAI/LLaDa-Image_ComfyUI
|
| 120 |
+
|
| 121 |
+
Typical model placement:
|
| 122 |
+
|
| 123 |
+
```text
|
| 124 |
+
ComfyUI/
|
| 125 |
+
βββ models/
|
| 126 |
+
βββ diffusion_models/
|
| 127 |
+
β βββ LLaDA-Image-Turbo-transformer-INT8.safetensors
|
| 128 |
+
β βββ LLaDA-Image-Turbo-transformer-BF16.safetensors
|
| 129 |
+
βββ text_encoders/
|
| 130 |
+
β βββ LLaDA-Image-Turbo-text_encoder-Q4_K_M-v3.gguf
|
| 131 |
+
βββ vae/
|
| 132 |
+
βββ LLaDa_VAE.safetensors
|
| 133 |
+
```
|
| 134 |
+
|
| 135 |
+
The ComfyUI integration also loads the supporting LLaDA pipeline components required by the official architecture.
|
| 136 |
+
|
| 137 |
+
## Upstream
|
| 138 |
+
|
| 139 |
+
All model architecture and original model weights originate from inclusionAI's LLaDA-Image project.
|
| 140 |
+
|
| 141 |
+
Official LLaDA-Image-Turbo:
|
| 142 |
+
|
| 143 |
+
https://huggingface.co/inclusionAI/LLaDA-Image-Turbo
|
| 144 |
+
|
| 145 |
+
Official source:
|
| 146 |
+
|
| 147 |
+
https://github.com/inclusionAI/LLaDA-Image
|
| 148 |
+
|
| 149 |
+
Please refer to the upstream project for the original model documentation, research information, and applicable licensing terms.
|
| 150 |
+
|
| 151 |
+
## Credits
|
| 152 |
+
|
| 153 |
+
- **inclusionAI** β LLaDA-Image / LLaDA-Image-Turbo
|
| 154 |
+
- **RealRebelAI** β ComfyUI integration, INT8 runtime/weights, and GGUF text-encoder integration
|
| 155 |
+
|
| 156 |
+
## License
|
| 157 |
+
|
| 158 |
+
These files are derivatives of LLaDA-Image-Turbo. The upstream model's applicable license and usage terms continue to apply. The ComfyUI integration code is maintained separately in the GitHub repository linked above.
|