File size: 4,436 Bytes
1628bfb
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
66
67
68
69
70
71
72
73
74
75
76
77
78
79
80
81
82
83
84
85
86
87
88
89
90
91
92
93
94
95
96
97
98
99
100
101
102
103
104
105
106
107
108
109
110
111
112
113
114
115
116
117
118
119
120
121
122
123
124
125
126
127
128
129
130
131
132
133
134
135
136
137
138
139
140
141
142
143
144
145
146
147
148
149
150
151
152
153
154
155
156
157
158
159
160
161
162
163
164
165
166
167
168
169
170
171
172
173
174
175
176
177
178
179
180
181
182
183
184
185
186
187
188
189
190
191
192
193
194
195
196
197
198
199
200
201
202
203
204
205
206
207
208
209
210
211
212
213
214
215
216
217
218
219
220
221
222
---
base_model: inclusionAI/LLaDA-Image
base_model_relation: quantized
library_name: diffusers
pipeline_tag: text-to-image
tags:
- llada-image
- image-generation
- image-editing
- comfyui
- int8
- gguf
- quantized
- rebelai
---

# LLaDA-Image Base β€” ComfyUI INT8 + GGUF

Quantized **LLaDA-Image Base** weights for the **RealRebelAI LLaDA-Image ComfyUI** custom nodes.

Based on **inclusionAI/LLaDA-Image**, this release is intended for lower-memory ComfyUI inference while retaining the Base model workflow.

## ComfyUI Custom Nodes

https://github.com/RealRebelAI/LLaDa-Image_ComfyUI

Install into:

```text
ComfyUI/custom_nodes/ComfyUI-LLaDA-Image/
```

Restart ComfyUI after installation.

## Upstream Model

Official Base model:

https://huggingface.co/inclusionAI/LLaDA-Image

Official project:

https://github.com/inclusionAI/LLaDA-Image

This is an unofficial quantized derivative and is not affiliated with inclusionAI.

## Included Weights

### Base INT8 Transformer

```text
LLaDA-Image-Base-transformer-INT8.safetensors
```

Place in:

```text
ComfyUI/models/diffusion_models/
```

This is a native INT8 **Safetensors transformer**, not a transformer GGUF.

### Base Q4_K_M Text Encoder

```text
LLaDA-Image-Base-text_encoder-Q4_K_M.gguf
```

Place in:

```text
ComfyUI/models/text_encoders/
```

The custom runtime uses City96 ComfyUI-GGUF support for the quantized LLaDA2 MoE text encoder.

## Requirement

Install City96 ComfyUI-GGUF:

https://github.com/city96/ComfyUI-GGUF

## VAE

Place the compatible LLaDA VAE in:

```text
ComfyUI/models/vae/
```

Then select it in **LLaDA Image Loader**.

The current custom nodes expose native VAE tiled decoding:

```text
vae_tiling:
On
Auto
Off
```

For low-VRAM GPUs, **On** is a good starting point.

## Base Recommended Settings

```text
Steps: 50
CFG:   5.0
```

LLaDA-Image Base is the full model rather than the distilled few-step Turbo variant.

## Text-to-Image

```text
LLaDA Image Loader
        |
        v
LLaDA Image Text to Image
        |
        v
    Save Image
```

Loader example:

```text
diffusion_model: LLaDA-Image-Base-transformer-INT8.safetensors
text_encoder:    LLaDA-Image-Base-text_encoder-Q4_K_M.gguf
vae:             your LLaDA VAE
dtype:           bfloat16
vae_tiling:      On
```

Start with **50 steps / CFG 5.0**.

## Native Image Editing

The custom nodes support LLaDA-Image's **native image-editing mode**.

This is not conventional img2img or denoise-strength emulation. The source image is passed through LLaDA-Image's native:

```text
generation_mode="editing"
```

using its image-conditioning/SigVQ path.

```text
LLaDA Image Loader -----------+
                              |
Load Image -------------------+--> LLaDA Image Edit --> Save Image
```

Example:

```text
Turn the fox into a white arctic fox while preserving the forest composition and realistic photography.
```

For Base editing, start with:

```text
Steps: 50
CFG:   5.0
```

Editing width and height must be divisible by **32**.

The same Base transformer and text encoder are used for generation and editing. No separate editing checkpoint is required.

## Base vs Turbo

```text
Base:
Steps: ~50
CFG:   ~5.0

Turbo:
Steps: ~4
CFG:   ~1.0
```

## Low-VRAM Notes

- Use the INT8 transformer.
- Use the Q4_K_M GGUF text encoder.
- Enable VAE tiling.
- Use CPU offload when necessary.
- Do not load the original full text encoder alongside the quantized encoder.

## Model Placement

```text
ComfyUI/
└── models/
    β”œβ”€β”€ diffusion_models/
    β”‚   └── LLaDA-Image-Base-transformer-INT8.safetensors
    β”œβ”€β”€ text_encoders/
    β”‚   └── LLaDA-Image-Base-text_encoder-Q4_K_M.gguf
    └── vae/
        └── <LLaDA VAE>.safetensors
```

## Important

Do **not** load the INT8 transformer through a GGUF diffusion loader. It is native INT8 Safetensors.

The `.gguf` file in this release is the **text encoder**, not the diffusion transformer.

## Credits

- **inclusionAI** β€” LLaDA-Image architecture and official model
- **Hugging Face Diffusers** β€” pipeline/component infrastructure
- **City96** β€” ComfyUI-GGUF / GGML support
- **ComfyUI** β€” node and inference ecosystem
- **RealRebelAI** β€” custom ComfyUI integration and quantized release

## License

The upstream LLaDA-Image model and components remain subject to their original licenses and terms. Review the upstream license before redistribution or commercial use.