mlx-FLUX.1-schnell-4bit-quantized
![FLUX.1 [schnell] Grid](./fluxonmac.png) Note: This checkpoint features 4-bit quantization of the mmdit module using MLX's nn.quantize function with default settings (groupsize=64). - ## Create conda environment shell conda create -n diffusionkit python=3.11 -y conda activate diffusionkit pip install diffusionkit - ## Run the cli command shell diffusionkit-cli --prompt "detailed cinematic dof render of a \ detailed...
Params
—
Context
—
Downloads 30d
28 K
Likes
37
Download history
daily snapshots · 22 days30 K20 K
Aug 31Sep 7Sep 14Sep 21
Can you run it?
Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.
| File | Quant | Size | Est. VRAM | Verdict on RTX 4090 · 24 GB |
|---|---|---|---|---|
| model.safetensors | fp16 | 7.0 GB | 8.2 GB | ✅ Runs comfortably |
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.
Specifications
- Licence
- apache-2.0
- First seen on the Hub
- 2024-08-20
- Training datasets
- undisclosed
- Added to our catalog
- 2026-08-31
Compare with any text-to-image model