ComfyUI GGUF: تشغيل نماذج الانتشار الكبيرة على وحدات معالجة رسومات منخفضة الـ VRAM
GGUF quantisation shrinks large diffusion models like FLUX.1 from ~24 GB to 5–12 GB, letting them run on consumer GPUs with 6–16 GB VRAM.Install the ComfyUI-GGUF custom node by city96, place .gguf files in ComfyUI/models/unet/, and use the UnetLoaderGGUF node instead of the standard UNETLoader.Q4_K_S or Q5_K_S offer the best quality-to-VRAM ratio for most cards.













