Nunchaku brings 4-bit diffusion inference into Diffusers
Hugging Face detailed integrating Nunchaku's 4-bit quantized inference into the Diffusers library, cutting the memory and compute needed to run diffusion models. For anyone serving image generation on constrained hardware, native 4-bit support in the standard library lowers the barrier considerably. A practical…
Source ↗