From the source
A workflow for optimizing Stable Diffusion models for Intel CPUs using OpenVINO NNCF and Optimum, achieving 5.1x inference acceleration and 4x model footprint reduction compared to PyTorch.
The method involves quantization-aware training, knowledge distillation, and token merging.





