Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face and Intel announce that SetFit inference can be accelerated by 7.8x on Intel CPUs using Optimum Intel, an open-source library that includes post-training static quantization to convert models to INT8 precision, enabling production-grade deployment on Intel Xeon CPUs.
From the source
we'll explain how you can accelerate inference with SetFit by 7.8x on Intel CPUs, by optimizing your SetFit model with 🤗 Optimum Intel
huggingface.co