Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face and Intel announce CPU-optimized embeddings using Optimum Intel and fastRAG, focusing on BGE models (small, base, large) to accelerate semantic search and RAG pipelines on Xeon CPUs.
From the source
In this blog, we will show how to unlock significant performance boost on Xeon based CPUs, and show how easy it is to integrate optimized models into existing RAG pipelines using fastRAG.
huggingface.co