Welcome Mixtral - a SOTA Mixture of Experts on Hugging Face
Mistral released Mixtral 8x7b, a Mixture-of-Experts open-access model that outperforms GPT-3.5 and Llama 2 70B on many benchmarks, with Apache 2.0 license and 32k context length. Hugging Face announced comprehensive integration including models on the Hub, Transformers, Inference Endpoints, and Text Generation Inference.
