Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face published a blog post detailing how to run Mistral 7B with Core ML on a Mac, using new features from WWDC 24 such as Swift Tensor and Stateful Buffers, and achieving less than 4GB memory usage.
From the source
By the end of this blog post, you will have learnt all the new goodies accompanying the latest macOS release AND you will have successfully run a 7B parameter model using less than 4GB of memory on your Mac.
huggingface.co