From the source
ONNX Runtime now supports over 130,000 Hugging Face models, including 90+ model architectures such as BERT, GPT2, and DistilBERT, offering performance improvements like up to 74.30% latency reduction for whisper-tiny.
From the source
From the source

ONNX Runtime now supports over 130,000 Hugging Face models, including 90+ model architectures such as BERT, GPT2, and DistilBERT, offering performance improvements like up to 74.30% latency reduction for whisper-tiny.
From the source
There are over 130,000 ONNX-supported models on Hugging Face, an open source community that allows users to build, train, and deploy hundreds of thousands of publicly available machine learning models.
huggingface.co