# Hugging Face — Benchmarking Language Model Performance on 5th Gen Xeon at GCP

- Company: Hugging Face (huggingface.co)
- Announced: 2024-12-17T00:00:00+00:00
- Subject: Platform
- Models affected: WhereIsAI/UAE-Large-V1, meta-llama/Llama-3.2-3
- Source: https://huggingface.co/blog/intel-gcp-c4
- Record: https://forck.live/items/1781-benchmarking-language-model-performance-on-5th-gen-xeon-at-gcp

The post benchmarks text embedding and text generation on Google Cloud C4 (5th Gen Xeon) vs N2 (3rd Gen Xeon), showing C4 has 10x-24x higher throughput for embedding and 2.3x-3.6x for generation, with TCO advantages.

## Evidence

Verbatim from https://huggingface.co/blog/intel-gcp-c4:

> The results consistently shows that C4 has 10x to 24x higher throughput over N2 in text embedding and 2.3x to 3.6x higher throughput over N2 in text generation.

---

Record: https://forck.live/items/1781-benchmarking-language-model-performance-on-5th-gen-xeon-at-gcp
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
