# Hugging Face — NPHardEval Leaderboard: Unveiling the Reasoning Abilities of Large Language Models through Complexity Classes and Dynamic Updates

- Company: Hugging Face (huggingface.co)
- Announced: 2024-02-02T00:00:00+00:00
- Category: research-paper
- Subject: Platform
- Source: https://huggingface.co/blog/leaderboard-nphardeval
- Record: https://forck.live/items/1945-nphardeval-leaderboard-unveiling-the-reasoning-abilities-of-large-language

Hugging Face introduces the NPHardEval leaderboard, a dynamic benchmark for evaluating LLM reasoning abilities using complexity classes, with 900 algorithmic questions updated monthly.

## Evidence

Verbatim from https://huggingface.co/blog/leaderboard-nphardeval:

> We're happy to introduce the NPHardEval leaderboard, using NPHardEval, a cutting-edge benchmark developed by researchers from the University of Michigan and Rutgers University.

---

Record: https://forck.live/items/1945-nphardeval-leaderboard-unveiling-the-reasoning-abilities-of-large-language
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
