# Hugging Face — The Open Medical-LLM Leaderboard: Benchmarking Large Language Models in Healthcare

- Company: Hugging Face (huggingface.co)
- Announced: 2024-04-19T00:00:00+00:00
- Category: research-paper
- Subject: Platform
- Models affected: GPT-3, GPT-4, Med-PaLM 2
- Source: https://huggingface.co/blog/leaderboard-medicalllm
- Record: https://forck.live/items/1903-the-open-medical-llm-leaderboard-benchmarking-large-language-models-in

Hugging Face announces the Open Medical-LLM Leaderboard, a standardized platform for evaluating and comparing large language models on medical tasks using datasets like MedQA, MedMCQA, PubMedQA, and MMLU subsets.

## Evidence

Verbatim from https://huggingface.co/blog/leaderboard-medicalllm:

> The Open Medical-LLM Leaderboard aims to address these challenges and limitations by providing a standardized platform for evaluating and comparing the performance of various large language models on a diverse range of medical tasks and datasets.

---

Record: https://forck.live/items/1903-the-open-medical-llm-leaderboard-benchmarking-large-language-models-in
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
