# Hugging Face — StackLLaMA: A hands-on guide to train LLaMA with RLHF

- Company: Hugging Face (huggingface.co)
- Announced: 2023-04-05
- Category: new-model
- Coverage: not counted
- Announcement: no
- Group: routine
- Source: https://huggingface.co/blog/stackllama
- Record: https://forck.live/items/2078-stackllama-a-hands-on-guide-to-train-llama-with-rlhf
- Subject: Platform
- Models affected: StackLLaMA

Hugging Face releases the StackLLaMA model, a version of LLaMA fine-tuned with RLHF on Stack Exchange data, and provides a detailed guide on the training process including supervised fine-tuning, reward modeling, and reinforcement learning.

## Evidence

Verbatim from https://huggingface.co/blog/stackllama:

> we are releasing the StackLLaMA model.

---

Record: https://forck.live/items/2078-stackllama-a-hands-on-guide-to-train-llama-with-rlhf
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
