# OpenAI — Improving Model Safety Behavior with Rule-Based Rewards

- Company: OpenAI (openai.com)
- Announced: 2024-07-24T09:00:00+00:00
- Category: research-paper
- Subject: GPT / ChatGPT / API
- Source: https://openai.com/index/improving-model-safety-behavior-with-rule-based-rewards
- Record: https://forck.live/items/731-improving-model-safety-behavior-with-rule-based-rewards

OpenAI developed a new method called Rule-Based Rewards (RBRs) to align models for safe behavior without needing extensive human data collection.

## Evidence

Verbatim from https://openai.com/index/improving-model-safety-behavior-with-rule-based-rewards:

> We’ve developed and applied a new method leveraging Rule-Based Rewards (RBRs) that aligns models to behave safely without extensive human data collection.

---

Record: https://forck.live/items/731-improving-model-safety-behavior-with-rule-based-rewards
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
