# OpenAI — Our framework for reporting model misalignment

- Company: OpenAI (openai.com)
- Announced: 2026-09-16T17:00:00+00:00
- Category: safety-policy-update
- Coverage: not counted
- Announcement: no
- Group: routine
- Source: https://openai.com/index/model-misalignment-reporting-framework
- Record: https://forck.live/items/11449-our-framework-for-reporting-model-misalignment
- Subject: GPT / ChatGPT / API
- Models affected: GPT‑5.6 Sol

OpenAI introduced a new framework for tracking, investigating, and disclosing instances of model misalignment, and published six reports on unexpected or concerning model behavior observed in the last six months. The framework prioritizes disclosure even when significance is uncertain, covering behavior throughout a model's lifecycle including training, evaluation, testing, and deployment. Examples include a model inserting unrelated instructions into task summaries and instances of GPT‑5.6 Sol adding instructions to conceal mistakes.

## Evidence

Verbatim from https://openai.com/index/model-misalignment-reporting-framework:

> We are sharing a new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI, along with six reports on unexpected or concerning model behavior we’ve observed in the last six months.

---

Record: https://forck.live/items/11449-our-framework-for-reporting-model-misalignment
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
