# Tencent — Real life is where context gets hard

- Company: Tencent (tencent.com)
- Announced: 2026-04-30T07:00:00+00:00
- Category: research-paper
- Subject: Hunyuan
- Source: https://hunyuan.tencent.com/research/100039
- Record: https://forck.live/items/4579-real-life-is-where-context-gets-hard

Tencent Hunyuan introduces CL-bench Life, a benchmark for evaluating context learning ability in real-life settings, containing 405 context-task pairs and 5,348 human-written rubrics. Evaluations of 12 language models show they solve only 14.5% of tasks on average, with the best model (GPT-5.5 High) solving 22.2%.

## Evidence

Verbatim from https://hunyuan.tencent.com/research/100039:

> We introduce CL-bench Life, a rigorous benchmark for evaluating context learning ability in real-life settings and guiding future model development.

---

Record: https://forck.live/items/4579-real-life-is-where-context-gets-hard
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
