# Together AI — Aurora

- Company: Together AI (together.ai)
- Announced: 2026-03-31T00:00:00+00:00
- Category: developer-tool-release
- Subject: Inference platform
- Models affected: Qwen3, Llama3, MiniMax M2.5, Qwen3-Coder-Next-FP8
- Source: https://www.together.ai/blog/aurora
- Record: https://forck.live/items/2365-aurora

is an open-source, RL-based framework for adaptive speculative decoding that learns from live inference traces and continuously updates the speculator without interrupting serving, achieving an additional 1.25x speedup over a well-trained static speculator.

## Evidence

Verbatim from https://www.together.ai/blog/aurora:

> Today, we're releasing Aurora, an open-source, RL-based framework that learns from live inference traces and updates the speculator asynchronously—turning speculative decoding from a static, one-time setup into a dynamic, self-improving flywheel.

---

Record: https://forck.live/items/2365-aurora
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
