# Together AI — Open, convenient and predictable: Introducing Provisioned Throughput

- Company: Together AI (together.ai)
- Announced: 2026-07-08T00:00:00+00:00
- Category: product-launch
- Subject: Inference platform
- Models affected: MiniMax M3, GLM-5.2
- Pricing: $0.05 per PTU per minute
- Source: https://www.together.ai/blog/provisioned-throughput
- Record: https://forck.live/items/2337-open-convenient-and-predictable-introducing-provisioned-throughput

Together AI introduces Provisioned Throughput, a reserved inference capacity for open models with token-based pricing and a 99% uptime SLA, available for MiniMax M3 and GLM-5.2 with a one-month minimum term and discounts at higher commitment levels. Costs are up to 90% below Claude Opus 4.8 at list price. Provisioned Throughput Units (PTUs) are priced at $0.05 per PTU per minute, with different burn rates for input, cached input, and output tokens. The service is available in North America, EMEA, and beyond.

## Evidence

Verbatim from https://www.together.ai/blog/provisioned-throughput:

> We're excited to introduce Provisioned Throughput, reserved inference capacity for frontier open models with token-based pricing and a 99% uptime SLA.

---

Record: https://forck.live/items/2337-open-convenient-and-predictable-introducing-provisioned-throughput
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
