# Together AI — Improved Batch Inference API: Enhanced UI, Expanded Model Support, and 3000× Rate Limit Increase

- Company: Together AI (together.ai)
- Announced: 2025-09-15T00:00:00+00:00
- Category: capability-change
- Subject: Inference platform
- Pricing: 50% the cost of our real-time API
- Source: https://www.together.ai/blog/batch-inference-api-updates-2025
- Record: https://forck.live/items/2408-improved-batch-inference-api-enhanced-ui-expanded-model-support-and-3000-rate

Together AI announced major improvements to its Batch Inference API, including a streamlined UI, support for all serverless models and private deployments, a 3000x rate limit increase (from 10M to 30B tokens), and 50% cost reduction compared to the real-time API.

## Evidence

Verbatim from https://www.together.ai/blog/batch-inference-api-updates-2025:

> Rate limits are up from 10M to 30B enqueued tokens per model per user, a 3000× increase.

---

Record: https://forck.live/items/2408-improved-batch-inference-api-enhanced-ui-expanded-model-support-and-3000-rate
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
