# Together AI — ThunderAgent: 2x Faster Agentic Inference for Synthetic Data Generation at Scale

- Company: Together AI (together.ai)
- Announced: 2026-07-29T00:00:00+00:00
- Category: infrastructure-release
- Subject: Inference platform
- Source: https://www.together.ai/blog/thunderagent
- Record: https://forck.live/items/2329-thunderagent-2x-faster-agentic-inference-for-synthetic-data-generation-at-scale

Together AI introduces ThunderAgent, a system for high throughput agentic inference that achieves up to 2.5x higher single-node throughput and 2.4x speedup on 8 nodes by introducing a novel program abstraction for scheduling agentic workflows to mitigate KV cache thrashing.

## Evidence

Verbatim from https://www.together.ai/blog/thunderagent:

> By introducing a novel program abstraction for agentic LLM request scheduling, ThunderAgent achieves up to 2.5× higher single-node throughput in our synthetic data generation pipeline, and delivers 2.4× speedup on an 8-node cluster with near-linear throughput scaling with respect to GPU nodes.

---

Record: https://forck.live/items/2329-thunderagent-2x-faster-agentic-inference-for-synthetic-data-generation-at-scale
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
