# Inception — Mercury 2 and the Rise of Real-time Subagents

- Company: Inception (inceptionlabs.ai)
- Announced: 2026-05-12
- Category: not stated
- Coverage: not counted
- Announcement: yes
- Group: announcements
- Source: https://www.inceptionlabs.ai/blog/rise-of-realtime-subagents
- Record: https://forck.live/items/18521-mercury-2-and-the-rise-of-real-time-subagents
- Subject: Mercury

How Inception's diffusion LLMs power fast, parallel agent workflows in production at Augment Code Inception Team The future of AI systems is not a single agent doing everything. It is multiple agents working together. That is already happening in coding. What looks like one coding agent from the outside is actually a system made up of specialized components. One part plans, another explores the codebase, another writes the implementation, and another compresses context so the session can continue. This is not just an implementation detail. It is becoming the core architecture of agentic systems. According to inference platforms like Baseten, customers run 7-10 models in production for targeted tasks, each powering a different part of the pipeline. Why multi-agent systems win There are structural, unavoidable tradeoffs between speed, quality, and cost for LLMs. The model you want for the hardest reasoning step is often not the model you want for every supporting operation. If you use the most expensive model everywhere, latency and cost become a problem. If you use the cheapest model everywhere, quality breaks at critical moments. …

---

Record: https://forck.live/items/18521-mercury-2-and-the-rise-of-real-time-subagents
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
