# Amazon — Scaling MoE reinforcement learning on Amazon EKS with EFA and DeepEP with 40% more throughput

- Company: Amazon (amazon.com)
- Announced: 2026-09-25T16:29:50+00:00
- Category: not stated
- Coverage: not counted
- Announcement: yes
- Group: announcements
- Source: https://aws.amazon.com/blogs/machine-learning/scaling-moe-reinforcement-learning-on-amazon-eks-with-efa-and-deepep-with-40-more-throughput/
- Record: https://forck.live/items/13891-scaling-moe-reinforcement-learning-on-amazon-eks-with-efa-and-deepep-with-40
- Subject: Bedrock / Nova

When you post-train a Mixture-of-Experts (MoE) model with Reinforcement Learning from Human Feedback (RLHF) or Group Relative Policy Optimization (GRPO) at scale, three simultaneous challenges emerge. The first requires coordinating heterogeneous compute for rollout generation and policy training. Second, sustaining high-throughput communication across hundreds of accelerators. And third, dynamically orchestrating every subsystem to keep them in balance. On AWS, you can address these challenges using Amazon Elastic Kubernetes Service (Amazon EKS), Elastic Fabric Adapter (EFA), and DeepEP. Mixture-of-Experts (MoE) has become a standard architecture for scaling large language models (LLMs) to hundreds of billions or even trillions of parameters, while maintaining efficient inference through sparsity. However, sparsity doesn’t remove infrastructure complexity in training. As part of the standard training pipeline, these models must undergo pre-training, mid-training, supervised fine-tuning (SFT), and reinforcement learning (RL). …

---

Record: https://forck.live/items/13891-scaling-moe-reinforcement-learning-on-amazon-eks-with-efa-and-deepep-with-40
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
