Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Amazon announces custom reward functions for multi-turn reinforcement learning with Amazon Nova Forge, including a serverless multi-turn RL option now generally available, and describes how to design composite rewards for GRPO.
From the source
Nova Forge also offers a serverless multi-turn RL option, now generally available, for teams that prefer not to manage that environment.
aws.amazon.com