Amazon Nova Forge introduces custom reward functions for multi-turn reinforcement learning (RFT) using Group Relative Policy Optimization (GRPO), allowing developers to define reward logic through Bring Your Own Orchestration (BYOO) or a serverless option.
From the source
Amazon Nova offers multiple customization approaches, with reinforcement fine-tuning (RFT) standing out because it can teach models the behaviors you want through iterative feedback.
aws.amazon.com