Amazon Nova Forge's Reward Stack Exposes RL's Silent Killer
A new AWS blog post on Nova Forge reveals that composite, instrumented reward functions are the difference between RL training that converges and training that silently collapses. Here's what changed, who's affected, and the operational playbook.