AI SA Sam Austin AI What Is Reward Shaping and How Does It Improve RL Training? You know that moment when you’re training an RL agent and it just… sits there? Doing nothing? Or worse, it discovers some bizarre exploit…