No rulebook, no instructions, just trial and error: that's how reinforcement learning works, and it's exactly what Year 5 to 9 students in Harrow explored this Saturday. The same concept behind self-driving cars and game-playing AI, tested first on a whiteboard grid, then shaped in Python by setting the rewards and penalties that taught their AI to learn from its own mistakes.