Agentic Racing: A Pilot That Never Learned to Drive, a Team Boss That Did Learn to Reason
I built a racing demo where a reactive pilot obeys a team-boss LLM that reasons per event. Eight attempts at training the pilot with reinforcement learning didn't work — the track was broken before the algorithm ever had a chance — and measuring how much the strategist actually contributes turned into the more interesting experiment: seven runs, real and partially fixable reasoning biases, and a lesson about trusting the full benchmark over the quick desk-check.