🚨 New episode of Into AI Safety is live.
@seanmcgregor.bsky.social joins the show to discuss why current AI benchmarks often fail, how incentive structures shape safety work, and what it would take to build evaluations that actually support real-world decisions.
Episode links in thread!