New episode of Into AI Safety is live!
Sean McGregor joins bsky.app/profile/jaco... to discuss why current AI benchmarks often fail, how incentive structures shape safety work, and what it would take to build evaluations that actually support real-world decisions.
kairos.fm/intoaisafety...
kairos.fm
Sobering up on AI Progress w/ Dr. Sean McGregor | Kairos.fm
Why AI benchmarks fail, how safety gets measured wrong, and what real evaluation should look like with Dr. Sean McGregor