Find the full piece here: www.antonleicht.me/writing/evals (12/12)
antonleicht.me
Problems in the AI Eval Political Economy — Anton Leicht
Evaluations of new AI models’ capabilities and risks are an important cornerstone of safety-focused AI policy. Currently, their future as part of the policy platform faces peril for four reasons: Entanglements with a broad AI safety ecosystem, structural incentives favouring less helpful evals, susc