A month left in the AIMO #Interpretability Challenge @neuripsconf.bsky.social 🧠
📊 1k+ submissions predicting whether #LLMs reason robustly on Olympiad-level math problems.
⚠️ Remember: high eval acc ≠ high test score.
⏳ Still time until 1 Nov.
👉 aimo-interp.github.io