🚀 Training a Large Reasoning Model, but high-quality data is scarce?
Check out our paper: “Leveraging Online Olympiad-Level Math Problems for LLM Training & Contamination-Resistant Evaluation.” 📖
TL;DR:
🔹 LLM-powered high-quality data collection 🤖
🔹 647K Math QA pairs 📊
🔹 A Live Math Benchmark ⏳
arxiv.org
Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation
Advances in Large Language Models (LLMs) have sparked interest in their ability to solve Olympiad-level math problems. However, the training and evaluation of these models are constrained by the limit...