Sign in

AlgoPerf

@algoperf.bsky.social
35 followers 15 following 14 posts

AlgoPerf benchmark for faster neural network training via better training algorithms

PostsRepliesMedia
AlgoPerf @algoperf.bsky.social · 12/02/2026
Learn more about AlgoPerf and this new release here: github.com/mlcommons/al...
010
AlgoPerf @algoperf.bsky.social · 12/02/2026
We just released AlgoPerf v1.0.1! 📚 New workload: AlgoPerf now includes a decoder-only 150M Language Model. ⚙️ Hardware switch: moving from 8xV100 to 4xA100. 🐞 Bug fixes and improved API. Coming soon: updated leaderboard with new optimizers! 👀
100
AlgoPerf @algoperf.bsky.social · 08/09/2025
We just released AlgoPerf v0.6! 🎉 ✅ Rolling leaderboard ✅ Lower compute costs ✅ JAX jit migration ✅ Bug fixes & flexible API Coming soon: More contemporary baselines + an LM workload… github.com/mlcommons/al...
github.com
GitHub - mlcommons/algorithmic-efficiency: MLCommons Algorithmic Efficiency is a benchmark and competition measuring neural network training speedups due to algorithmic improvements in both training a...
MLCommons Algorithmic Efficiency is a benchmark and competition measuring neural network training speedups due to algorithmic improvements in both training algorithms and models. - mlcommons/algori...
032
AlgoPerf @algoperf.bsky.social · 03/04/2025
The explainer video: www.youtube.com/watch?v=_yX1...
youtube.com
ICLR 2025: Accelerating Neural Network Training (AlgoPerf)
YouTube video by Tübingen Machine Learning
072
AlgoPerf @algoperf.bsky.social · 03/04/2025
We're all about acceleration! 😉 Watch @priya-kasimbeg.bsky.social & @fsschneider.bsky.social speedrun an explanation of the AlgoPerf benchmark, rules, and results all within a tight 5 minutes for our #ICLR2025 paper video on "Accelerating Neural Network Training". See you in Singapore!
154
AlgoPerf @algoperf.bsky.social · 14/03/2025
More details: 📄 ICLR 2025 results paper: openreview.net/pdf?id=CtM5x... 📄 Benchmark paper: arxiv.org/abs/2306.07179 💻 AlgoPerf codebase: github.com/mlcommons/al...
openreview.net
010
AlgoPerf @algoperf.bsky.social · 14/03/2025
This is just the beginning! With the help of the community, we want to test even more training recipes, push the SOTA for neural network training methods even further, and improve the AlgoPerf benchmark. Follow us & join the working group (mlcommons.org/working-grou...).
mlcommons.org
Algorithms - MLCommons
The Algorithms working group creates a set of rigorous and relevant benchmarks to measure neural network training speedups due to algorithmic improvements.
110
AlgoPerf @algoperf.bsky.social · 14/03/2025
And the winner in the self-tuning ruleset, based on Schedule Free AdamW, demonstrated a new level of effectiveness for completely hyperparameter-free neural network training. Roughly ~10% faster training, compared to a NadamW baseline with well-tuned default hyperparameters.
100
AlgoPerf @algoperf.bsky.social · 14/03/2025
Then, we asked the community to submit training algorithms. The results? The winner of the external tuning ruleset, using Distributed Shampoo, reduced training time by ~30% over our well-tuned baseline—showing that non-diagonal methods can beat Adam, even in wall-clock time!
110
AlgoPerf @algoperf.bsky.social · 14/03/2025
(3) Training algorithms must perform across 8 realistic deep learning workloads (ResNet-50, Conformer, ViT, etc.). (4) Submissions compete on the runtime to reach a given performance threshold. (5) Hyperparameter tuning is explicitly accounted for with our tuning rulesets.
100
AlgoPerf @algoperf.bsky.social · 14/03/2025
Over the past years, we've built the AlgoPerf: Training Algorithms benchmark. The core ideas: (1) Only the training algorithm changes; everything else (hardware, model, data) stays fixed. (2) Submissions compete directly; no more weak baselines!
100
AlgoPerf @algoperf.bsky.social · 14/03/2025
These choices are often critical, but reliable empirical guidance is scarce. Instead, we rely on expert intuition, anecdotal evidence, and babysitting. Check out this learning rate schedule from the OPT paper, which was manually determined. There has to be a better way!
100
AlgoPerf @algoperf.bsky.social · 14/03/2025
Currently, training neural nets is a complicated & fragile process with many important choices: How should I set/tune the learning rate? Using what schedule? Should I use SGD or Adam (or maybe Nadam/Amos/Shampoo/SOAP/Muon/... the list is virtually endless)?
100
AlgoPerf @algoperf.bsky.social · 14/03/2025
Hi there! This account will post about the AlgoPerf benchmark and leaderboard updates for faster neural network training via better training algorithms. But let's start with what AlgoPerf is, what we have done so far, and how you can train neural nets ~30% faster.
163