MLPerf Training v6.1 adds the suite's first LLM post-training benchmark: agentic RL that teaches a 397B-parameter open-weight model to repair real software, scored on pass@4 quality - not just throughput.
Details from the task force:
mlcommons.org/2026/09/mlperf-traini…