🧵 1/
@MLPerf Inference v6.1 is here.
Record participation, new benchmarks, and substantial performance gains.
30 organizations submitted — the most ever.
Details: mlcommons.org/2026/09/mlperf…
MLPerf Training v6.1 adds the suite's first LLM post-training benchmark: agentic RL that teaches a 397B-parameter open-weight model to repair real software, scored on pass@4 quality - not just throughput.
Details from the task force:
mlcommons.org/2026/09/mlperf…
MLCommons joins the EU-funded AIRIS consortium to benchmark next-generation biomedical AI.
We're building an evaluation framework to measure mechanism-informed generative AI across accuracy, robustness, fairness, interpretability, and usability.
More: mlcommons.org/2026/09/join-a…
MLPerf Inference v6.1 WG chairs Miro Hodak and Frank Han walk through the record round, new silicon (AMD MI350P, Intel Arc Pro B70, NVIDIA Vera Rubin), the first cross-vendor heterogeneous submission, and what it signals for inference engineering.
mlcommons.org/2026/09/chairs…
AI Infra 2026: MLPerf Inference v6.1 - Technical deep dive.
David Kanter, Miro Hodak (AMD), and Frank Han (Dell) on serving scenarios, reasoning workloads, edge measurements, and what's driving the biggest gains this round. Plus: a preview of MLPerf Endpoints.
#AIInfraSummit