Blog · corpus.blog/blogs/mlcommons.org/posts
mlcommons.org
mlcommons.org
2026
MLCommons Joins EU-Funded AIRIS Project to Build and Benchmark Next-Generation Biomedical AIoriginal ↗
23 Sept 2026
16 Sept 2026
MLCommons Agent Reliability Profile Named a Finalist in Global Agentic Regulator Hackathonoriginal ↗
15 Sept 2026
18 Aug 2026
MedPerf Meets Google Cloud Confidential Computing: Secure AI Benchmarking for Brain Tumor Researchoriginal ↗
22 Jul 2026
MLCommons Releases MLPerf Mobile v6.0 with New Generative AI Benchmarks for On-Device LLMsoriginal ↗
15 Jun 2026
9 Jun 2026
Chakra Comes of Age: A Standardized Trace Ecosystem for AI Systems Benchmarking and Co-designoriginal ↗
2 Jun 2026
Fresh Benchmarks, Reliable Scores: Introducing Continuous Prompt Stewardship for AI Risk Evaluationoriginal ↗
20 Apr 2026
MLCommons Releases MLPerf Client v1.6 with Performance Optimizations and Enhanced User Experienceoriginal ↗
6 Apr 2026
24 Mar 2026
19 Mar 2026
Global Standards, Local Ground Truths: Piloting Multilingual, Multimodal AI Safety Understanding in APACoriginal ↗
13 Mar 2026
MedPerf enhances User Experience with improved Data Preparation Pipelines in Federated Clinical Studiesoriginal ↗
11 Mar 2026
A New Standard for AI Risk: How the AILuminate Global Assurance Program Is Reshaping Reliabilityoriginal ↗
20 Feb 2026
2025
9 Dec 2025
MLCommons Unveils New Jailbreak Benchmark, Quantifying AI’s “Resilience Gap” to Adversarial Attacksoriginal ↗
15 Oct 2025
New MLPerf Storage v2.0 Benchmark Results Demonstrate the Critical Role of Storage Performance in AI Training Systemsoriginal ↗
4 Aug 2025
MLCommons Releases MLPerf Client v1.0: A New Standard for AI PC and Client LLM Benchmarkingoriginal ↗
30 Jul 2025
MLCommons Builds New Agentic Reliability Evaluation Standard in Collaboration with Industry Leadersoriginal ↗
27 Jun 2025
Peter Mattson Joins Leaders to Discuss AI Innovation at Asia Tech x Conference in Singaporeoriginal ↗
6 Jun 2025
New MLCommons MLPerf Training v5.0 Benchmark Results Reflect Rapid Growth and Evolution of the Field of AIoriginal ↗
4 Jun 2025
3 Jun 2025
28 Apr 2025
MedPerf enhances transparency and integrity of real-world benchmarks using smart contractsoriginal ↗
10 Mar 2025
4 Mar 2025
MLCommons Releases AILuminate LLM v1.1, Adding French Language Capabilities to Industry-Leading AI Safety Benchmarkoriginal ↗
11 Feb 2025
MLCommons Medical Working Group Co-authors Book Chapter on Collaborative Evaluation for Medical Imagingoriginal ↗
27 Jan 2025
15 Jan 2025
2024
MLCommons Launches AILuminate, First-of-its-Kind Benchmark to Measure the Safety of Large Language Models original ↗
4 Dec 2024
New MLPerf Training v4.1 Benchmarks Highlight Industry’s Focus on New Systems and Generative AI Applicationsoriginal ↗
13 Nov 2024
New MLPerf Storage v1.0 Benchmark Results Show Storage Systems Play a Critical Role in AI Model Training Performanceoriginal ↗
25 Sept 2024
New MLPerf Inference v4.1 Benchmark Results Highlight Rapid Hardware and Software Innovations in Generative AI Systemsoriginal ↗
28 Aug 2024
Announcing the results of the inaugural AlgoPerf: Training Algorithms benchmark competitionoriginal ↗
1 Aug 2024
26 Jun 2024
20 Jun 2024