1. X
  2. MLCommons
Log inSign up
MLCommons
839 posts
user avatar
MLCommons
@MLCommons
Better Artificial Intelligence for Everyone
mlcommons.org
Joined September 2020
149
Following
3,653
Followers
RepliesRepliesMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    user avatar
    MLCommons
    @MLCommons
    Jul 28
    MLPerf Endpoints v0.7 is live - a foundation release for AI inference benchmarking. Initial results from Coreweave, Google, Intel, KRAI, and NVIDIA. Four principles: Current, Comprehensive, Comparable, Commentary. Blog: mlcommons.org/2026/07/mlperf…
    6.6K
  • user avatar
    MLCommons
    @MLCommons
    Jul 22
    Benchmark a brain tumor AI model on real patient MRI data. No data leaves the hospital. No model weights exposed. That's MedPerf + @googlecloud Confidential Space, demonstrated live at #GoogleCloudNext 2026. mlcommons.org/2026/06/medper…
    263
  • user avatar
    MLCommons
    @MLCommons
    Jul 14
    There's an AI Reliability Map. Most of it is still empty. Benchmarking clusters in a few cells. Enterprise AI readiness needs the full grid. AIRR is mapping what others skip: mlcommons.org/2026/04/airr-m…
    284
  • user avatar
    MLCommons
    @MLCommons
    Jul 9
    MLCommons is introducing an Edge Agentic Inference benchmark for MLPerf Inference v6.1. Single accelerator. One user. Multi-turn tool-calling. Hard 32K context wall. Model: Qwen3.6-27B Q4_K_M Accuracy gate: BFCL v4 Deadline: July 31, 2026 Read more: mlcommons.org/2026/07/mlperf…
    479
  • user avatar
    MLCommons
    @MLCommons
    Jul 8
    MLPerf Inference now measures multi-turn agents. 990 trajectories, Kimi K2.6 + Qwen3.6-35B-A3B, Pareto-curve performance, three-level accuracy. Built on MLPerf Endpoints. mlcommons.org/2026/07/agenti… MLPerf #AgenticAI #LLM #Inference #Kimi #Qwen #MLCommons
    242K