AIData brief

China's best open model leads the US by 46 Arena points, down from 72

Best open-weight text score by lab country (LMArena)Arena score
1,1001,2001,3001,4001,500Feb ’24Jul ’24Dec ’24May ’25Oct ’25Mar ’26Aug ’261,491China1,445US1,428Europe1,331OtherDeepSeek-R1 era: China jumpsMistral Large 3 lifts EuropeGemma 4 lifts the US
1,2001,400Feb ’24Jan ’25Dec ’251,491China1,445US1,428Europe1,331OtherDeepSeek-R1 era: Chin…Mistral Large 3 lifts…Gemma 4 lifts the US
Source: LMArena text leaderboard history (overall category), cleaned snapshot table; The Decoder, Reflection Beam report. Chart: Tec N Trend.
Best Chinese open model, 2 Oct 2026
1491 (mimo-v2.6-pro)
Best US open model
1445
China lead over US, Sep 2025 to Oct 2026
72 to 46 pts
Best European open model (Mistral Large 3)
1428

tecntrend.com/briefs/2026-10-07/open-beam

X Bluesky LinkedIn Reddit

The best US open model on the board is Nvidia's Nemotron 3 Ultra. Reflection's Beam, billed as the top non-Chinese open model, is not yet scored.

The Decoder reports that Reflection AI's Beam, a 501B-parameter model (23B active) due under Apache 2.0 later this month, is now the strongest open-weight model built outside China. Reflection says it matches GLM 5.2 on key benchmarks. Beam is not on LMArena yet.

LMArena's text ranking shows what it has to beat. On 2 October the top open-weight model from a Chinese lab was Xiaomi's mimo-v2.6-pro at 1491, ahead of the best US open model (nvidia-nemotron-3-ultra-550b-a55b-nvfp4, 1445) by 46 points and Europe's Mistral Large 3 (1428) by 63. A year earlier the China-US gap was 72.

The gap shrank because US open models jumped (Gemma 4 in March, Nemotron 3 Ultra in July), not because China stalled; its best rose 63 points. A Beam score above about 1445 would make it the top non-Chinese open model on the board.

Data, method and caveats
  • Reflection Beam is not in the Arena data (last snapshot 2026-10-02, before its release). The news supplies its size and licence only; the chart plots no Beam point.
  • "Open weight" uses the leaderboard's licence field, filled from each model's other snapshots where a snapshot had it blank; any licence other than Proprietary counts, including non-commercial ones.
  • Lab country is by company headquarters, set by us: China (Alibaba, DeepSeek, Moonshot, Xiaomi, Z.ai and others), US, Europe (Mistral, Stability), Other (Cohere, AI21, TII, Upstage and similar). A model named Trinity with no listed org is counted as US (Arcee).
  • Points are the last snapshot of each month from Jan 2024, plus 2 Oct 2026. Arena scores drift as new models and votes arrive, and are only comparable within a snapshot.
  • Scores are the default overall text view, not style-controlled or category-specific.

This brief as JSON or Markdown. Reuse the chart with credit to Tec N Trend and the sources above.

More data briefs