415.tech
AI & tech, from the frontlines of Silicon Valley
Z.ai's GLM-5.3-Flash scores 57 on Artificial Analysis at $0.15 per 1M input tokens

Z.ai's GLM-5.3-Flash scores 57 on Artificial Analysis at $0.15 per 1M input tokens

Z.ai's open-weights GLM-5.3-Flash, released August 26, 2026, scores 57 on the Artificial Analysis Intelligence Index against a median of 27 for comparable open-weight models, at $0.15 per 1M input and $0.50 per 1M output tokens under an MIT license. The MoE design — 320B total parameters, 18B active, 1M-token context — puts frontier-adjacent reasoning and image input inside a self-hostable model that any team can deploy commercially without a vendor contract. Output runs 50.2 tokens per second, below average for the class, and the benchmark page is live so the figures will move.

Source: artificialanalysis.ai

Post on XEmail

GLM-5.3-Flash scores 57 on the Artificial Analysis Intelligence Index, placing it well above average among other open weight models of similar size (median: 27).

Artificial Analysis

Why this matters

  • → Open-weights model scores frontier-adjacent reasoning at $0.15/1M input tokens—self-hostable without vendor lo
  • → MoE design (320B total, 18B active) makes frontier capabilities deployable on constrained hardware
  • → MIT license + 1M context window enable commercial use cases previously locked behind proprietary APIs
Open weights break through
Also in this edition