GLM-5.3 Flash ("Ox Alpha") official: Benchmarks attached. This...

GLM-5.3-Flash might be one of the most impressive efficiency releases yet.
It is a 320B MoE with only 18B parameters active per token, yet Zai reports:
- 84.3 on Terminal-Bench 2.1, nearly matching Claude Opus 4.8 at 85.0
- 63.4 on DeepSWE, ahead of Opus 4.8 and DeepSeek V4 Vision Exp
- 48.8 on AutomationBench, ahead of Opus 4.8 and GPT-5.6 Terra
- The highest GDPval-AA v2 score in its comparisonIt also beats the much larger GLM-5.2 across all six reported benchmarks while costing one-tenth as much to serve.
Open weights, MIT licensed, natively multimodal, 1M context.
Important caveat: 18B active parameters does not make it a normal local 18B model. All 320B weights still need to be stored. But in terms of intelligence per active parameter, this looks exceptional!

Outperforming GLM-5.2 at 1/10th of its price and approaching Opus 4.8 on coding and agentic benchmarks.
Big things incoming!

