Introducing GLM-5.3-Flash - Leading capabilities at a highly...

@Zai_org
Z.ai@Zai_org
17 views Aug 26, 2026 ~1 min read
Advertisement
1
Introducing GLM-5.3-Flash

- Leading capabilities at a highly competitive price
- Natively multimodal with a 1M-token context window
- A 320B-A18B model released under the MIT License
- Previously previewed as Ox Alpha, running entirely on Chinese AI chips

Blog: z.ai/blog/glm-5.3-f…

Available now across all official platforms:

Weights: huggingface.co/zai-org/GLM-5.…
API: docs.z.ai/guides/llm/glm…
Coding Plan: z.ai/subscribe
ZCode: zcode.z.ai/en
Chat: chat.z.ai
AutoClaw: autoclaw.z.ai
Media image
2
Standard API Pricing for GLM-5.3-Flash (per 1M tokens)

- Input: $0.15
- Output: $0.50
- Cached input: $0.03
3
On the Z.ai Code Bench, which measures real-world coding performance, GLM-5.3-Flash clearly outperforms GLM-5.2 at every effort level and performs on par with Claude Opus 4.8.
Media image
4
Architectural enhancements, combined with an optimized pre-training corpus, enable GLM-5.3-Flash to deliver greater intelligence with less compute.
Media image
Actions
What You Can Do
  • Export as PDF or Markdown
  • Batch Export to Notion
  • Bookmark & Highlight
  • LinkedIn & Instagram Carousel Maker
Create Free Account

Includes 7-day Premium trial

Advertisement