Superhuman/X ArchiveView on X
Z.ai

@Zai_org

Introducing GLM-5.3-Flash

- Leading capabilities at a highly competitive price
- Natively multimodal with a 1M-token context window
- A 320B-A18B model released under the MIT License
- Previously previewed as Ox Alpha, running entirely on Chinese AI chips

Blog: z.ai/blog/glm-5.3-flash

Available now across all official platforms:

Weights: huggingface.co/zai-org/GLM-5.3-Flash
API: docs.z.ai/guides/llm/glm-5.3-flash
Coding Plan: z.ai/subscribe
ZCode: zcode.z.ai/en
Chat: chat.z.ai
AutoClaw: autoclaw.z.ai
Image from the post
7872K18.4K3.1K
Z.ai

@Zai_org

Standard API Pricing for GLM-5.3-Flash (per 1M tokens)

- Input: $0.15
- Output: $0.50
- Cached input: $0.03
60651.9K136
Z.ai

@Zai_org

On the Z.ai Code Bench, which measures real-world coding performance, GLM-5.3-Flash clearly outperforms GLM-5.2 at every effort level and performs on par with Claude Opus 4.8.
Image from the post
122081154
Z.ai

@Zai_org

Architectural enhancements, combined with an optimized pre-training corpus, enable GLM-5.3-Flash to deliver greater intelligence with less compute.
Image from the post
425795102
End of thread