Zhipu
Topic archive • 1 matches
AI-generated: summaries written by AI from the linked sources. How we use AI
2026-09-18
Technology
Zhipu AI launched GLM-5.3-FlashX, claiming nearly 200 tokens per second inference on ~100,000 domestic accelerators. The base Flash model is a 320B MoE with 18B active parameters, optimized by an Infra Agent serving stack.
Model Release • Pandaily
Permalink