Skip to content
AIインテリジェンスSep 18, 2026AIインテリジェンス
記事

Zhipu AI launched GLM-5.3-FlashX, claiming nearly 200 tokens per second inference on ~100,000 domestic accelerators.

The base Flash model is a 320B MoE with 18B active parameters, optimized by an Infra Agent serving stack.

AI生成:要約はリンク先の情報源をもとにAIが作成しています。 AIの利用について

AI生成出典: Pandaily
01

ソース要約

Zhipu AI launched GLM-5.3-FlashX, claiming nearly 200 tokens per second inference on ~100,000 domestic accelerators. The base Flash model is a 320B MoE with 18B active parameters, optimized by an Infra Agent serving stack.