Skip to main content
Back to News
Zhipu releases GLM-5.3-Flash with SenseTime-backed domestic inference
Partnership
2 min read
CN

Zhipu releases GLM-5.3-Flash with SenseTime-backed domestic inference

The AMW Read

The open-source native multimodal release meaningfully advances Zhipu's known model strategy and explicitly ties open-weight distribution to model-scale and serving-cost economics.
NoveltySignificance
Foundation Models · Player MapScaling Laws
Zhipu AI
Zhipu AI

Foundation Models / LLMs

View Company Profile

Named counterparties: SenseTime

Zhipu releases GLM-5.3-Flash with SenseTime-backed domestic inference

Zhipu AI has released and open-sourced GLM-5.3-Flash, a 320 billion-parameter model it describes as the GLM-5 series’ first native multimodal release. The company said the model was pretrained on 30 trillion multimodal tokens and designed for low serving cost. Before launch, it was tested anonymously as Ox-Alpha on OpenCode and OpenRouter; the source reports 62 trillion tokens of calls were served on domestic chips. SenseTime provided the heterogeneous infrastructure and token-serving support for the launch.

The release puts deployment economics alongside model capability in the contest for developer adoption. The source says GLM-5.3-Flash achieved a 57 score on the Artificial Analysis Intelligence Index, matching Claude Opus 4.8, while the supported cluster improved end-to-end performance threefold versus its initial baseline. It also claims hardware efficiency and per-token cost comparable with mainstream Nvidia GPUs. Open-sourcing a native multimodal model makes those serving-cost claims strategically important: lower-cost inference can determine whether an open model becomes a practical production alternative rather than only a benchmark contender.

Builders evaluating GLM-5.3-Flash should test latency, output quality, tool compatibility, and effective token pricing on their own multimodal workloads before shifting production traffic. Investors should watch whether the reported infrastructure gains translate into repeatable third-party demand and margins, especially as Zhipu’s model distribution and domestic-chip deployment become more tightly linked.

#Zhipu #GLM53Flash #OpenSourceAI #MultimodalAI #Inference

#Zhipu AI#GLM-5.3-Flash#SenseTime#open-source models#related:SenseTime

How This Connects

Based on Foundation Models · Player Map

  1. 5h agoAnthropic launches Model Hardware Standard (MHS), a preview protocol connecting AI agents to lab and industrial hardware.Anthropic
  2. 11h agoAnthropic's Reported $45B Nscale Deal Secures Vera Rubin CapacityAnthropic
  3. 19h agoZhipu releases GLM-5.3-Flash with SenseTime-backed domestic inference · THIS ARTICLE
  4. 1w agoAlibaba's Qwen team has open-sourced Qwen3.8-27B, a 27-billion-parameter multimodal model designed f...Qwen
  5. 1w agoAlibaba has released Qwen 3.8 27B, an Apache 2.0-licensed open-weight dense model with 27 billion pa...Alibaba Qwen 3.8 27B launch
  6. 1mo agoMoonshot AI launches Kimi K3, a 2.8 trillion-parameter open-weight model, claiming performance near US frontier labsMoonshot AI

Related News

More news from Zhipu AI

Stay updated with the latest news and announcements from Zhipu AI.

View all Zhipu AI news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard