Zhipu AI
What Zhipu AI is up to now
New 08-31 PM: the afternoon re-check holds the GLM-5.5 negative and adds a finding about the expectation's own framing. The zai-org Hugging Face org carries GLM-5.3 and its Flash variants with checkpoint uploads hours old, and no GLM-5.5 model card exists anywhere checked; Chinese coverage says the numbering stayed on the 5.x track and that "5.5" was a rumoured designation Zhipu never used. This does NOT mean GLM-5.3 satisfies the expectation: the claim is a >1T-parameter, 1M-context model and GLM-5.3 is 753B. So the claimed model does not exist under any name, and the label it was logged under appears abandoned by the vendor. New 08-31: GLM-5.5 still has not shipped, and today's negative is stronger than yesterday's. Neither Hugging Face nor Z.AI's own documentation carries it — while the company is demonstrably publishing this morning, shipping GLM-5.3-Flash updates. So this is a live vendor channel shipping something else, not an unreachable one; yesterday's check rested on secondary sources after a DNS failure on z.ai. The `glm-5-5-release` expectation is due today and is held pending until the window closes. New 08-15→08-16: still quiet — both days' only Zhipu/GLM mentions are explicit cross-references back to the 08-14 GLM-5.3 launch already captured below (08-16's TLDR AI Monday-lead item is confirmed weekday-publication lag on the same 08-14 story, not a new event). No movement yet on the ~08-28 Hugging Face weight release or the pending, later GLM-5.5 (due 08-31). Standing picture below holds. New 08-14 (the only substantive development found across the full ~3-week gap since 07-24): Zhipu, operating as Z.ai, launched its flagship GLM-5.3, claiming it edges out Anthropic's restricted-access Claude Mythos 5 on the CyberGym cybersecurity benchmark (84.5% vs. 83.8%) while trailing badly on ExploitBench (54.4% vs. 78%). Runs on the same 743B-parameter MoE base as GLM-5.2 (~40B active parameters/token). Notably, Zhipu is holding the weights back from Hugging Face until roughly 08-28 for its own safety review, with the most sensitive cybersecurity functions gated behind a "trusted access" program for verified users — the first time a Chinese open-weight lab has voluntarily gated a release this way, mirroring the access controls Anthropic itself uses for Mythos. Distinct from the still-pending, later GLM-5.5 release (also due 08-31) — not a flip of that entry. Standing picture unchanged: the 1GW all-domestic-chip training site (Z.AI) and the GLM line remain the clearest proof China can train at frontier scale without Nvidia; GLM-5.2 was the model Hugging Face used to investigate OpenAI's containment breach. New 08-26: confirmed it built "Ox Alpha," the stealth reasoning/coding model released uncredited over the previous weekend that had climbed past DeepSeek to top OpenRouter usage, and said it would open the weights the same night. Anonymity was the launch tactic: the model earned its leaderboard position with no national label attached, and the label arrived only once the position was unarguable.
as of 2026-08-31
The metrics
- Posture med
- expanding 3 src →
- Capital · available med
- ~$7-8B cumulative capital raised across all rounds (2023 series ~$350M -> May-2024 $400M @ ~$3B valuation -> 2025 state-linked rounds ~$1.4-1.9B -> Jan-2026 HK IPO $558M -> Jul-2026 $4B HK … 7 src →
- Capital · operating low
- revenue figures reported but inconsistent across sources — 2024 ARR reported ~$168M; a post-IPO earnings report cited revenue growth of 132% yoy while total revenue stayed 'under $105M' … 2 src →
- Capital · deployed low
- 1-gigawatt AI datacenter built on all-domestic chips (multiple 10,000-chip clusters, zero Nvidia silicon) completed/announced Jul-2026, alongside a self-developed inference-chip effort — the … 1 src →
- Capital · in med
- VC + Big-Tech strategic (Alibaba/Tencent/Meituan/Ant/Xiaomi/HongShan 2023; Prosperity7 2024) -> Chinese state/government guidance funds (Shanghai, 2025) -> public capital markets (HK IPO … 2 src →
- Capital · out med
- 1GW domestic-chip datacenter buildout + self-developed inference chips; compute shortage reported Feb-2026 (user signup restrictions) suggests spend has not kept pace with demand 1 src →
- Optionality med
- constrained — Chinese state/municipal guidance-fund investors (Shanghai, ~2025) sit alongside Big-Tech strategics (Alibaba/Tencent/Meituan/Ant); added to the US Commerce Dept Entity List … 3 src →
- Gravity low
- no clean $ figure — sub-$200M/yr disclosed revenue (2024 ARR ~$168M, growing fast per 2026 earnings) is the only hard number; broader touch includes open-weight GLM model adoption … 3 src →
Holdings
This week — what changed
- GLM-5.5 has not shipped, and today's negative is stronger than yesterday's. Neither Hugging Face nor Z.AI's own documentation carries it. Z.AI is demonstrably publishing this morning — it is shipping GLM-5.3-Flash updates — so this is a live vendor channel shipping something else, not an unreachable one. Yesterday's check rested on secondary sources after a DNS failure on z.ai.
- Zhipu (Z.ai) posted its first earnings as a public company, missing the growth Wall Street had priced in even as revenue quadrupled. First-half 2026 revenue: 953.9 million yuan (~$142 million), up ~400% year-on-year, against a Bloomberg consensus modeling 514% full-year growth; net loss narrowed 12.1% to 2.07 billion yuan. Cloud-deployment revenue rose over 2,700% YoY and open-platform/API revenue rose 27-fold, now 86.5% of the total. Shares rose ~9.6% on the print but remain ~60% below their June peak; annual recurring revenue hit $1.6bn by end of August, ahead of rival MiniMax's $800M ARR. Published ~7:43am ET 08-31, before this digest's cutoff, and missed until this catch. (SCMP)
Projects & threads
Full item history for Zhipu AI →
