- 🔒 Runs in a Trusted Execution Environment with hardware attestation.
- 🧠 Post-training upgrade to GLM-5 for agentic engineering work.
- 📏 200,000-token context window, FP8 precision.
- 🏢 Large Mixture-of-Experts design per NVIDIA's model card.
- ⚡ Z.ai documents up to 8 hours of autonomous task execution.
- 🔧 Reasoning, web search, and tool-calling agent capabilities.
- 📚 Open weights under the permissive MIT license.
- 🌐 From Z.ai (formerly Zhipu AI), released April 2026.
Z.ai, formally Knowledge Atlas Technology Joint Stock Co., Ltd., is a Chinese technology company specializing in artificial intelligence. Previously known internationally as Zhipu AI, the company rebranded to Z.ai in 2025. Its core focus is the GLM family of large language…
Explore 14 more models by Z.ai →GLM 5.1 is the privacy-hardened deployment of Z.ai's long-horizon flagship, served inside a Trusted Execution Environment where hardware attestation lets users independently verify enclave identity and configuration, so even the host cannot read prompts. It runs in FP8 precision under the MIT license, with a 200,000-token context window, and is released by Z.ai (formerly Zhipu AI) in April 2026.
It builds directly on Z.ai's GLM-5 foundation as a post-training upgrade rather than a new architecture. NVIDIA's model card describes GLM-5.1 as a large Mixture-of-Experts model and states it has significantly stronger coding capabilities along with improved results across coding, math, and agentic tasks versus its same-family predecessor [[sibling:zai-org-glm-5|GLM 5]].
Z.ai positions the model for sustained autonomous work. Its documentation describes planning, execution, testing, and fixing loops running up to eight hours, and reports a 3.6× geometric-mean speedup on KernelBench Level 3 optimization, compared with 1.49× for torch.compile in max-autotune mode. These figures are vendor-reported.
Within the broader lineage GLM 5.1 sits between [[sibling:zai-org-glm-4.7|GLM 4.7]] and the newer, longer-context [[sibling:zai-org-glm-5-2|GLM 5.2]], reflecting Z.ai's steady iteration on its open-weight GLM family.
This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.
| Seller | Reputation↓ | Routing | Input $/M | Cached $/M | Output $/M | Categories | API |
|---|---|---|---|---|---|---|---|
| Venice.ai Proxy 0x1f22…18c9 | 99 | #2 | $0.55 | $0.55 | $2.075 | chat,reasoning,web-search,e2ee | openai-chat-completions |
| surplusintelligence.ai 0x0e49…8927 | 68 | #1 | $0.33 | $0.33 | $1.245 | anon,chat,cheap,e2ee,frontier,privacy,reasoning,research,tasks,tee,translate,web-search | openai-chat-completions |
| Fire Ant 🔥🐜 0xbe05…bc5d | 37 | gated | $0.99 | $0.99 | $3.735 | base-usdc,chat,code,coding,general,github,glm,json,low-cost,math,model,monitored,openai-compatible,reasoning,response-auth,surplus,tasks,tee,tools,value,verified,zhipu | — |
| antseed-neon-puma-944e 0x6650…944e | 3 | gated | $0.3751 | $0.3751 | $1.4152 | chat,math | openai-chat-completions |
| antseed-opal-badger-2580 0xc85d…2580 | 2 | gated | $0.55 | $0.55 | $2.075 | chat | openai-chat-completions |
| Leftermute 0x388b…5389 | 0 | gated | $0.2222 | $0.2222 | $0.8383 | chat,coding,json,tools | openai-chat-completions |
"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.