Z.aiZ.ai·text

E2EE GLM 5.1

E2eeReasoningWeb search
Advertised as e2ee-glm-5-1e2ee-glm-5.1
Quick reference
GLM 5.1 — TLDR
  • 🔒 Runs in a Trusted Execution Environment with hardware attestation.
  • 🧠 Post-training upgrade to GLM-5 for agentic engineering work.
  • 📏 200,000-token context window, FP8 precision.
  • 🏢 Large Mixture-of-Experts design per NVIDIA's model card.
  • ⚡ Z.ai documents up to 8 hours of autonomous task execution.
  • 🔧 Reasoning, web search, and tool-calling agent capabilities.
  • 📚 Open weights under the permissive MIT license.
  • 🌐 From Z.ai (formerly Zhipu AI), released April 2026.
💰 Best price on AntSeed
$0.222 / $0.83877%
per 1M · cheapest in / out
📏 Context
200K tokens
🐜 Sellers
6
advertising on AntSeed
Provider

Z.ai, formally Knowledge Atlas Technology Joint Stock Co., Ltd., is a Chinese technology company specializing in artificial intelligence. Previously known internationally as Zhipu AI, the company rebranded to Z.ai in 2025. Its core focus is the GLM family of large language…

Explore 14 more models by Z.ai
About this model

GLM 5.1 is the privacy-hardened deployment of Z.ai's long-horizon flagship, served inside a Trusted Execution Environment where hardware attestation lets users independently verify enclave identity and configuration, so even the host cannot read prompts. It runs in FP8 precision under the MIT license, with a 200,000-token context window, and is released by Z.ai (formerly Zhipu AI) in April 2026.

It builds directly on Z.ai's GLM-5 foundation as a post-training upgrade rather than a new architecture. NVIDIA's model card describes GLM-5.1 as a large Mixture-of-Experts model and states it has significantly stronger coding capabilities along with improved results across coding, math, and agentic tasks versus its same-family predecessor [[sibling:zai-org-glm-5|GLM 5]].

Z.ai positions the model for sustained autonomous work. Its documentation describes planning, execution, testing, and fixing loops running up to eight hours, and reports a 3.6× geometric-mean speedup on KernelBench Level 3 optimization, compared with 1.49× for torch.compile in max-autotune mode. These figures are vendor-reported.

Within the broader lineage GLM 5.1 sits between [[sibling:zai-org-glm-4.7|GLM 4.7]] and the newer, longer-context [[sibling:zai-org-glm-5-2|GLM 5.2]], reflecting Z.ai's steady iteration on its open-weight GLM family.

View source on GitHub ↗View model card on HuggingFace ↗
Sources
docs.z.aiGLM-5.1 - Overview - Z.AI DEVELOPER DOCUMENT· docs.z.aidocs.api.nvidia.comz-ai / glm5.1· docs.api.nvidia.com

This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.

Usage on AntSeed
Tokens served
4.54k
input + output
Requests
3
settled calls
Buyers
2
distinct, on this model
Sellers used
1
of 6 advertising
Settled
$0.00
gross USDC, this model
Sellers serving E2EE GLM 5.1 (6)compare on the network explorer →
SellerReputationRoutingInput $/MCached $/MOutput $/MCategoriesAPI

"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.