- 🧠 Next-generation flagship built for agentic engineering and long-horizon tasks.
- 🔧 Designed to keep using tools across extended autonomous runs.
- 🏢 Built by Z.ai (formerly Zhipu AI), released under MIT license.
- 📏 Roughly 200K-token context window for extended documents and repos.
- 🆕 Incremental refresh of GLM 5 with stronger coding and reasoning.
- ⚡ Ships as an FP8 checkpoint for efficient serving.
- 💬 Supports reasoning traces, function calling, and web search.
- 🌐 Open weights published on Hugging Face under MIT.
Z.ai, formally Knowledge Atlas Technology Joint Stock Co., Ltd., is a Chinese technology company specializing in artificial intelligence. Previously known internationally as Zhipu AI, the company rebranded to Z.ai in 2025. Its core focus is the GLM family of large language…
Explore 14 more models by Z.ai →GLM 5.1 is Z.ai's next-generation flagship large language model, positioned for agentic engineering, complex software tasks, and long-horizon planning with extended tool use. It is released under the MIT license, continuing the open-weight strategy the company (formerly Zhipu AI) has applied across its GLM family. The catalog lists a roughly 200K-token context window, and the model ships in an FP8 quantization for efficient deployment.
Z.ai describes GLM 5.1 as an evolution of [[sibling:zai-org-glm-5|GLM 5]], built to keep revising and improving across long autonomous runs rather than plateauing early. In its release notes, the company states that, relative to GLM 5, version 5.1 delivers gains in coding, agentic tool use, reasoning, and long-horizon agentic tasks. It is further tuned for agentic coding workflows, with adjustments to its chat template for deferred tool loading.
On Z.ai's own reported benchmarks, GLM 5.1 scores 58.4 on SWE-Bench Pro, and the company says it improves over its predecessor GLM 5 across major coding, mathematical reasoning, and agentic evaluations. Capabilities exposed through the catalog include reasoning, function calling, and web search.
In the family timeline, GLM 5.1 follows GLM 5 and earlier releases such as [[sibling:zai-org-glm-4.7|GLM 4.7]], and precedes the later [[sibling:zai-org-glm-5-2|GLM 5.2]], which Z.ai positions as a further step in long-horizon capability.
This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.
| Seller | Reputation↓ | Routing | Input $/M | Cached $/M | Output $/M | Categories | API |
|---|---|---|---|---|---|---|---|
| Venice.ai Proxy 0x1f22…18c9 | 99 | #4 | $0.875 | $0.1625 | $2.75 | chat,reasoning,web-search | openai-chat-completions |
| ChainScout AI 0x1734…e621 | 88 | gated | $2.00 | $2.00 | $6.00 | chat,coding,reasoning,model,general,zhipu,glm | openai-chat-completions |
| ▲ Apex Ant 0x73b4…e736 | 83 | #3 | $0.495 | $0.097 | $1.8675 | chat,open-source,coding,reasoning,long-context,agents | openai-chat-completions |
| Open Forge 0x1d90…b0aa | 75 | #5 | $1.05 | $0.21 | $3.50 | chat,coding,reasoning | openai-chat-completions |
| surplusintelligence.ai 0x0e49…8927 | 68 | #2 | $0.462 | $0.0858 | $1.452 | agents,anon,chat,cheap,frontier,function-calling,reasoning,research,tasks,tools,translate,web-search | openai-chat-completions |
| Open Bird 0xc0f1…8183 | 64 | #1 | $0.385 | $0.0715 | $1.21 | chat | openai-chat-completions |
| NovaRoute AI 0xc50d…ed7b | 60 | gated | $0.3812 | $0.3812 | $1.1979 | chat,coding,code,reasoning,tasks,glm,value,surplus,openai-compatible,low-cost,verified,github,response-auth,base-usdc,monitored | openai-chat-completions |
| Chutes 0xded6…657c | 43 | gated | $1.078 | $0.1078 | $3.388 | chat,reasoning,coding,tee | openai-chat-completions |
| Super Seeder 0xd19f…41f3 | 40 | gated | $0.77 | $0.143 | $2.42 | chat,reasoning,tools | openai-chat-completions |
| Fire Ant 🔥🐜 0xbe05…bc5d | 37 | gated | $1.369 | $0.2574 | $4.356 | base-usdc,chat,code,coding,general,github,glm,json,low-cost,math,model,monitored,openai-compatible,reasoning,response-auth,surplus,tasks,tee,tools,value,verified,zhipu | — |
| Meridian AI 0x8c8c…06f5 | 22 | gated | $0.1497 | $0.1497 | $0.4704 | chat,reasoning | openai-chat-completions |
| uomi.ai 0x87df…48e3 | 19 | gated | $1.379 | $1.379 | $4.374 | chat,math,coding | openai-chat-completions |
| antseed-neon-puma-944e 0x6650…944e | 3 | gated | $0.5968 | $0.325 | $1.8755 | chat,math | openai-chat-completions |
| minion0x 0x215e…e2e3 | 1 | gated | $1.03 | $0.20 | $3.48 | chat,coding,math,finance | openai-chat-completions |
| antseed-opal-badger-2580 0xc85d…2580 | 1 | gated | $0.77 | $0.77 | $2.42 | chat | openai-chat-completions |
| Skeffo Inference 0x1af8…e2b5 | 0 | gated | $1.50 | $1.50 | $5.00 | chat,reasoning,agent,function-calling | openai-chat-completions |
| Leftermute 0x388b…5389 | 0 | gated | $0.2121 | $0.2121 | $0.6666 | chat,coding,json,tools | openai-chat-completions |
| Apex TEE Test 0xe672…7955 | 0 | gated | $0.1944 | $0.189 | $0.648 | chat,open-source,coding,reasoning,long-context,agents | openai-chat-completions |
"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.