Qwen 3.5 35B A3B
qwen-3.5-35b-a3bqwen3-5-35b-a3bqwen3.5-35b-a3bqwen35-35b-a3b- 🧠 35B-parameter mixture-of-experts activating only ~3B per token
- 📏 Native 256K-token context window
- 👁️ Multimodal model handling both text and vision inputs
- 🔧 Built for reasoning, coding, agents, function calling, web search
- 🏢 From Alibaba's Qwen team, released under Apache 2.0
- 🆕 Provider reports it surpasses the larger Qwen3-235B-A22B
- ⚡ Sparse activation targets efficient, lower-cost inference
Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research, developing the Qwen…
Explore 39 more models by Alibaba Group →Qwen 3.5 35B A3B is a sparse mixture-of-experts model from Alibaba's Qwen team, released in February 2026 under the Apache 2.0 license. It carries roughly 35 billion total parameters but activates only about 3 billion per token, a design intended to deliver large-model behavior at a fraction of the compute cost. The model is multimodal, accepting text and vision inputs, and supports a native 256K-token context window alongside reasoning, code-optimized generation, function calling, and web search.
Within the Qwen family, this release sits between smaller and larger 3.5-generation siblings such as Qwen 3.5 9B and the much larger Qwen 3.5 397B. The provider describes the 35B A3B model as surpassing the earlier, denser Qwen 3 235B A22B Instruct 2507 while being roughly 6.7 times smaller in total parameters — a generational efficiency gain attributed to the company's own description.
In practice, the model targets reasoning, coding, and general-knowledge tasks where its small active-parameter footprint keeps latency and serving costs low. Later Qwen releases such as Qwen 3.6 27B continue this compact mixture-of-experts direction. Because primary benchmark documentation specific to this checkpoint is limited, the description here stays to verifiable architectural and licensing facts.
This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.
| Seller | Reputation↓ | Routing | Input $/M | Cached $/M | Output $/M | Categories | API |
|---|---|---|---|---|---|---|---|
| Venice.ai Proxy 0x1f22…18c9 | 100.00 | #2 | $0.1563 | $0.0781 | $0.625 | chat,reasoning,coding,vision,video,multimodal,web-search | openai-chat-completions |
| ▲ Apex Ant 0x73b4…e736 | 92.13 | #1 | $0.14 | $0.07 | $0.57 | chat,open-source,fast,cheap,vision,multimodal,reasoning,long-context,agents | openai-chat-completions |
| Fire Ant 🔥🐜 0xbe05…bc5d | 41.29 | gated | $0.2554 | $0.1277 | $1.0215 | — | — |
| Open Bird 0xc0f1…8183 | 18.32 | gated | $0.125 | $0.125 | $1.00 | chat,open-source | openai-chat-completions |
| Meridian AI 0x8c8c…06f5 | 5.64 | gated | $0.0304 | $0.0304 | $0.1215 | chat,coding,reasoning | openai-chat-completions |
| Apex TEE Test 0xe672…7955 | 0.51 | gated | $0.2331 | $0.2331 | $0.9408 | chat,open-source,fast,cheap,vision,multimodal,reasoning,long-context,agents | openai-chat-completions |
| antseed-electric-heron-4939 0xdc56…4939 | 0.08 | gated | $0.09 | $0.02 | $0.35 | chat,coding,reasoning | openai-chat-completions |
| Skeffo Inference 0x1af8…e2b5 | 0.02 | gated | $0.30 | $0.30 | $1.25 | chat,vision,function-calling | openai-chat-completions |
| antseed-neon-puma-944e 0x6650…944e | 0.00 | gated | $0.1066 | $0.1563 | $0.4263 | chat,coding,math | openai-chat-completions |
| Leftermute 0x388b…5389 | 0.00 | gated | $0.0284 | $0.0284 | $0.1136 | chat,coding,json | openai-chat-completions |
"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.