Qwen 3 Coder 480B Turbo

CodeWeb searchFunction calling
Advertised as qwen-3-coder-480b-turboqwen/qwen3-coder-480b-a35b-instruct-turboqwen3-coderqwen3-coder-480b-a35b-instruct-turboqwen3-coder-480b-turboqwen3-coder-turbo
Quick reference
Qwen 3 Coder 480B Turbo — TLDR
  • 🧠 Mixture-of-Experts coder: 480B total, 35B active parameters
  • ⚡ Turbo, FP8-quantized build tuned for faster code inference
  • 📏 256K native context, extendable toward 1M via extrapolation
  • 🔧 Agentic coding with a purpose-built function-call format
  • 💬 Instruct, non-thinking model — no reasoning-trace blocks
  • 🌐 Works with Qwen Code, CLINE, and similar agent tools
  • 🏢 Built by Alibaba's Qwen team (Alibaba Cloud)
  • 🎯 Capabilities here include function calling and web search
💰 Best price on AntSeed
$0.0050 / $0.015
per 1M · cheapest in / out
📏 Context
256K tokens
🐜 Sellers
13
advertising on AntSeed
Provider

Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research, developing the Qwen…

Explore 33 more models by Alibaba Group
About this model

Qwen 3 Coder 480B Turbo is a code-optimized large language model from Alibaba's Qwen team, served as a Turbo, FP8-quantized variant of the Qwen3-Coder-480B-A35B-Instruct base. The underlying model is a Mixture-of-Experts design with 480 billion total parameters and 35 billion active per inference, which the Qwen team frames as delivering high performance at lower compute cost than dense models of comparable scale. It supports a 256K-token context natively, with extrapolation methods reaching up to roughly 1M tokens.

The "Turbo" designation reflects an inference-optimized deployment: FP8 weights and provider-side serving aimed at faster, cheaper code workloads, which is the focus of this catalog entry. Functionally, it is an instruct, non-thinking model — it does not emit separate reasoning-trace blocks — and ships with a specially designed function-call format for agentic coding across tools like Qwen Code and CLINE.

Within Venice's broader Qwen lineup, it sits alongside general-purpose siblings such as [[sibling:qwen3-235b-a22b-instruct-2507|Qwen 3 235B A22B Instruct 2507]] and the efficiency-focused [[sibling:qwen3-next-80b|Qwen 3 Next 80b]], but this checkpoint is specialized purely for coding and agentic tool use. As deployed here it adds function-calling and web-search capabilities for developer workflows.

Sources
build.nvidia.comqwen3-coder-480b-a35b-instruct Model by Qwen· build.nvidia.comqwenlm.github.ioQwen3-Coder: Agentic Coding in the World | Qwen· qwenlm.github.iohuggingface.coQwen/Qwen3-Coder-480B-A35B-Instruct · Hugging Face· huggingface.co

This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.

Usage on AntSeed
Tokens served
2.77M
input + output
Requests
1,422
settled calls
Buyers
7
distinct, on this model
Sellers used
4
of 13 advertising
Settled
$0.47
gross USDC, this model
Sellers serving Qwen 3 Coder 480B Turbo (13)compare on the network explorer →
SellerReputationRoutingInput $/MCached $/MOutput $/MCategoriesAPI

"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.