Skip to content
Alibaba logo

Qwen3 32B

qwen/qwen3-32bPOST /v1/chat/completions

Dense 32B model. A solid mid-tier open-weights default.

Tool callingStructured outputStreaming
Input / 1M
$0.0800
via DeepInfra
Output / 1M
$0.2800
via DeepInfra
Cached input / 1M
not published
Context
131K
max out 33K
Providers
2
failover available

2 providers serve this model

Multigrid picks between them on every request using your policy. Each price is what that route costs us, and we add no spread of our own to it.

ProviderUpstream model idIn / 1MOut / 1MBilling
DeepInfraCheapestQwen/Qwen3-32B$0.0800$0.2800Credit
OpenRouterqwen/qwen3-32b$0.0800$0.2800Credit
No performance data published yet
Time to first token, throughput and uptime are measurements, and Multigrid has not served enough traffic through Qwen3 32B to publish honest ones. When it has, they will appear here per provider, taken from real requests. Nothing on this page is estimated: prices, context window and capabilities all come from the provider’s own documentation.

Call it

Same request shape as every other model in the catalogue.

example.ts
const res = await client.chat.completions.create({
  model: "qwen/qwen3-32b",
  messages: [{ role: "user", content: "Hello" }],
  route: {
    // pin, exclude or order the providers below
    providers: ["deepinfra", "openrouter"],
    sort: "price",
  },
});

Send Qwen3 32B a real request

Signing up takes no card and commits you to nothing — you start at a zero balance, so nothing can be charged until you decide to add credit. The playground opens on Qwen3 32B and prices every answer as it arrives, at the rate shown above — what the route costs us, with nothing of ours added. Already have an account? The second button skips the form.

Qwen3 32B — API and pricing · Multigrid