Skip to content
Microsoft logo

WizardLM-2 8x22B

microsoft/wizardlm-2-8x22bPOST /v1/chat/completions

WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...

Structured outputStreamingPrompt caching
Input / 1M
$0.6541
via OpenRouter
Output / 1M
$0.6541
via OpenRouter
Cached input / 1M
$0.6541
on a cache hit
Context
66K
max out 8K
Providers
1
single route

1 provider serves this model

Multigrid picks between them on every request using your policy. Each price is what that route costs us, and we add no spread of our own to it.

ProviderUpstream model idIn / 1MOut / 1MBillingUptime
OpenRoutermicrosoft/wizardlm-2-8x22b$0.6541$0.6541Credit100.0% up0.0% degradedReachability probes: 99.98% of 5666, 100% of the window observed.
This model is reached through OpenRouter, and the rate reflects that

We hold no direct key for Microsoft, so every route above goes through OpenRouter. We do not buy tokens from them. We buy their credit, and a dollar of it costs $1.055. The rate shown here is their price with that 5.5% included, because that is what the route costs us. We add nothing on top of it.

The consequence, plainly: for this model we cannot be cheaper than OpenRouter. Their fee sits underneath ours. What you get here instead is one balance and one key across the whole catalogue, failover, spend caps and the audit trail, and if you already hold a Microsoft contract, bring that key and we take nothing at all.

What the uptime column is, and is not
It is availability of the provider, not of WizardLM-2 8x22B in particular: reachability probes we ran, and the success rate of requests we routed to that provider across the whole catalogue. Time to first token and throughput are still absent, because Multigrid has not served enough traffic through this model to publish honest ones. Nothing here is estimated: prices, context window and capabilities come from the provider’s own documentation, and the ribbon comes from rows we wrote.

Call it

Same request shape as every other model in the catalogue.

example.ts
const res = await client.chat.completions.create({
  model: "microsoft/wizardlm-2-8x22b",
  messages: [{ role: "user", content: "Hello" }],
});

Send WizardLM-2 8x22B a real request

Signing up takes no card and commits you to nothing. You start at a zero balance, so nothing can be charged until you decide to add credit. The playground opens on WizardLM-2 8x22B and prices every answer as it arrives, at the rate shown above: what the route costs us, with nothing of ours added. Already have an account? The second button skips the form.