Skip to content

OpenAI

Live·14 models

The broadest frontier line-up, and the wire format every other provider ended up copying.

Frontier reasoningTool callingStructured output

14 models on OpenAI

OpenAI's own rate for each of these. Where another provider serves the same model it may charge differently. The model page shows every route side by side.

ModelPriceContext
Embedding 3 Smallopenai/text-embedding-3-small$0.020 inFree out8K
Embedding 3 Largeopenai/text-embedding-3-large$0.130 inFree out8K
GPT-5 nanoopenai/gpt-5-nano$0.050 in$0.400 out400K
GPT-4o miniopenai/gpt-4o-mini$0.150 in$0.600 out128K
GPT-5.6 Lunaopenai/gpt-5.6-luna$0.200 in$1.20 out1.0M
GPT-4.1 miniopenai/gpt-4.1-mini$0.400 in$1.60 out1.0M
GPT-5 miniopenai/gpt-5-mini$0.250 in$2.00 out400K
o4-miniopenai/o4-mini$1.10 in$4.40 out200K
GPT-4.1openai/gpt-4.1$2.00 in$8.00 out1.0M
o3openai/o3$2.00 in$8.00 out200K
GPT-5.1openai/gpt-5.1$1.25 in$10.00 out400K
GPT-4oopenai/gpt-4o$2.50 in$10.00 out128K
GPT-5.6 Terraopenai/gpt-5.6-terra$2.00 in$12.00 out1.0M
GPT-5.6 Solopenai/gpt-5.6-sol$4.00 in$20.00 out1.0M

Availability, from our vantage point

What OpenAI looked like from Multigrid's servers, not OpenAI's own uptime figure, which only they can publish.

100.0% up0.0% degradedReachability probes: 99.98% of 5658, 100% of the window observed.
Probes 99.98%Real traffic no traffic yet

The ring is reachability probes, measured across 100% of the last 30 days a share of what we watched, not of the month. We probe when someone loads this page, so the gaps are ours.

  • Succeeded
  • Failed on a degraded day
  • Failed on a bad day (< 50%)
  • Nothing measured yet

The two lanes are never averaged. Probes ask whether OpenAI is answering at all; traffic is the share of real requests we routed there that completed. A provider can pass every probe while completing nothing, which is why they are stacked rather than blended, and why a green top lane over a red bottom one is the pattern worth looking for.

Coverage is ours, not theirs: we probe when someone loads the status page, at most once every five minutes, so an empty slot means we were not looking rather than that anything was wrong. For an authoritative account of an incident, read OpenAI’s own status page. They are the only ones who can write it.

None of this is latency. What happened to your own requests (timing, errors and cost per provider) is on your analytics page, measured from traffic you actually sent.

Reach OpenAI on one balance

We hold a platform key for OpenAI, so everything above is reachable on Multigrid credit: one balance across every vendor, at the prices listed. Signing up takes no card and commits you to nothing. If you already have a OpenAI contract, bring that key instead and we take no percentage on the traffic.