Skip to content
Together AI logo

Together AI

Live·11 models

The widest open-weight catalogue, including models nobody else bothers to host.

Widest open catalogueFine-tuningDedicated endpoints

11 models on Together AI

Together AI's own rate for each of these. Where another provider serves the same model it may charge differently — the model page shows every route side by side.

ModelPriceContext
Multilingual E5 Large Instructintfloat/multilingual-e5-large-instruct$0.020 inFree out1K
FLUX.1 [schnell]black-forest-labs/flux-1-schnell$0.002700 / image
Gemma 3n E4B Instructgoogle/gemma-3n-e4b$0.060 in$0.120 out33K
GPT-OSS 20Bopenai/gpt-oss-20b$0.050 in$0.200 out128K
Qwen3.5 9Bqwen/qwen3.5-9b$0.170 in$0.250 out262K
Llama 3.3 70B Instructmeta/llama-3.3-70b-instruct$1.04 in$1.04 out131K
MiniMax M3minimax/minimax-m3$0.300 in$1.20 out524K
Qwen3.7 Plusqwen/qwen3.7-plus$0.320 in$1.28 out1M
DeepSeek V4 Prodeepseek/deepseek-v4-pro$1.74 in$3.48 out512K
GLM-5.2zai/glm-5.2$1.40 in$4.40 out262K
Kimi K3moonshot/kimi-k3$3.00 in$15.00 out1M

Availability and latency

Multigrid does not measure Together AI's availability, and will not publish numbers it did not measure.

For live status, read Together AI’s own status page — it is the only authoritative source, and the one they update during an incident. Our own reachability probes against every provider are on the status page, and they answer “is anyone home”, not “is it fast”.

What Multigrid can tell you is what happened to your own requests: latency, errors and cost per provider are on your analytics page, measured from traffic you actually sent.

Reach Together AI on one balance

We hold a platform key for Together AI, so everything above is reachable on Multigrid credit — one balance across every vendor, at the prices listed. Signing up takes no card and commits you to nothing. If you already have a Together AI contract, bring that key instead and we take no percentage on the traffic.

Together AI — provider · Multigrid