Together AI
The widest open-weight catalogue, including models nobody else bothers to host.
11 models on Together AI
Together AI's own rate for each of these. Where another provider serves the same model it may charge differently — the model page shows every route side by side.
| Model | Price | Context |
|---|---|---|
| Multilingual E5 Large Instructintfloat/multilingual-e5-large-instruct | $0.020 inFree out | 1K |
| FLUX.1 [schnell]black-forest-labs/flux-1-schnell | $0.002700 / image | — |
| Gemma 3n E4B Instructgoogle/gemma-3n-e4b | $0.060 in$0.120 out | 33K |
| GPT-OSS 20Bopenai/gpt-oss-20b | $0.050 in$0.200 out | 128K |
| Qwen3.5 9Bqwen/qwen3.5-9b | $0.170 in$0.250 out | 262K |
| Llama 3.3 70B Instructmeta/llama-3.3-70b-instruct | $1.04 in$1.04 out | 131K |
| MiniMax M3minimax/minimax-m3 | $0.300 in$1.20 out | 524K |
| Qwen3.7 Plusqwen/qwen3.7-plus | $0.320 in$1.28 out | 1M |
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | $1.74 in$3.48 out | 512K |
| GLM-5.2zai/glm-5.2 | $1.40 in$4.40 out | 262K |
| Kimi K3moonshot/kimi-k3 | $3.00 in$15.00 out | 1M |
Availability and latency
Multigrid does not measure Together AI's availability, and will not publish numbers it did not measure.
For live status, read Together AI’s own status page — it is the only authoritative source, and the one they update during an incident. Our own reachability probes against every provider are on the status page, and they answer “is anyone home”, not “is it fast”.
What Multigrid can tell you is what happened to your own requests: latency, errors and cost per provider are on your analytics page, measured from traffic you actually sent.
Reach Together AI on one balance
We hold a platform key for Together AI, so everything above is reachable on Multigrid credit — one balance across every vendor, at the prices listed. Signing up takes no card and commits you to nothing. If you already have a Together AI contract, bring that key instead and we take no percentage on the traffic.