← LLM Directory
Together AI
Hosted PlatformServerless GPU cloud for open-weight models. Fast inference, competitive pricing, 200+ models.
Flagship1 model
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Llama 4 Maverick Meta's latest MoE on fast serverless inference. | $0.180 | $0.540 | 1.0M |
Reasoning1 model
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
DeepSeek R1 DeepSeek reasoning on Together. | $0.550 | $2.19 | 131K |
Coding1 model
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Codestral 22B Mistral's coding specialist. | $0.150 | $0.450 | 66K |
Popular6 models
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Llama 3.3 70B Stable 70B on fast inference. | $0.880 | $0.880 | 131K |
Qwen3 235B-A22B Qwen's MoE model on Together. | $0.800 | $1.60 | 131K |
Qwen3 32B Dense 32B at low cost. | $0.300 | $0.600 | 131K |
DeepSeek V3 672B MoE at unbeatable pricing. | $0.140 | $0.280 | 66K |
Gemma 3 27B Google's 27B open-weight on Together. | $0.240 | $0.240 | 131K |
Mistral Small 3.1 24B Mistral's latest small model. | $0.100 | $0.100 | 131K |
Image2 models
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
FLUX.2 Pro Top open-weight image model. $0.03/MP. | $0.030 | $0.030 | — |
FLUX.2 Dev Open-weight image generation. Great quality per dollar. | $0.010 | $0.010 | — |