← LLM Directory
Google
LLM Provider
Google DeepMind's Gemini model family. 1M+ context windows, multimodal, competitive pricing.
Flagship2 models
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Gemini 3.1 Pro Google's high-capability multimodal model for long context, coding agents, and Vertex / AI Studio deployments. | $2.50 | $15.00 | 1.0M |
Gemini 2.5 Pro 1M token context. Strong coding and reasoning. Best Gemini model. | $1.25 | $10.00 | 1.0M |
Mid-range2 models
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Gemini 2.5 Flash Fast, cost-efficient Gemini with 1M context window. | $0.150 | $0.600 | 1.0M |
Gemini 2.0 Flash Stable 2.0 Flash for production workloads. | $0.100 | $0.400 | 1.0M |
Budget1 model
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Gemini 2.0 Flash Lite Ultra-lightweight model for high-volume tasks. | $0.075 | $0.300 | 1.0M |
Open Source5 models
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Gemma 3 27B Open-weight instruction-tuned model. Free to download and run locally. | $0.000 | $0.000 | 128K |
Gemma 3 12B Mid-size open-weight model. Good balance for local deployment. | $0.000 | $0.000 | 128K |
Gemma 3 4B Small open-weight model. Runs on consumer hardware. | $0.000 | $0.000 | 128K |
Gemma 4 QAT 32B Quantization-aware training release. 72% less VRAM needed. Runs on 16GB GPUs. | $0.000 | $0.000 | 256K |
Gemma 4 QAT 14B Mid-size QAT model. Efficient inference for 8GB GPUs. | $0.000 | $0.000 | 128K |
Legacy2 models
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Gemini 1.5 Pro 2M context window model. Good for long document analysis. | $1.25 | $5.00 | 2.1M |
Gemini 1.5 Flash Fast, affordable 1M context model. Superseded by 2.5 Flash. | $0.075 | $0.300 | 1.0M |
Image1 model
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
Imagen 3 Text-to-image generation model. Priced per image. | $0.020 | $0.020 | — |