Gemini vs Llama
The short answer. The cheapest Gemini model we track, Gemini 2.5 Flash-Lite, costs $0.400 per million output tokens and an input price that depends on the type of content. The cheapest Llama model, Llama 3.1 8B Instruct, costs $0.080 per million output tokens and $0.050 per million input tokens, as sold by OpenRouter. At the entry level, Llama is 5.0 times cheaper than Gemini on output tokens. Prices read on 18 September 2026.
This page compares what Google and Meta charge through their APIs. Every price below carries the day we read it and links to the page it came from. We do not rank quality: we tell you what each model costs, and since when.
At the top of each range
- Among Gemini models that carry a relayed score, the highest on the Artificial Analysis intelligence index is Gemini 3.8 Flash, at 41.2, sold by OpenRouter at $3.75 per million output tokens.
- Among Llama models that carry a relayed score, the highest on the Artificial Analysis intelligence index is Llama 4 Maverick, at 9.3, sold by OpenRouter at $0.6525 per million output tokens.
The index is produced by Artificial Analysis and relayed through OpenRouter's catalogue. It is not our measurement; we record it as a level 4 source and show it next to the price we read. A model without a score here has not been scored yet, which is not the same as scoring low.
Every model, side by side
Gemini models
12 standard offers, from the prices Google publishes itself.
| Model | Input | Output | Context | Read on | Source |
|---|---|---|---|---|---|
| Gemini 2.5 Flash-Lite | by content type | $0.400 | 18 September 2026 | ||
| Gemini 3.1 Flash-Lite | by content type | $1.50 | 18 September 2026 | ||
| Gemini 2.5 Flash | by content type | $2.50 | 18 September 2026 | ||
| Gemini 3.5 Flash-Lite | by content type | $2.50 | 18 September 2026 | ||
| Gemini 3 Flash Preview | by content type | $3.00 | 18 September 2026 | ||
| Gemini 3.6 Flash | $0.750 | $3.75 | 18 September 2026 | ||
| Gemini 3.7 Flash | $0.750 | $3.75 | 18 September 2026 | ||
| Gemini 3.8 Flash | $0.750 | $3.75 | 18 September 2026 | ||
| Gemini Robotics ER 1.6 Preview | by content type | $5.00 | 11 September 2026 | ||
| Gemini Robotics ER 2 Preview | by content type | $5.00 | 17 September 2026 | ||
| Gemini Robotics ER 2 Streaming Preview | by content type | $5.00 | 17 September 2026 | ||
| Gemini 3.5 Flash | $1.50 | $9.00 | 18 September 2026 |
Llama models
8 standard offers, from OpenRouter's prices: we track no price list published by Meta itself for these models.
| Model | Input | Output | Context | Read on | Source |
|---|---|---|---|---|---|
| Llama 3.1 8B Instruct | $0.050 | $0.080 | 131 072 | 18 September 2026 | OpenRouter |
| Llama Guard 4 12B | $0.180 | $0.180 | 163 840 | 18 September 2026 | OpenRouter |
| Llama 3.2 1B Instruct | $0.027 | $0.201 | 60 000 | 18 September 2026 | OpenRouter |
| Llama 4 Scout | $0.100 | $0.300 | 1 310 720 | 18 September 2026 | OpenRouter |
| Llama 3.3 70B Instruct | $0.100 | $0.320 | 131 072 | 18 September 2026 | OpenRouter |
| Llama 3.2 3B Instruct | $0.050 | $0.330 | 131 072 | 18 September 2026 | OpenRouter |
| Llama 3.1 70B Instruct | $0.400 | $0.400 | 131 072 | 18 September 2026 | OpenRouter |
| Llama 4 Maverick | $0.1875 | $0.6525 | 1 048 576 | 18 September 2026 | OpenRouter |
How these prices have moved
We hold dated captures of these models' prices back to 5 September 2024. We did not observe these values ourselves: they were recovered from public captures of OpenRouter's catalogue, and each links to its capture so you can check it without trusting us. Largest moves on output price:
| Model | Before | After | Change | Captured |
|---|---|---|---|---|
| Meta: Llama 3.2 1B Instruct | $0.010 | $0.200 | +1900% | 14 November 2025 |
| Meta: Llama 3.2 3B Instruct | $0.020 | $0.340 | +1600% | 3 March 2026 |
| Meta: Llama 3.2 3B Instruct | $0.006 | $0.024 | +300% | 4 September 2025 |
| Meta: Llama Guard 4 12B | $0.050 | $0.180 | +260% | 12 August 2025 |
| Meta: Llama 3.3 70B Instruct | $0.036 | $0.120 | +233% | 24 September 2025 |
| Meta: Llama 3.3 70B Instruct | $0.120 | $0.390 | +225% | 12 October 2025 |
Questions
Which is cheaper, Gemini or Llama?
The cheapest Gemini model we track, Gemini 2.5 Flash-Lite, costs $0.400 per million output tokens and an input price that depends on the type of content. The cheapest Llama model, Llama 3.1 8B Instruct, costs $0.080 per million output tokens and $0.050 per million input tokens, as sold by OpenRouter. At the entry level, Llama is 5.0 times cheaper than Gemini on output tokens. Prices read on 18 September 2026.
What is the cheapest Gemini model?
Gemini 2.5 Flash-Lite, at $0.400 per million output tokens, read on 18 September 2026.
What is the cheapest Llama model?
Llama 3.1 8B Instruct, at $0.080 per million output tokens, read on 18 September 2026.
Where do these prices come from?
From the price lists Google and OpenRouter publish, read by our collectors. No figure on this page was produced by a language model.