Bertrande

Llama vs Qwen

The short answer. The cheapest Llama model we track, Llama 3.1 8B Instruct, costs $0.080 per million output tokens and $0.050 per million input tokens, as sold by OpenRouter. The cheapest Qwen model, Qwen3.7 Flash, costs $0.130 per million output tokens and $0.030 per million input tokens, as sold by OpenRouter. At the entry level, Llama is 1.6 times cheaper than Qwen on output tokens. Prices read on 18 September 2026.

This page compares what Meta and Alibaba charge through their APIs. Every price below carries the day we read it and links to the page it came from. We do not rank quality: we tell you what each model costs, and since when.

At the top of each range

The index is produced by Artificial Analysis and relayed through OpenRouter's catalogue. It is not our measurement; we record it as a level 4 source and show it next to the price we read. A model without a score here has not been scored yet, which is not the same as scoring low.

Every model, side by side

Llama models

8 standard offers, from OpenRouter's prices: we track no price list published by Meta itself for these models.

ModelInputOutputContextRead onSource
Llama 3.1 8B Instruct$0.050$0.080131 07218 September 2026OpenRouter
Llama Guard 4 12B$0.180$0.180163 84018 September 2026OpenRouter
Llama 3.2 1B Instruct$0.027$0.20160 00018 September 2026OpenRouter
Llama 4 Scout$0.100$0.3001 310 72018 September 2026OpenRouter
Llama 3.3 70B Instruct$0.100$0.320131 07218 September 2026OpenRouter
Llama 3.2 3B Instruct$0.050$0.330131 07218 September 2026OpenRouter
Llama 3.1 70B Instruct$0.400$0.400131 07218 September 2026OpenRouter
Llama 4 Maverick$0.1875$0.65251 048 57618 September 2026OpenRouter

Qwen models

52 standard offers, from OpenRouter's prices: we track no price list published by Alibaba itself for these models.

ModelInputOutputContextRead onSource
Qwen3.7 Flash$0.030$0.1301 000 00018 September 2026OpenRouter
Qwen3.5-9B$0.100$0.150262 14418 September 2026OpenRouter
Qwen3 30B A3B Instruct 2507$0.04815$0.19305262 14418 September 2026OpenRouter
Qwen2.5 7B Instruct$0.100$0.20032 76818 September 2026OpenRouter
Qwen3 14B$0.120$0.240131 07218 September 2026OpenRouter
Qwen3.5-Flash$0.065$0.2601 000 00018 September 2026OpenRouter
Qwen3 Coder 30B A3B Instruct$0.070$0.280262 14418 September 2026OpenRouter
Qwen3 32B$0.080$0.280131 07218 September 2026OpenRouter
Qwen3 235B A22B Instruct 2507$0.0875$0.350262 14418 September 2026OpenRouter
Qwen2.5 72B Instruct$0.360$0.40032 76818 September 2026OpenRouter
Qwen3 VL 32B Instruct$0.104$0.416131 07218 September 2026OpenRouter
Qwen3 8B$0.117$0.455131 07218 September 2026OpenRouter
Qwen3 VL 8B Instruct$0.117$0.455262 14418 September 2026OpenRouter
Qwen3.8 Flash$0.150$0.4701 000 00018 September 2026OpenRouter
Qwen3 30B A3B$0.120$0.500131 07218 September 2026OpenRouter
Qwen3 VL 30B A3B Instruct$0.130$0.520262 14418 September 2026OpenRouter
Qwen Plus 0728$0.260$0.7801 000 00018 September 2026OpenRouter
Qwen-Plus$0.260$0.7801 000 00018 September 2026OpenRouter
Qwen3 Coder Next$0.120$0.800262 14418 September 2026OpenRouter
Qwen3.6 35B A3B$0.100$0.900262 14418 September 2026OpenRouter
Qwen3 Coder Flash$0.195$0.9751 000 00018 September 2026OpenRouter
Qwen3 Coder 480B A35B$0.300$1.00262 14418 September 2026OpenRouter
Qwen2.5 Coder 32B Instruct$0.660$1.0032 76818 September 2026OpenRouter
Qwen2.5 VL 72B Instruct$0.800$1.00128 00018 September 2026OpenRouter
Qwen3 Next 80B A3B Instruct$0.090$1.10262 14418 September 2026OpenRouter
Qwen3.6 Flash$0.1875$1.1251 000 00018 September 2026OpenRouter
Qwen3 Next 80B A3B Thinking$0.150$1.20262 14418 September 2026OpenRouter
Qwen3.7 Plus$0.320$1.281 000 00018 September 2026OpenRouter
Qwen3.5-35B-A3B$0.1625$1.30262 14418 September 2026OpenRouter
Qwen3.5-27B$0.195$1.56262 14418 September 2026OpenRouter
Qwen3.5 Plus 2026-02-15$0.260$1.561 000 00018 September 2026OpenRouter
Qwen3.5 Plus 2026-04-20$0.300$1.801 000 00018 September 2026OpenRouter
Qwen3 235B A22B$0.455$1.82131 07218 September 2026OpenRouter
Qwen3 VL 235B A22B Instruct$0.210$1.90262 14418 September 2026OpenRouter
Qwen3.6 Plus$0.325$1.951 000 00018 September 2026OpenRouter
Qwen3.6 27B$0.300$2.00262 14418 September 2026OpenRouter
Qwen3.5-122B-A10B$0.260$2.08262 14418 September 2026OpenRouter
Qwen3 VL 8B Thinking$0.180$2.10131 07218 September 2026OpenRouter
Qwen3 235B A22B Thinking 2507$0.230$2.30131 07218 September 2026OpenRouter
Qwen3 30B A3B Thinking 2507$0.200$2.4081 92018 September 2026OpenRouter
Qwen3 VL 30B A3B Thinking$0.200$2.40262 14418 September 2026OpenRouter
Qwen3.8 27B$0.214$2.551 000 00018 September 2026OpenRouter
Qwen3 Coder Plus$0.650$3.251 000 00018 September 2026OpenRouter
Qwen3.5 397B A17B$0.550$3.50262 14418 September 2026OpenRouter
Qwen3 Max$0.780$3.90262 14418 September 2026OpenRouter
Qwen3 Max Thinking$0.780$3.90262 14418 September 2026OpenRouter
Qwen3 VL 235B A22B Thinking$0.400$4.00131 07218 September 2026OpenRouter
Qwen3.7 Max$1.475$4.4251 000 00018 September 2026OpenRouter
Qwen3.8 2.4T A95B$2.00$6.001 048 57618 September 2026OpenRouter
Qwen3.8 Max (0902)$2.00$6.001 000 00016 September 2026OpenRouter
Qwen3.8 Max$2.00$6.001 000 0004 September 2026OpenRouter
Qwen3.6 Max Preview$1.027$6.162262 14418 September 2026OpenRouter

How these prices have moved

We hold dated captures of these models' prices back to 5 September 2024. We did not observe these values ourselves: they were recovered from public captures of OpenRouter's catalogue, and each links to its capture so you can check it without trusting us. Largest moves on output price:

ModelBeforeAfterChangeCaptured
Meta: Llama 3.2 1B Instruct$0.010$0.200+1900%14 November 2025
Meta: Llama 3.2 3B Instruct$0.020$0.340+1600%3 March 2026
Qwen: Qwen3 235B A22B Thinking 2507$0.100$1.495+1395%17 July 2026
Qwen: Qwen3 235B A22B$0.100$0.600+500%9 May 2025
Qwen: Qwen3 235B A22B Instruct 2507$0.100$0.600+500%23 September 2025
Qwen: Qwen3 235B A22B Instruct 2507$0.100$0.550+450%17 July 2026

Questions

Which is cheaper, Llama or Qwen?

The cheapest Llama model we track, Llama 3.1 8B Instruct, costs $0.080 per million output tokens and $0.050 per million input tokens, as sold by OpenRouter. The cheapest Qwen model, Qwen3.7 Flash, costs $0.130 per million output tokens and $0.030 per million input tokens, as sold by OpenRouter. At the entry level, Llama is 1.6 times cheaper than Qwen on output tokens. Prices read on 18 September 2026.

What is the cheapest Llama model?

Llama 3.1 8B Instruct, at $0.080 per million output tokens, read on 18 September 2026.

What is the cheapest Qwen model?

Qwen3.7 Flash, at $0.130 per million output tokens, read on 18 September 2026.

Where do these prices come from?

From the price lists OpenRouter and OpenRouter publish, read by our collectors. No figure on this page was produced by a language model.

Other comparisons

All comparisons