Bertrande

GLM vs Llama

The short answer. The cheapest GLM model we track, GLM-OCR, costs $0.030 per million output tokens and $0.030 per million input tokens. The cheapest Llama model, Llama 3.1 8B Instruct, costs $0.080 per million output tokens and $0.050 per million input tokens, as sold by OpenRouter. At the entry level, GLM is 2.7 times cheaper than Llama on output tokens. Prices read on 11 September 2026 and 18 September 2026.

This page compares what Z.ai and Meta charge through their APIs. Every price below carries the day we read it and links to the page it came from. We do not rank quality: we tell you what each model costs, and since when.

At the top of each range

The index is produced by Artificial Analysis and relayed through OpenRouter's catalogue. It is not our measurement; we record it as a level 4 source and show it next to the price we read. A model without a score here has not been scored yet, which is not the same as scoring low.

Every model, side by side

GLM models

17 standard offers, from the prices Z.ai publishes itself.

ModelInputOutputContextRead onSource
GLM-OCR$0.030$0.03011 September 2026Z.ai
GLM-4-32B-0414-128K$0.100$0.10011 September 2026Z.ai
GLM-4.6V-FlashX$0.040$0.40011 September 2026Z.ai
GLM-4.7-FlashX$0.070$0.40011 September 2026Z.ai
GLM-5.3-Flash$0.150$0.50018 September 2026Z.ai
GLM-4.6V$0.300$0.90018 September 2026Z.ai
GLM-4.5-Air$0.200$1.1018 September 2026Z.ai
GLM-4.5V$0.600$1.8018 September 2026Z.ai
GLM-4.5$0.600$2.2016 September 2026Z.ai
GLM-4.6$0.600$2.2018 September 2026Z.ai
GLM-4.7$0.600$2.2011 September 2026Z.ai
GLM-5$1.00$3.2018 September 2026Z.ai
GLM-5.1$1.40$4.4015 September 2026Z.ai
GLM-5.2$1.40$4.4018 September 2026Z.ai
GLM-5.3$1.40$4.4018 September 2026Z.ai
GLM-4.5-AirX$1.10$4.5011 September 2026Z.ai
GLM-4.5-X$2.20$8.9011 September 2026Z.ai

Llama models

8 standard offers, from OpenRouter's prices: we track no price list published by Meta itself for these models.

ModelInputOutputContextRead onSource
Llama 3.1 8B Instruct$0.050$0.080131 07218 September 2026OpenRouter
Llama Guard 4 12B$0.180$0.180163 84018 September 2026OpenRouter
Llama 3.2 1B Instruct$0.027$0.20160 00018 September 2026OpenRouter
Llama 4 Scout$0.100$0.3001 310 72018 September 2026OpenRouter
Llama 3.3 70B Instruct$0.100$0.320131 07218 September 2026OpenRouter
Llama 3.2 3B Instruct$0.050$0.330131 07218 September 2026OpenRouter
Llama 3.1 70B Instruct$0.400$0.400131 07218 September 2026OpenRouter
Llama 4 Maverick$0.1875$0.65251 048 57618 September 2026OpenRouter

How these prices have moved

We hold dated captures of these models' prices back to 5 September 2024. We did not observe these values ourselves: they were recovered from public captures of OpenRouter's catalogue, and each links to its capture so you can check it without trusting us. Largest moves on output price:

ModelBeforeAfterChangeCaptured
Meta: Llama 3.2 1B Instruct$0.010$0.200+1900%14 November 2025
Meta: Llama 3.2 3B Instruct$0.020$0.340+1600%3 March 2026
Z.ai: GLM 4.5$0.200$0.800064+300%12 August 2025
Meta: Llama 3.2 3B Instruct$0.006$0.024+300%4 September 2025
Z.ai: GLM 4.5 Air$0.220$0.850+286%19 February 2026
Meta: Llama Guard 4 12B$0.050$0.180+260%12 August 2025

Questions

Which is cheaper, GLM or Llama?

The cheapest GLM model we track, GLM-OCR, costs $0.030 per million output tokens and $0.030 per million input tokens. The cheapest Llama model, Llama 3.1 8B Instruct, costs $0.080 per million output tokens and $0.050 per million input tokens, as sold by OpenRouter. At the entry level, GLM is 2.7 times cheaper than Llama on output tokens. Prices read on 11 September 2026 and 18 September 2026.

What is the cheapest GLM model?

GLM-OCR, at $0.030 per million output tokens, read on 11 September 2026.

What is the cheapest Llama model?

Llama 3.1 8B Instruct, at $0.080 per million output tokens, read on 18 September 2026.

Where do these prices come from?

From the price lists Z.ai and OpenRouter publish, read by our collectors. No figure on this page was produced by a language model.

Other comparisons

All comparisons