uttapen

GLM 4.5 vs GLM 5

Both sit behind the same key and the same code on uttapen — the only thing that changes is the model string, so you can run each of them against your own workload without touching anything else.

FeatureGLM 4.5GLM 5
ProviderZ.ai (GLM)Z.ai (GLM)
Model idz-ai/glm-4.5z-ai/glm-5
Context131,072 tokens204,800 tokens
Max output98,304 tokens128,000 tokens
Input / 1M tokens174,075 toman174,075 toman
Output / 1M tokens638,275 toman557,040 toman
≈ one 1,000-word request1,056 toman950 toman
Input caching31,914 toman / 1M34,815 toman / 1M
Image inputNoNo
File inputNoNo
Tool callingYesYes
JSON outputYesYes
Reasoning modeYesYes

Which one for what?

GLM 4.5

  • Multi-step problems, maths, and tracking down a bug
  • Agents that call your own APIs and database
1,056 toman per 1,000 words · Model page

GLM 5

  • Multi-step problems, maths, and tracking down a bug
  • Whole documents or a codebase in a single request
  • Agents that call your own APIs and database
950 toman per 1,000 words · Model page

Run both with the same code

Swap the model value between the two ids and send the same request twice; what each answer cost comes back in the X-Uttapen-Cost-Toman header.

curl https://api.uttapen.ir/v1/chat/completions \
  -H "Authorization: Bearer sk-up-…" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "z-ai/glm-4.5",
    "messages": [{"role": "user", "content": "Introduce yourself in one sentence."}],
    "stream": true
  }'

model = "z-ai/glm-5"

More comparisons: all pairs · model rankings · full catalogue

base_url = https://api.uttapen.ir/v1