uttapen

Cheap models, ranked

At a few requests a day the price per million tokens is noise. At a hundred thousand requests a month it is the whole budget. This list is ordered by the cost of one realistic 1,000-word round trip — roughly 1,300 tokens in and the same out — rather than by a headline rate.

Most used in this category

  1. 1GPT-4o-mini118 tok174,075

Best value in this category

Score out of 100: 50% cheapness, 30% capabilities, 20% context size — not a quality benchmark.

#ModelContext≈ 1,000 wordsScore
1Gemini 2.5 Flash Lite (batch)1,048,5769483.8
2Qwen3.7 Flash1,000,0006082.5
3Muse Spark 1.2 Contributor1,048,57611381.8
4Muse Spark 1.3 Contributor1,048,57611381.8
5GPT-5 Nano (batch)400,0008581.5
6Nex-N2-Mini262,1444780.5
7Ling 3.0 Flash262,1443278.7
8GPT-4.1 Nano (batch)1,047,5769477.8
9GLM Flash Latest1,310,72011676.2
10Gemini 2.5 Flash Lite1,048,57618976.2
11GLM 5.3 Flash1,310,72012375.6
12Solar Pro 4524,2885774.8
13Qwen3.5-Flash1,000,00012374.7
14DeepSeek V4 Flash Latest1,310,7207974.4
15Mistral Nemo131,0721874

Full list of cheap models for high-volume work with prices →