uttapen

Llama 3 8B Lunaris vs MythoMax 13B

Both sit behind the same key and the same code on uttapen — the only thing that changes is the model string, so you can run each of them against your own workload without touching anything else. On cost, Llama 3 8B Lunaris comes out about 1.3× cheaper.

FeatureLlama 3 8B LunarisMythoMax 13B
ProviderSao10kGryphe
Model idsao10k/l3-lunaris-8bgryphe/mythomax-l2-13b
Context8,192 tokens8,192 tokens
Max output7,372 tokens3,686 tokens
Input / 1M tokens11,605 toman17,408 toman
Output / 1M tokens14,506 toman17,408 toman
≈ one 1,000-word request34 toman45 toman
Input cachingNoNo
Image inputNoNo
File inputNoNo
Tool callingNoNo
JSON outputYesYes
Reasoning modeNoNo

Which one for what?

Llama 3 8B Lunaris

  • High-volume work where cost per call decides
34 toman per 1,000 words · Model page

MythoMax 13B

  • High-volume work where cost per call decides
45 toman per 1,000 words · Model page

Run both with the same code

Swap the model value between the two ids and send the same request twice; what each answer cost comes back in the X-Uttapen-Cost-Toman header.

curl https://api.uttapen.ir/v1/chat/completions \
  -H "Authorization: Bearer sk-up-…" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "sao10k/l3-lunaris-8b",
    "messages": [{"role": "user", "content": "Introduce yourself in one sentence."}],
    "stream": true
  }'

model = "gryphe/mythomax-l2-13b"

More comparisons: all pairs · model rankings · full catalogue

base_url = https://api.uttapen.ir/v1