Cheapest Good LLMs

Value-ranked with real ratings and real prices — because a cheap model that fails the task is the most expensive one.

Live data · updated 2026-09-15 18:18 · sources: LMArena + OpenRouter

Our verdict

glm-5.3-flash: arena-top-20 quality at $0.25/1M output — the best deal in AI right now.

ModelOutput $/1MArena ratingValueLicense
glm-5.3-flash$0.251588★★★★★MIT
qwen3.8-flash-next$0.471635★★★☆☆qwen-community-1.0
gemini-3.8-flash-high$3.751493★☆☆☆☆Proprietary
muse-spark-1.3-max$4.251645★☆☆☆☆Proprietary
muse-spark-1.2 (xHigh)$4.251500★☆☆☆☆Proprietary
qwen3.8-max-0902$61681★☆☆☆☆Proprietary
grok-4.6-high$61596★☆☆☆☆Proprietary
gemini-omni-1.1-flash$91515★☆☆☆☆Proprietary
kimi-k3-max$151674★☆☆☆☆Kimi K3 license
gpt-5.6-sol-xhigh (codex-harness)$201604★☆☆☆☆Proprietary

Why cheap no longer means bad

As of 2026-09-15 18:18, the #20 model on LMArena (muse-spark-1.2 (xHigh), 1500) would have topped the board a year ago — and it costs $4.25 per million output tokens. The value tier now clears bars that frontier models set last year.

Where budget models still lose

Long agentic runs, subtle instruction following and tool-call reliability still separate the top 5 from the pack. For interactive chat and batch transforms, value models are fine; for autonomous agents, measure failure cost before celebrating the price.

Ready to build? Access all of these models through one API key. Try OpenRouter →