Pricing guide

Cheapest AI Models That Still Hold Up

Cheap is only useful if the model is still competent. These are the low-cost options that still deserve production trials.

ModelInputOutputTotal price signalBest for
GPT-OSS-120B
OpenAI
$0$0$0.00Self-hosted
DeepSeek V4 Flash
DeepSeek
$0.14$0.28$0.42Low-cost agent experiments
DeepSeek V4 Pro
DeepSeek
$0.44$1.32$1.76Budget coding
GPT-5.6 Luna
OpenAI
$0.4$1.8$2.20High-volume workflows
Gemini 3.5 Flash-Lite
Google
$0.3$2.5$2.80High-volume automation
GLM-5
Zhipu AI
$1$3.2$4.20Bilingual (CN/EN)
Gemini 3.7 Flash
Google
$0.75$3.75$4.50Agentic coding
GLM-5.3
Z.ai
$1.4$4.4$5.80Complex software engineering
Claude Haiku 4.5
Anthropic
$1$5$6.00Fast responses
Grok 4.5
xAI
$2$6$8.00Agentic coding
Grok 4.6
xAI
$2$6$8.00Agentic coding
Gemini 3.6 Flash
Google
$1.5$7.5$9.00Agentic coding
Mistral Medium 3.5
Mistral
$1.5$7.5$9.00European compliance
Gemini 3.5 Flash
Google
$1.5$9$10.50Fast multimodal agents
Claude Sonnet 5
Anthropic
$2$10$12.00Balanced performance
GPT-5.2-Codex
OpenAI
$1.75$14$15.75Coding-focused tasks
GPT-5.2
OpenAI
$1.75$14$15.75General-purpose
GPT-5.4
OpenAI
$2.5$15$17.50Coding
Kimi K3
Moonshot AI
$3$15$18.00Long-horizon coding
GPT-5.6 Terra
OpenAI
$4$18$22.00Production agents
Claude Opus 5
Anthropic
$5$25$30.00Complex reasoning
Claude Opus 4.8
Anthropic
$5$25$30.00Complex reasoning
GPT-5.5
OpenAI
$5$30$35.00Complex reasoning
GPT-5.6 Sol
OpenAI
$10$45$55.00Complex production workflows
Claude Fable 5
Anthropic
$10$50$60.00Long-running agents

Best ultra-budget option

DeepSeek V4 Pro combines low listed API pricing, 1M context, and hybrid thinking modes for cost-sensitive production trials.

Best value frontier model

MiniMax M2.5 is the best balance of serious quality and low token cost.

Best bilingual value

GLM-5 is still one of the strongest value models if Chinese/English matters.

When cheap is the right choice