Best Cheap LLM in 2026
Active AI models ranked by input price per million tokens. Prices from official provider pages.
Updated automatically as pricing changes. Full model database →
| # | Model | Input /1M |
|---|---|---|
| 1 | Baidu Qianfan: CoBuddyBaidu AI | free |
| 2 | Gemma 4 31B InstructGoogle DeepMind | free |
| 3 | Llama 3.2 3B InstructMeta AI | free |
| 4 | Owl AlphaOpenrouter | free |
| 5 | Qwen 3.6 Plus PreviewAlibaba / Qwen | free |
| 6 | Nex-N2-ProNex Agi | free |
| 7 | Google Lyria 3 Pro PreviewGoogle DeepMind | free |
| 8 | Hy3Tencent | free |
| 9 | Llama Guard 4 12BMeta AI | free |
| 10 | Google Lyria 3 Clip PreviewGoogle DeepMind | free |
| 11 | North Mini CodeCohere | free |
| 12 | Gemma 4 26B A4B ITGoogle DeepMind | free |
| 13 | Elephant AlphaOpenrouter | free |
| 14 | Nemotron 3.5 Content SafetyNVIDIA | free |
| 15 | Ling-2.6-flashInclusionai | free |
| 16 | Gemma 4 E2BGoogle DeepMind | free |
| 17 | Ling-2.6-1TInclusionai | free |
| 18 | Trinity Large PreviewArcee Ai | free |
| 19 | Trinity Large ThinkingArcee Ai | free |
| 20 | Nex-N2-MiniNex Agi | $0.025 |
| 21 | Qwen-TurboAlibaba / Qwen | $0.033 |
| 22 | Command R7B (12-2024)Cohere | $0.037 |
| 23 | Granite 4.1 8BIbm | $0.05 |
| 24 | Mistral Small 3Mistral AI | $0.05 |
| 25 | Qwen3.5-FlashAlibaba / Qwen | $0.065 |
| 26 | Hy3 PreviewTencent | $0.066 |
| 27 | Phi-4-miniMicrosoft | $0.07 |
| 28 | ERNIE 4.5 21B A3BBaidu AI | $0.07 |
| 29 | ERNIE 4.5 21B A3B ThinkingBaidu AI | $0.07 |
| 30 | ERNIE 4.5 VL 28B A3BBaidu AI | $0.07 |
Finding the best value LLM
Price alone doesn't tell the whole story. A model that costs twice as much but solves problems in half the calls is actually cheaper. When evaluating cost, consider:
- Input vs output pricing — for chat and RAG, input is usually 80%+ of your tokens. For generation-heavy tasks (writing, summarization), output price matters more.
- Context window — larger contexts let you process more in a single call, reducing round trips and total token usage.
- Open-weight models — if you can self-host, models like Mistral and LLaMA can cost near zero at scale. Check the “open weights” column on the model database.
- Quality vs cost — compare benchmark scores on the benchmark leaderboard to find the best performance per dollar.
Also see: Best Coding LLM, Best Reasoning LLM, Compare any two models.