Get the app

Which LLM should I use?

The right model depends on the job, so these lists answer the common questions from current API prices: what is cheapest, what reads the most, and what is the best value when you need reasoning and tools. They are computed from the 40 models we track and refresh weekly (last synced 4 Oct 2026).

The cheapest models that support tool use

Tool use (function calling) is what lets a model drive an agent or call your APIs. Ranked by price per million output tokens, since output is usually the larger cost.

ModelInputOutputContext
Ministral 3 14B 2512
Mistral AI
$0.200$0.200262K
Llama 4 Scout
Meta
$0.100$0.3001.3M
DeepSeek V4 Pro 0423
DeepSeek
$0.209$0.4181.0M
Qwen3.8 Omni Flash
Qwen (Alibaba)
$0.150$0.4701M
Qwen3.8 Flash
Qwen (Alibaba)
$0.150$0.4701M
GPT-6 Luna Pro
OpenAI
$0.100$0.5001.1M

The cheapest reasoning models

Reasoning models spend extra tokens thinking before they answer, which helps on maths, code and multi-step tasks. These support both reasoning and tools.

ModelInputOutputContext
DeepSeek V4 Pro 0423
DeepSeek
$0.209$0.4181.0M
Qwen3.8 Omni Flash
Qwen (Alibaba)
$0.150$0.4701M
Qwen3.8 Flash
Qwen (Alibaba)
$0.150$0.4701M
GPT-6 Luna Pro
OpenAI
$0.100$0.5001.1M
GPT-6 Luna
OpenAI
$0.100$0.5001.1M
GLM 5.3 Flash
Z.ai
$0.150$0.5001.0M

The largest context windows

Context is how much text a model can consider at once: whole codebases, long contracts, large document sets. A bigger window costs more to fill, so check the input price too.

ModelInputOutputContext
Llama 4 Scout
Meta
$0.100$0.3001.3M
GPT-6.1 Sol Pro
OpenAI
$2.00$10.001.1M
GPT-6.1 Sol
OpenAI
$2.00$10.001.1M
GPT-6 Luna Pro
OpenAI
$0.100$0.5001.1M
GPT-6 Luna
OpenAI
$0.100$0.5001.1M
Gemini 3.8 Flash
Google
$0.750$3.751.0M

Best value for serious work

Models with reasoning, tool use and at least a 200K-token context, cheapest first by output price. A reasonable shortlist to test on your own task.

ModelInputOutputContext
DeepSeek V4 Pro 0423
DeepSeek
$0.209$0.4181.0M
Qwen3.8 Omni Flash
Qwen (Alibaba)
$0.150$0.4701M
Qwen3.8 Flash
Qwen (Alibaba)
$0.150$0.4701M
GPT-6 Luna Pro
OpenAI
$0.100$0.5001.1M
GPT-6 Luna
OpenAI
$0.100$0.5001.1M
GLM 5.3 Flash
Z.ai
$0.150$0.5001.0M

By provider

How to choose

Price is only one input. Run your own prompts on two or three candidates and compare quality, latency and cost on your task, because benchmark rankings rarely predict that for you. Prices here are OpenRouter's listings and can differ from a provider's direct pricing; see the full comparison table for every model.

The daily AI brief, on your phone.

Feed, daily deep-dive and bytes — readable offline, with push alerts for the topics you follow.

Get it on Google Play