LLM Price Comparison
Input, output and cache pricing for every model, sorted per 1M tokens with filters.
Showing 458 of 464 models
Space Bunny Alpha stealth | Free | Free | Free | — | 1.00M | ToolsReasoningVision | |
Ling 3.0 Flash Sante (free) inclusionAI | Free | Free | Free | — | 262K | ToolsReasoning | |
Qwen3.8 27B (free) Qwen | Free | Free | Free | — | 262K | ToolsReasoningVision | |
Dots3-Note Preview (free) Dots Studio | Free | Free | Free | — | 512K | ToolsReasoningVision | |
LFM2.5-2.6B (free) LiquidAI | Free | Free | Free | — | 65.5K | ToolsReasoning | |
Nemotron 3.5 Lightning (free) NVIDIA | Free | Free | Free | — | 1.00M | ToolsReasoning | |
Inkling Small (free) Thinking Machines | Free | Free | Free | — | 1.05M | ToolsReasoningVision | |
Laguna S 2.1 (free) Poolside | Free | Free | Free | — | 262K | ToolsReasoning | |
Inkling (free) Thinking Machines | Free | Free | Free | — | 1.05M | ToolsReasoningVision | |
Laguna XS 2.1 (free) Poolside | Free | Free | Free | — | 262K | ToolsReasoning | |
North Mini Code (free) Cohere | Free | Free | Free | — | 256K | ToolsReasoning | |
Nemotron 3.5 Content Safety (free) NVIDIA | Free | Free | Free | — | 128K | ReasoningVision | |
Nemotron 3 Ultra (free) NVIDIA | Free | Free | Free | — | 1.00M | ToolsReasoning | |
Nemotron 3 Nano Omni (free) NVIDIA | Free | Free | Free | — | 256K | ToolsReasoningVision | |
Gemma 4 26B A4B (free) | Free | Free | Free | — | 262K | ToolsReasoningVision | |
Gemma 4 31B (free) | Free | Free | Free | — | 262K | ToolsReasoningVision | |
Lyria 3 Pro Preview | Free | Free | Free | — | 1.05M | Vision | |
Lyria 3 Clip Preview | Free | Free | Free | — | 1.05M | Vision | |
Nemotron 3 Super (free) NVIDIA | Free | Free | Free | — | 262K | ToolsReasoning | |
Free Models Router openrouter | Free | Free | Free | — | 200K | ToolsReasoningVision | |
Mistral Nemo Mistral | $0.019 | $0.030 | $0.022 | — | 131K | Tools | |
Ling 3.0 Flash VL inclusionAI | $0.021 | $0.062 | $0.031 | $0.019 | 262K | ToolsReasoningVision | |
Ling 3.0 Flash inclusionAI | $0.021 | $0.063 | $0.032 | $0.019 | 262K | ToolsReasoning | |
gpt-oss-20b OpenAI | $0.018 | $0.090 | $0.036 | $0.029 | 131K | ToolsReasoning | |
DeepSeek V4 Flash Latest DeepSeek | $0.010 | $0.131 | $0.040 | $0.034 | 1.05M | ToolsReasoning | |
Granite 4.0 Micro IBM | $0.017 | $0.112 | $0.041 | — | 131K | ||
Llama 3 8B Lunaris Sao10K | $0.040 | $0.050 | $0.042 | — | 8.19K | ||
Nex-N2.5-Mini Nex AGI | $0.025 | $0.100 | $0.044 | $0.027 | 262K | ReasoningVision | |
gpt-oss-20b (batch) OpenAI | $0.024 | $0.112 | $0.046 | — | 131K | ToolsReasoning | |
Qwen3.7 Flash Qwen | $0.030 | $0.130 | $0.055 | $0.037 | 1.00M | ToolsReasoningVision | |
gpt-oss-120b (batch) OpenAI | $0.030 | $0.136 | $0.056 | — | 131K | ToolsReasoning | |
Mistral Small 3 Mistral | $0.050 | $0.080 | $0.057 | — | 32.8K | ||
Llama 3.1 8B Instruct Meta | $0.050 | $0.080 | $0.057 | $0.039 | 131K | Tools | |
Schematron V2 Turbo Inference.net | $0.030 | $0.150 | $0.060 | $0.060 | 128K | ||
Nova Micro 1.0 Amazon | $0.035 | $0.140 | $0.061 | — | 128K | Tools | |
Gemma 3 4B | $0.050 | $0.100 | $0.063 | — | 131K | Vision | |
Command R7B (12-2024) Cohere | $0.037 | $0.150 | $0.066 | — | 128K | ||
Mercury 2.5 Inception | $0.040 | $0.150 | $0.068 | $0.041 | 260K | ToolsReasoning | |
GPT-5 Nano (batch) OpenAI | $0.025 | $0.200 | $0.069 | $0.052 | 400K | ToolsReasoningVision | |
gpt-oss-120b OpenAI | $0.037 | $0.170 | $0.070 | — | 131K | ToolsReasoning | |
Llama 3.2 1B Instruct Meta | $0.027 | $0.201 | $0.071 | — | 60.0K | ||
Laguna XS 2.1 Poolside | $0.060 | $0.120 | $0.075 | $0.052 | 262K | ToolsReasoning | |
Ministral 3 8B 2512 (batch) Mistral | $0.075 | $0.075 | $0.075 | $0.024 | 262K | ToolsVision | |
Gemma 3 12B | $0.050 | $0.150 | $0.075 | — | 131K | ToolsVision | |
GLM Flash Latest Z.ai | $0.020 | $0.248 | $0.077 | $0.069 | 1.05M | ToolsReasoningVision | |
Hy-MT2-1.8B Tencent | $0.044 | $0.177 | $0.077 | — | 8.19K | ||
Qwen3 30B A3B Instruct 2507 Qwen | $0.048 | $0.193 | $0.084 | — | 262K | Tools | |
Nemotron 3.5 Lightning NVIDIA | $0.060 | $0.160 | $0.085 | $0.063 | 262K | ToolsReasoning | |
Solar Mini 4 Upstage | $0.050 | $0.200 | $0.087 | $0.054 | 524K | ToolsReasoning | |
Nemotron 3 Nano 30B A3B NVIDIA | $0.050 | $0.200 | $0.087 | $0.072 | 262K | ToolsReasoning | |
Gemini 2.5 Flash Lite (batch) | $0.050 | $0.200 | $0.087 | $0.057 | 1.05M | ToolsReasoningVision | |
GPT-4.1 Nano (batch) OpenAI | $0.050 | $0.200 | $0.087 | $0.059 | 1.05M | ToolsVision | |
MythoMax 13B gryphe | $0.080 | $0.110 | $0.087 | — | 8.19K | ||
Phi 4 Microsoft | $0.070 | $0.140 | $0.088 | — | 16.4K | ||
Ling 3.0 Flash Fin inclusionAI | $0.060 | $0.180 | $0.090 | $0.054 | 262K | ToolsReasoning | |
Schematron V2 Small Inference.net | $0.050 | $0.230 | $0.095 | $0.095 | 128K | ||
GLM 5.3 Flash (batch) Z.ai | $0.060 | $0.200 | $0.095 | $0.059 | 1.05M | ToolsReasoningVision | |
Reka Edge rekaai | $0.100 | $0.100 | $0.100 | — | 16.4K | ToolsReasoningVision | |
Ministral 3 3B 2512 Mistral | $0.100 | $0.100 | $0.100 | $0.033 | 131K | ToolsVision | |
GPT-6 Luna Pro (batch) OpenAI | $0.050 | $0.250 | $0.100 | $0.066 | 1.05M | ToolsReasoningVision |
Cost estimator
Enter one task's usage to convert pricing into per-call and monthly cost.
No model selected yet. Use the estimate button on a row to add one.
Prices are USD per 1M tokens, converted from OpenRouter's official rates. The cache price applies to reused long prompts: a cache hit bills the input at the discounted rate.
Input, output and cache pricing for every model, sorted per 1M tokens with filters.
Every model on OpenRouter bills input and output separately, and the gap between the two is routinely four or five times. Sorting on output alone picks the wrong model; so does sorting on input alone. This table puts both prices, the cache-read rate, context length and capability tags on one row so you can match a model to your own usage shape — and the estimator underneath converts pricing into what one task, and one month, actually cost.
How to use
- 1Narrow the list with the search box or the vendor dropdown, then stack filters for price band, context length and capabilities.
- 2Click a column header to sort by input, output or blended price — blended weights input three to one, which tracks a real bill more closely.
- 3Enter one task's input and output tokens in the estimator to compare monthly cost across candidates.