Live model catalog
Every model below is normalized into one shape: price per million input and output tokens, context window, and capabilities. Upstream schemas never reach the UI directly.
Live· 432458 upstream records · 57 providers
- OpenRouterlivehttps://openrouter.ai/api/v1/models
458 records · 159ms
- models.devlivehttps://models.dev/api.json
8,468 records · 428ms
Capability coverage
tools
432
of 432 models
reasoning
77
of 432 models
vision
284
of 432 models
json
432
of 432 models
longContext
312
of 432 models
Cheapest first
first 60 of 432
| Model | Provider | In /MTok | Out /MTok | Context | Capabilities | Weights |
|---|---|---|---|---|---|---|
| Mistral: Mistral Nemomistralai:mistral-nemo | mistralai | $0.029 | $0.03 | 131k | toolsjson | hosted |
| inclusionAI: Ling 3.0 Flash VLinclusionai:ling-3.0-flash-vl | inclusionai | $0.021 | $0.0616 | 262k | toolsvisionjsonlongContext | hosted |
| inclusionAI: Ling 3.0 Flashinclusionai:ling-3.0-flash | inclusionai | $0.021 | $0.063 | 262k | toolsjsonlongContext | hosted |
| Sao10K: Llama 3 8B Lunarissao10k:l3-lunaris-8b | sao10k | $0.04 | $0.05 | 8k | toolsjson | hosted |
| OpenAI: gpt-oss-20bopenai:gpt-oss-20b | openai | $0.018 | $0.09 | 131k | toolsjson | hosted |
| Nex AGI: Nex-N2.5-Mininex-agi:nex-n2.5-mini | nex-agi | $0.025 | $0.1 | 262k | toolsvisionjsonlongContext | hosted |
| IBM: Granite 4.0 Microibm-granite:granite-4.0-h-micro | ibm-granite | $0.017 | $0.112 | 131k | toolsjson | hosted |
| Mistral: Mistral Small 3mistralai:mistral-small-24b-instruct-2501 | mistralai | $0.05 | $0.08 | 33k | toolsjson | hosted |
| Meta: Llama 3.1 8B Instructmeta-llama:llama-3.1-8b-instruct | meta-llama | $0.05 | $0.08 | 131k | toolsjson | hosted |
| OpenAI: gpt-oss-20b (batch)openai:gpt-oss-20b:batch | openai | $0.024 | $0.112 | 131k | toolsjson | hosted |
| Mistral: Ministral 3 8B 2512 (batch)mistralai:ministral-8b-2512:batch | mistralai | $0.075 | $0.075 | 262k | toolsvisionjsonlongContext | hosted |
| Google: Gemma 3 4Bgoogle:gemma-3-4b-it | $0.05 | $0.1 | 131k | toolsvisionjson | hosted | |
| Qwen: Qwen3.7 Flashqwen:qwen3.7-flash | qwen | $0.03 | $0.13 | 1000k | toolsvisionjsonlongContext | hosted |
| inclusionAI: Ling 3.0 Flash Santeinclusionai:ling-3.0-flash-sante | inclusionai | $0.042 | $0.1232 | 262k | toolsjsonlongContext | hosted |
| inclusionAI: Ling 3.0 Flash Fininclusionai:ling-3.0-flash-fin | inclusionai | $0.042 | $0.1232 | 262k | toolsjsonlongContext | hosted |
| OpenAI: gpt-oss-120b (batch)openai:gpt-oss-120b:batch | openai | $0.0296 | $0.136 | 131k | toolsjson | hosted |
| DeepSeek: DeepSeek Flash Latest~deepseek:deepseek-flash-latest | ~deepseek | $0.0348 | $0.1392 | 1049k | toolsvisionjsonlongContext | hosted |
| Amazon: Nova Micro 1.0amazon:nova-micro-v1 | amazon | $0.035 | $0.14 | 128k | toolsjson | hosted |
| Inference.net: Schematron V2 Turboinference-net:schematron-v2-turbo | inference-net | $0.03 | $0.15 | 128k | toolsjson | hosted |
| Poolside: Laguna XS 2.1poolside:laguna-xs-2.1 | poolside | $0.06 | $0.12 | 262k | toolsjsonlongContext | hosted |
| Cohere: Command R7B (12-2024)cohere:command-r7b-12-2024 | Cohere | $0.0375 | $0.15 | 128k | toolsjson | open |
| Inception: Mercury 2.5inception:mercury-2.5 | Inception | $0.04 | $0.15 | 260k | toolsreasoningjsonlongContext | hosted |
| MythoMax 13Bgryphe:mythomax-l2-13b | gryphe | $0.08 | $0.11 | 8k | toolsjson | hosted |
| Reka Edgerekaai:reka-edge | rekaai | $0.1 | $0.1 | 16k | toolsvisionjson | hosted |
| Mistral: Ministral 3 3B 2512mistralai:ministral-3b-2512 | mistralai | $0.1 | $0.1 | 131k | toolsvisionjson | hosted |
| Google: Gemma 3 12Bgoogle:gemma-3-12b-it | $0.05 | $0.15 | 131k | toolsvisionjson | hosted | |
| OpenAI: gpt-oss-120bopenai:gpt-oss-120b | openai | $0.037 | $0.17 | 131k | toolsjson | hosted |
| Microsoft: Phi 4microsoft:phi-4 | microsoft | $0.07 | $0.14 | 16k | toolsjson | hosted |
| Tencent: Hy-MT2-1.8Btencent:hy-mt2-1.8b | tencent | $0.044 | $0.177 | 8k | toolsjson | hosted |
| OpenAI: GPT-5 Nano (batch)openai:gpt-5-nano:batch | openai | $0.025 | $0.2 | 400k | toolsvisionjsonlongContext | hosted |
| Meta: Llama 3.2 1B Instructmeta-llama:llama-3.2-1b-instruct | meta-llama | $0.027 | $0.201 | 60k | toolsjson | hosted |
| Upstage: Solar Mini 4upstage:solar-mini4 | upstage | $0.05 | $0.2 | 524k | toolsjsonlongContext | hosted |
| Qwen: Qwen3.5-9Bqwen:qwen3.5-9b | qwen | $0.1 | $0.15 | 262k | toolsvisionjsonlongContext | hosted |
| Google: Gemini 2.5 Flash Lite (batch)google:gemini-2.5-flash-lite:batch | $0.05 | $0.2 | 1049k | toolsvisionjsonlongContext | hosted | |
| OpenAI: GPT-4.1 Nano (batch)openai:gpt-4.1-nano:batch | openai | $0.05 | $0.2 | 1048k | toolsvisionjsonlongContext | hosted |
| Z.ai: GLM 5.3 Flash (batch)z-ai:glm-5.3-flash:batch | z-ai | $0.06 | $0.2 | 1049k | toolsvisionjsonlongContext | hosted |
| NVIDIA: Nemotron 3.5 Lightningnvidia:nemotron-3.5-lightning | nvidia | $0.07 | $0.2 | 262k | toolsjsonlongContext | hosted |
| Poolside: Laguna S 2.1poolside:laguna-s-2.1 | poolside | $0.09 | $0.18 | 1049k | toolsjsonlongContext | hosted |
| Inference.net: Schematron V2 Smallinference-net:schematron-v2-small | inference-net | $0.05 | $0.23 | 128k | toolsjson | hosted |
| Google: Gemma 4 26B A4B google:gemma-4-26b-a4b-it | $0.0675 | $0.225 | 262k | toolsreasoningvisionjson | open | |
| Anthropic: Claude Haiku 5.5 (batch)anthropic:claude-haiku-5.5:batch | anthropic | $0.05 | $0.25 | 1000k | toolsvisionjsonlongContext | hosted |
| OpenAI: GPT-6 Luna Pro (batch)openai:gpt-6-luna-pro:batch | openai | $0.05 | $0.25 | 1050k | toolsvisionjsonlongContext | hosted |
| OpenAI: GPT-6 Luna (batch)openai:gpt-6-luna:batch | openai | $0.05 | $0.25 | 1050k | toolsvisionjsonlongContext | hosted |
| NVIDIA: Nemotron 3 Nano 30B A3Bnvidia:nemotron-3-nano-30b-a3b | nvidia | $0.06 | $0.24 | 262k | toolsjsonlongContext | hosted |
| Mistral: Ministral 3 8B 2512mistralai:ministral-8b-2512 | mistralai | $0.15 | $0.15 | 262k | toolsvisionjsonlongContext | hosted |
| Amazon: Nova Lite 1.0amazon:nova-lite-v1 | amazon | $0.06 | $0.24 | 300k | toolsvisionjsonlongContext | hosted |
| Meta: Muse Spark 1.2 Contributormeta:muse-spark-1.2-contributor | Meta | $0.1 | $0.2 | 1049k | toolsreasoningvisionjson | hosted |
| Meta: Muse Spark 1.3 Contributormeta:muse-spark-1.3-contributor | Meta | $0.1 | $0.2 | 1049k | toolsreasoningvisionjson | hosted |
| ByteDance: UI-TARS 7B bytedance:ui-tars-1.5-7b | bytedance | $0.1 | $0.2 | 128k | toolsvisionjson | hosted |
| Reka Flash 3rekaai:reka-flash-3 | rekaai | $0.1 | $0.2 | 66k | toolsjson | hosted |
| Qwen: Qwen2.5 7B Instructqwen:qwen-2.5-7b-instruct | qwen | $0.1 | $0.2 | 33k | toolsjson | hosted |
| IBM: Granite 4.2 8Bibm-granite:granite-4.2-8b | ibm-granite | $0.06 | $0.25 | 131k | toolsjson | hosted |
| Nex AGI: Nex-N2.5-Pronex-agi:nex-n2.5-pro | nex-agi | $0.075 | $0.25 | 262k | toolsvisionjsonlongContext | hosted |
| Qwen: Qwen3.5-Flashqwen:qwen3.5-flash-02-23 | qwen | $0.065 | $0.26 | 1000k | toolsvisionjsonlongContext | hosted |
| Mistral: Mistral Small 3.2 24Bmistralai:mistral-small-3.2-24b-instruct | mistralai | $0.09375 | $0.25 | 256k | toolsvisionjsonlongContext | hosted |
| Qwen: Qwen3 Coder 30B A3B Instructqwen:qwen3-coder-30b-a3b-instruct | qwen | $0.07 | $0.28 | 262k | toolsjsonlongContext | hosted |
| Meta: Llama Guard 4 12Bmeta-llama:llama-guard-4-12b | meta-llama | $0.18 | $0.18 | 164k | toolsvisionjson | hosted |
| Qwen: Qwen3 14Bqwen:qwen3-14b | qwen | $0.12 | $0.24 | 41k | toolsjson | hosted |
| Qwen: Qwen3 32Bqwen:qwen3-32b | qwen | $0.08 | $0.28 | 131k | toolsjson | hosted |
| Tencent: Hy-MT2-30B-A3Btencent:hy-mt2-30b-a3b | tencent | $0.074 | $0.295 | 8k | toolsjson | hosted |
Which lane does a model belong on?
fast
Target reference request: $0.0008 · context ≥ 32k.
balanced
Target reference request: $0.0008 · context ≥ 128k.
deep
Target reference request: $0.0008 · context ≥ 200k.
Patch a model onto a lane from any agent page; the compiler prices it there immediately.