Skip to content
Patchbayagent routing bay

Live model catalog

Every model below is normalized into one shape: price per million input and output tokens, context window, and capabilities. Upstream schemas never reach the UI directly.

Live· 432458 upstream records · 57 providers

Capability coverage

  • tools

    432

    of 432 models

  • reasoning

    77

    of 432 models

  • vision

    284

    of 432 models

  • json

    432

    of 432 models

  • longContext

    312

    of 432 models

Cheapest first

first 60 of 432

ModelProviderIn /MTokOut /MTokContextCapabilitiesWeights
Mistral: Mistral Nemomistralai:mistral-nemomistralai$0.029$0.03131ktoolsjsonhosted
inclusionAI: Ling 3.0 Flash VLinclusionai:ling-3.0-flash-vlinclusionai$0.021$0.0616262ktoolsvisionjsonlongContexthosted
inclusionAI: Ling 3.0 Flashinclusionai:ling-3.0-flashinclusionai$0.021$0.063262ktoolsjsonlongContexthosted
Sao10K: Llama 3 8B Lunarissao10k:l3-lunaris-8bsao10k$0.04$0.058ktoolsjsonhosted
OpenAI: gpt-oss-20bopenai:gpt-oss-20bopenai$0.018$0.09131ktoolsjsonhosted
Nex AGI: Nex-N2.5-Mininex-agi:nex-n2.5-mininex-agi$0.025$0.1262ktoolsvisionjsonlongContexthosted
IBM: Granite 4.0 Microibm-granite:granite-4.0-h-microibm-granite$0.017$0.112131ktoolsjsonhosted
Mistral: Mistral Small 3mistralai:mistral-small-24b-instruct-2501mistralai$0.05$0.0833ktoolsjsonhosted
Meta: Llama 3.1 8B Instructmeta-llama:llama-3.1-8b-instructmeta-llama$0.05$0.08131ktoolsjsonhosted
OpenAI: gpt-oss-20b (batch)openai:gpt-oss-20b:batchopenai$0.024$0.112131ktoolsjsonhosted
Mistral: Ministral 3 8B 2512 (batch)mistralai:ministral-8b-2512:batchmistralai$0.075$0.075262ktoolsvisionjsonlongContexthosted
Google: Gemma 3 4Bgoogle:gemma-3-4b-itgoogle$0.05$0.1131ktoolsvisionjsonhosted
Qwen: Qwen3.7 Flashqwen:qwen3.7-flashqwen$0.03$0.131000ktoolsvisionjsonlongContexthosted
inclusionAI: Ling 3.0 Flash Santeinclusionai:ling-3.0-flash-santeinclusionai$0.042$0.1232262ktoolsjsonlongContexthosted
inclusionAI: Ling 3.0 Flash Fininclusionai:ling-3.0-flash-fininclusionai$0.042$0.1232262ktoolsjsonlongContexthosted
OpenAI: gpt-oss-120b (batch)openai:gpt-oss-120b:batchopenai$0.0296$0.136131ktoolsjsonhosted
DeepSeek: DeepSeek Flash Latest~deepseek:deepseek-flash-latest~deepseek$0.0348$0.13921049ktoolsvisionjsonlongContexthosted
Amazon: Nova Micro 1.0amazon:nova-micro-v1amazon$0.035$0.14128ktoolsjsonhosted
Inference.net: Schematron V2 Turboinference-net:schematron-v2-turboinference-net$0.03$0.15128ktoolsjsonhosted
Poolside: Laguna XS 2.1poolside:laguna-xs-2.1poolside$0.06$0.12262ktoolsjsonlongContexthosted
Cohere: Command R7B (12-2024)cohere:command-r7b-12-2024Cohere$0.0375$0.15128ktoolsjsonopen
Inception: Mercury 2.5inception:mercury-2.5Inception$0.04$0.15260ktoolsreasoningjsonlongContexthosted
MythoMax 13Bgryphe:mythomax-l2-13bgryphe$0.08$0.118ktoolsjsonhosted
Reka Edgerekaai:reka-edgerekaai$0.1$0.116ktoolsvisionjsonhosted
Mistral: Ministral 3 3B 2512mistralai:ministral-3b-2512mistralai$0.1$0.1131ktoolsvisionjsonhosted
Google: Gemma 3 12Bgoogle:gemma-3-12b-itgoogle$0.05$0.15131ktoolsvisionjsonhosted
OpenAI: gpt-oss-120bopenai:gpt-oss-120bopenai$0.037$0.17131ktoolsjsonhosted
Microsoft: Phi 4microsoft:phi-4microsoft$0.07$0.1416ktoolsjsonhosted
Tencent: Hy-MT2-1.8Btencent:hy-mt2-1.8btencent$0.044$0.1778ktoolsjsonhosted
OpenAI: GPT-5 Nano (batch)openai:gpt-5-nano:batchopenai$0.025$0.2400ktoolsvisionjsonlongContexthosted
Meta: Llama 3.2 1B Instructmeta-llama:llama-3.2-1b-instructmeta-llama$0.027$0.20160ktoolsjsonhosted
Upstage: Solar Mini 4upstage:solar-mini4upstage$0.05$0.2524ktoolsjsonlongContexthosted
Qwen: Qwen3.5-9Bqwen:qwen3.5-9bqwen$0.1$0.15262ktoolsvisionjsonlongContexthosted
Google: Gemini 2.5 Flash Lite (batch)google:gemini-2.5-flash-lite:batchgoogle$0.05$0.21049ktoolsvisionjsonlongContexthosted
OpenAI: GPT-4.1 Nano (batch)openai:gpt-4.1-nano:batchopenai$0.05$0.21048ktoolsvisionjsonlongContexthosted
Z.ai: GLM 5.3 Flash (batch)z-ai:glm-5.3-flash:batchz-ai$0.06$0.21049ktoolsvisionjsonlongContexthosted
NVIDIA: Nemotron 3.5 Lightningnvidia:nemotron-3.5-lightningnvidia$0.07$0.2262ktoolsjsonlongContexthosted
Poolside: Laguna S 2.1poolside:laguna-s-2.1poolside$0.09$0.181049ktoolsjsonlongContexthosted
Inference.net: Schematron V2 Smallinference-net:schematron-v2-smallinference-net$0.05$0.23128ktoolsjsonhosted
Google: Gemma 4 26B A4B google:gemma-4-26b-a4b-itGoogle$0.0675$0.225262ktoolsreasoningvisionjsonopen
Anthropic: Claude Haiku 5.5 (batch)anthropic:claude-haiku-5.5:batchanthropic$0.05$0.251000ktoolsvisionjsonlongContexthosted
OpenAI: GPT-6 Luna Pro (batch)openai:gpt-6-luna-pro:batchopenai$0.05$0.251050ktoolsvisionjsonlongContexthosted
OpenAI: GPT-6 Luna (batch)openai:gpt-6-luna:batchopenai$0.05$0.251050ktoolsvisionjsonlongContexthosted
NVIDIA: Nemotron 3 Nano 30B A3Bnvidia:nemotron-3-nano-30b-a3bnvidia$0.06$0.24262ktoolsjsonlongContexthosted
Mistral: Ministral 3 8B 2512mistralai:ministral-8b-2512mistralai$0.15$0.15262ktoolsvisionjsonlongContexthosted
Amazon: Nova Lite 1.0amazon:nova-lite-v1amazon$0.06$0.24300ktoolsvisionjsonlongContexthosted
Meta: Muse Spark 1.2 Contributormeta:muse-spark-1.2-contributorMeta$0.1$0.21049ktoolsreasoningvisionjsonhosted
Meta: Muse Spark 1.3 Contributormeta:muse-spark-1.3-contributorMeta$0.1$0.21049ktoolsreasoningvisionjsonhosted
ByteDance: UI-TARS 7B bytedance:ui-tars-1.5-7bbytedance$0.1$0.2128ktoolsvisionjsonhosted
Reka Flash 3rekaai:reka-flash-3rekaai$0.1$0.266ktoolsjsonhosted
Qwen: Qwen2.5 7B Instructqwen:qwen-2.5-7b-instructqwen$0.1$0.233ktoolsjsonhosted
IBM: Granite 4.2 8Bibm-granite:granite-4.2-8bibm-granite$0.06$0.25131ktoolsjsonhosted
Nex AGI: Nex-N2.5-Pronex-agi:nex-n2.5-pronex-agi$0.075$0.25262ktoolsvisionjsonlongContexthosted
Qwen: Qwen3.5-Flashqwen:qwen3.5-flash-02-23qwen$0.065$0.261000ktoolsvisionjsonlongContexthosted
Mistral: Mistral Small 3.2 24Bmistralai:mistral-small-3.2-24b-instructmistralai$0.09375$0.25256ktoolsvisionjsonlongContexthosted
Qwen: Qwen3 Coder 30B A3B Instructqwen:qwen3-coder-30b-a3b-instructqwen$0.07$0.28262ktoolsjsonlongContexthosted
Meta: Llama Guard 4 12Bmeta-llama:llama-guard-4-12bmeta-llama$0.18$0.18164ktoolsvisionjsonhosted
Qwen: Qwen3 14Bqwen:qwen3-14bqwen$0.12$0.2441ktoolsjsonhosted
Qwen: Qwen3 32Bqwen:qwen3-32bqwen$0.08$0.28131ktoolsjsonhosted
Tencent: Hy-MT2-30B-A3Btencent:hy-mt2-30b-a3btencent$0.074$0.2958ktoolsjsonhosted

Which lane does a model belong on?

  • fast

    Target reference request: $0.0008 · context ≥ 32k.

  • balanced

    Target reference request: $0.0008 · context ≥ 128k.

  • deep

    Target reference request: $0.0008 · context ≥ 200k.

Patch a model onto a lane from any agent page; the compiler prices it there immediately.