Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
OpenRouter
Provider ID: openrouter
Details
SDK package
@openrouter/ai-sdk-provider
Models
344
API
https://openrouter.ai/api/v1
Links
Documentation
Auth: OPENROUTER_API_KEY
Models (344)
Jamba Large 1.7
Flagship model for demanding analysis, coding, and production agent workflows
256,000 ctx
Aion-2.0
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
131,072 ctx
reasoning
Aion-3.0
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
131,072 ctx
reasoning
Aion-3.0-Mini
Efficient model for low-latency assistance, extraction, and routine automation
131,072 ctx
reasoning
Aion-RP 1.0 (8B)
Open Llama instruction model for multilingual chat, reasoning, and coding
32,768 ctx
Olmo 3 32B Think
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
65,536 ctx
reasoning
Nova 2 Lite
Multimodal reasoning model for visual analysis, planning, and tool use
1,000,000 ctx
reasoning
Nova Lite 1.0
Efficient model for low-latency assistance, extraction, and routine automation
300,000 ctx
Nova Micro 1.0
Efficient model for low-latency assistance, extraction, and routine automation
128,000 ctx
Nova Premier 1.0
Flagship model for demanding analysis, coding, and production agent workflows
1,000,000 ctx
Nova Pro 1.0
Flagship model for demanding analysis, coding, and production agent workflows
300,000 ctx
Magnum v4 72B
Open-weight instruction model for adaptable chat and self-hosted production workloads
16,384 ctx
Claude 3 Haiku
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
Claude Fable 5
Claude model for creative writing, analysis, and controlled agent workflows
1,000,000 ctx
reasoning
Claude Haiku 4.5 (latest)
Fast Claude lane for lightweight agents, office tasks, and responsive chat
200,000 ctx
reasoning
Claude Opus 4
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.1 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
High-end Claude for difficult coding, planning, and slower expert reasoning
1,000,000 ctx
reasoning
Claude Opus 4.7
Stronger Opus tier for advanced software work and high-stakes reasoning
1,000,000 ctx
reasoning
Claude Opus 4.7 (Fast)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.8 (Fast)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 5 (Fast)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Sonnet 4
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 4.5 (latest)
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
Claude Sonnet 5
Everyday Claude agent model for coding, planning, browsing, and general work
1,000,000 ctx
reasoning
Trinity Large Thinking
Flagship model for demanding analysis, coding, and production agent workflows
262,144 ctx
reasoning
Virtuoso Large
Flagship model for demanding analysis, coding, and production agent workflows
131,072 ctx
ERNIE 4.5 VL 424B A47B
Multimodal reasoning model for visual analysis, planning, and tool use
123,000 ctx
reasoning
Seed 1.6
Multimodal reasoning model for visual analysis, planning, and tool use
262,144 ctx
reasoning
Seed 1.6 Flash
Multimodal reasoning model for visual analysis, planning, and tool use
262,144 ctx
reasoning
Seed-2.0-Lite
Multimodal reasoning model for visual analysis, planning, and tool use
262,144 ctx
reasoning
Seed-2.0-Mini
Multimodal reasoning model for visual analysis, planning, and tool use
262,144 ctx
reasoning
UI-TARS 7B
Multimodal model for analyzing text, images, documents, and rich media
128,000 ctx
Uncensored
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
128,000 ctx
Command A
Cohere command model for multilingual enterprise agents, tools, and chat
256,000 ctx
Command R
Cohere retrieval model for long-context chat and enterprise RAG workflows
128,000 ctx
Command R+
Cohere's RAG workhorse for long-context enterprise search and tool use
128,000 ctx
Command R7B
Cohere retrieval model for long-context chat and enterprise RAG workflows
128,000 ctx
North Mini Code (free)
Cohere coding model for practical software engineering and agentic edits
256,000 ctx
reasoning
Cogito v2.1 671B
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
128,000 ctx
reasoning
DeepSeek Chat
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
DeepSeek V3 0324
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
DeepSeek V3.1
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek-R1
Classic open reasoning model for transparent math, coding, and deliberate problem solving
163,840 ctx
reasoning
R1 0528
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
163,840 ctx
reasoning
R1 Distill Llama 70B
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
8,192 ctx
reasoning
DeepSeek V3.1 Terminus
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek V3.2
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek V3.2 Exp
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,048,576 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,048,576 ctx
reasoning
Gemini 2.5 Flash
Fast Gemini workhorse for multimodal apps where latency and price matter
1,048,576 ctx
reasoning
Nano Banana
Nano Banana image model for fast generation, edits, and character-consistent assets
32,768 ctx
Gemini 2.5 Flash-Lite
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
1,048,576 ctx
reasoning
Gemini 2.5 Pro
Google's proven reasoning model for coding, math, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 2.5 Pro Preview 06-05
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 2.5 Pro Preview 05-06
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3 Flash Preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
1,048,576 ctx
reasoning
Nano Banana Pro
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
131,072 ctx
reasoning
Nano Banana Pro
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
65,536 ctx
reasoning
Nano Banana 2
Image model for prompt-driven generation, editing, and visual design workflows
131,072 ctx
reasoning
Nano Banana 2
Image model for prompt-driven generation, editing, and visual design workflows
131,072 ctx
reasoning
Gemini 3.1 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Nano Banana 2 Lite
Image model for prompt-driven generation, editing, and visual design workflows
65,536 ctx
reasoning
Gemini 3.1 Flash Lite Preview
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview Custom Tools
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemini 3.5 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.6 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemma 2 27B
Open Gemma instruction model for efficient chat and self-hosted deployments
8,192 ctx
Gemma 3 12B
Open Gemma instruction model for efficient chat and self-hosted deployments
131,072 ctx
Gemma 3 27B
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
Gemma 3 4B
Open Gemma instruction model for efficient chat and self-hosted deployments
131,072 ctx
Gemma 3n 4B
Open Gemma instruction model for efficient chat and self-hosted deployments
32,768 ctx
Gemma 4 26B A4B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 26B A4B (free)
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
Gemma 4 31B (free)
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
Lyria 3 Clip Preview
Speech generation model for controllable voice, narration, and audio delivery
1,048,576 ctx
Lyria 3 Pro Preview
Speech generation model for controllable voice, narration, and audio delivery
1,048,576 ctx
MythoMax 13B
Open-weight instruction model for adaptable chat and self-hosted production workloads
8,192 ctx
Granite 4.0 Micro
Efficient model for low-latency assistance, extraction, and routine automation
131,000 ctx
Granite 4.1 8B
Open-weight instruction model for adaptable chat and self-hosted production workloads
131,072 ctx
Mercury 2
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
128,000 ctx
reasoning
Ling-2.6-1T
Tool-capable chat model for instruction following and agentic application workflows
262,144 ctx
Ling-2.6-flash
Efficient model for low-latency assistance, extraction, and routine automation
262,144 ctx
Ling-3.0-flash (free)
Efficient model for low-latency assistance, extraction, and routine automation
262,144 ctx
reasoning
Ring-2.6-1T
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
262,144 ctx
reasoning
Inflection 3 Pi
General-purpose chat model for instruction following, writing, and analysis
8,000 ctx
Inflection 3 Productivity
General-purpose chat model for instruction following, writing, and analysis
8,000 ctx
KAT-Coder-Air V2.5
Coding model for repository understanding, refactors, and agentic engineering tasks
256,000 ctx
KAT-Coder-Pro V2
Coding model for repository understanding, refactors, and agentic engineering tasks
262,144 ctx
KAT-Coder-Pro V2.5
Coding model for repository understanding, refactors, and agentic engineering tasks
256,000 ctx
Weaver (alpha)
General-purpose chat model for instruction following, writing, and analysis
8,000 ctx
LongCat 2.0
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
1,048,756 ctx
reasoning
Llama 3.1 70B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Llama 3.1 8B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Llama 3.2 1B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
60,000 ctx
Llama 3.2 3B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Llama-3.3-70B-Instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
131,072 ctx
Llama 4 Maverick
Open multimodal Llama model for strong reasoning and fast responses
1,048,576 ctx
Llama 4 Scout
Open multimodal Llama model for long-context analysis and efficient agents
1,310,720 ctx
Llama Guard 4 12B
Safety model for policy screening, moderation, and risk-aware routing workflows
1,048,576 ctx
Muse Spark 1.1
Open Llama multimodal model for image understanding and text reasoning
1,048,576 ctx
reasoning
Phi 4
Open-weight instruction model for adaptable chat and self-hosted production workloads
16,384 ctx
WizardLM-2 8x22B
Open-weight instruction model for adaptable chat and self-hosted production workloads
65,535 ctx
MiniMax-01
MiniMax multimodal coding model for long-context reasoning and agent tasks
1,000,192 ctx
MiniMax M1
MiniMax model for chat, coding, office work, and agentic tasks
1,000,000 ctx
reasoning
MiniMax-M2
Efficient open MiniMax model built for coding agents and tool-heavy workflows
204,800 ctx
reasoning
MiniMax M2-her
MiniMax model for chat, coding, office work, and agentic tasks
65,536 ctx
MiniMax-M2.1
Earlier MiniMax agent model for practical coding and productivity tasks
204,800 ctx
reasoning
MiniMax-M2.5
Prior MiniMax coding model for agent workflows, office edits, and automation
204,800 ctx
reasoning
MiniMax-M2.7
Open MiniMax flagship for coding agents, office automation, and complex environments
204,800 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
1,048,576 ctx
reasoning
Codestral 2508
Mistral coding model for code completion, generation, and developer workflows
256,000 ctx
Devstral 2
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
262,144 ctx
Ministral 3 14B 2512
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
262,144 ctx
Ministral 3 3B 2512
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
131,072 ctx
Ministral 3 8B 2512
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
262,144 ctx
Mistral Large
Flagship Mistral model for advanced reasoning, coding, and multilingual work
128,000 ctx
Mistral Large 2407
Flagship Mistral model for advanced reasoning, coding, and multilingual work
131,072 ctx
Mistral Large 3
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
262,144 ctx
Mistral Medium 3
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
131,072 ctx
Mistral Medium 3.5
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
262,144 ctx
reasoning
Mistral Medium 3.1
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
131,072 ctx
Mistral Nemo
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
131,072 ctx
Saba
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
32,768 ctx
Mistral Small 3
Efficient Mistral model for fast chat, extraction, and production assistants
32,768 ctx
Mistral Small 4
Fast Mistral production model for chat, extraction, and cost-sensitive agents
262,144 ctx
reasoning
Mistral Small 3.1 24B
Efficient Mistral model for fast chat, extraction, and production assistants
128,000 ctx
Mistral Small 3.2 24B
Efficient Mistral model for fast chat, extraction, and production assistants
256,000 ctx
Mixtral 8x22B Instruct
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
65,536 ctx
Voxtral Small 24B 2507
Efficient Mistral model for fast chat, extraction, and production assistants
32,000 ctx
Kimi K2 0711
Kimi model for long-context chat, coding, and agentic reasoning
131,072 ctx
Kimi K2 0905
Kimi model for long-context chat, coding, and agentic reasoning
262,144 ctx
Kimi K2 Thinking
Thinking Kimi model for slower research passes, planning, and hard technical questions
262,144 ctx
reasoning
Kimi K2.5
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
262,144 ctx
reasoning
Kimi K2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
262,144 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Kimi K3
Kimi multimodal agent model for visual understanding, coding, and planning
1,048,576 ctx
reasoning
Morph V3 Fast
Efficient model for low-latency assistance, extraction, and routine automation
81,920 ctx
Morph V3 Large
Flagship model for demanding analysis, coding, and production agent workflows
262,144 ctx
Nex-N2-Mini
Multimodal reasoning model for visual analysis, planning, and tool use
262,144 ctx
reasoning
Nex-N2-Pro
Multimodal reasoning model for visual analysis, planning, and tool use
262,144 ctx
reasoning
Hermes 3 405B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Hermes 3 70B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Hermes 4 405B
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
131,072 ctx
reasoning
Hermes 4 70B
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
131,072 ctx
reasoning
Nemotron 3 Nano 30B A3B
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
262,144 ctx
reasoning
Nemotron 3 Nano 30B A3B (free)
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
256,000 ctx
reasoning
Nemotron 3 Nano Omni (free)
Open Nemotron omni model combining reasoning with text, vision, and audio
256,000 ctx
reasoning
Nemotron 3 Super 120B A12B
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
1,000,000 ctx
reasoning
Nemotron 3 Super (free)
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
262,144 ctx
reasoning
Nemotron 3 Ultra 550B A55B
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
512,288 ctx
reasoning
Nemotron 3 Ultra (free)
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
1,000,000 ctx
reasoning
Nemotron 3.5 Content Safety (free)
Safety model for policy screening, moderation, and risk-aware routing workflows
128,000 ctx
reasoning
Nemotron Nano 12B 2 VL (free)
Nemotron multimodal model for visual reasoning and agentic AI workflows
128,000 ctx
reasoning
Nemotron Nano 9B V2 (free)
Compact Nemotron model for efficient reasoning and deployable AI agents
128,000 ctx
reasoning
GPT-3.5-turbo
Compact GPT model for low-latency assistance and high-volume workloads
16,385 ctx
GPT-3.5 Turbo (older v0613)
Compact GPT model for low-latency assistance and high-volume workloads
4,095 ctx
GPT-3.5 Turbo 16k
Compact GPT model for low-latency assistance and high-volume workloads
16,385 ctx
GPT-3.5 Turbo Instruct
Compact GPT model for low-latency assistance and high-volume workloads
4,095 ctx
GPT-4
GPT model for general reasoning, writing, coding, and tool-assisted tasks
8,191 ctx
GPT-4 Turbo
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4 Turbo Preview
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
GPT-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
128,000 ctx
GPT-4o (2024-05-13)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o (2024-08-06)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o (2024-11-20)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
128,000 ctx
GPT-4o-mini (2024-07-18)
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4o-mini Search Preview
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4o Search Preview
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-5
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
400,000 ctx
reasoning
GPT-5 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5-Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5 Image
Image model for prompt-driven generation, editing, and visual design workflows
400,000 ctx
reasoning
GPT-5 Image Mini
Image model for prompt-driven generation, editing, and visual design workflows
400,000 ctx
reasoning
GPT-5 Mini
Small GPT-5 for responsive agents, coding help, and everyday automation
400,000 ctx
reasoning
GPT-5 Nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
400,000 ctx
reasoning
GPT-5 Pro
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
400,000 ctx
reasoning
GPT-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.1 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.1 Codex
Codex GPT for repository edits, code review, and practical software agents
400,000 ctx
reasoning
GPT-5.1 Codex Max
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.1 Codex mini
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
400,000 ctx
reasoning
GPT-5.2 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.2 Codex
Code-specialist GPT for repository edits, reviews, and long-running software agents
400,000 ctx
reasoning
GPT-5.2 Pro
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
400,000 ctx
reasoning
GPT-5.3 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
1,050,000 ctx
reasoning
GPT-5.4 Image 2
Image model for prompt-driven generation, editing, and visual design workflows
272,000 ctx
reasoning
GPT-5.4 mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
400,000 ctx
reasoning
GPT-5.4 Pro
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
1,050,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
GPT-5.5 Pro
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
1,050,000 ctx
reasoning
GPT-5.6 Luna
GPT model for general reasoning, writing, coding, and tool-assisted tasks
1,050,000 ctx
reasoning
GPT-5.6 Luna Pro
Frontier GPT model for professional reasoning, coding, and multimodal work
1,050,000 ctx
reasoning
GPT-5.6 Sol
GPT model for general reasoning, writing, coding, and tool-assisted tasks
1,050,000 ctx
reasoning
GPT-5.6 Sol Pro
Frontier GPT model for professional reasoning, coding, and multimodal work
1,050,000 ctx
reasoning
GPT-5.6 Terra
GPT model for general reasoning, writing, coding, and tool-assisted tasks
1,050,000 ctx
reasoning
GPT-5.6 Terra Pro
Frontier GPT model for professional reasoning, coding, and multimodal work
1,050,000 ctx
reasoning
GPT Audio
Speech generation model for controllable voice, narration, and audio delivery
128,000 ctx
GPT Audio Mini
Speech generation model for controllable voice, narration, and audio delivery
128,000 ctx
GPT Chat Latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
400,000 ctx
GPT OSS 120B
Open GPT reasoning model for self-hosted agents and controllable deployments
131,072 ctx
reasoning
GPT OSS 20B
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
gpt-oss-20b (free)
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
gpt-oss-safeguard-20b
Safety model for policy screening, moderation, and risk-aware routing workflows
131,072 ctx
reasoning
o1
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o1-pro
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o3
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
200,000 ctx
reasoning
o3-deep-research
Research model for long-horizon investigation, synthesis, and analytical reports
200,000 ctx
reasoning
o3-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
200,000 ctx
reasoning
o3 Mini High
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o3-pro
High-effort o3 tier for difficult technical reasoning and careful answers
200,000 ctx
reasoning
o4-mini
Fast o-series model for compact reasoning, coding, and tool use
200,000 ctx
reasoning
o4-mini-deep-research
Research model for long-horizon investigation, synthesis, and analytical reports
200,000 ctx
reasoning
o4 Mini High
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
Auto Router
Image model for prompt-driven generation, editing, and visual design workflows
2,000,000 ctx
reasoning
Body Builder (beta)
Preview model for early access evaluation, prototyping, and compatibility testing
128,000 ctx
Free Models Router
Multimodal reasoning model for visual analysis, planning, and tool use
200,000 ctx
reasoning
Fusion
General-purpose chat model for instruction following, writing, and analysis
1,000,000 ctx
Pareto Code Router
Coding model for repository understanding, refactors, and agentic engineering tasks
2,000,000 ctx
Perceptron Mk1
Multimodal reasoning model for visual analysis, planning, and tool use
32,768 ctx
reasoning
Sonar
Sonar search model for current answers, retrieval, and citation-backed chat
127,072 ctx
Sonar Deep Research
Sonar search model for current answers, retrieval, and citation-backed chat
128,000 ctx
reasoning
Sonar Pro
Advanced Sonar search model for deeper research and cited synthesis
200,000 ctx
Sonar Pro Search
Advanced Sonar search model for deeper research and cited synthesis
200,000 ctx
reasoning
Sonar Reasoning Pro
Web-grounded reasoning model for multi-step research and cited answers
128,000 ctx
reasoning
Laguna M.1
Poolside's open-weight model for agentic coding and long-horizon work
262,144 ctx
reasoning
Laguna M.1 (free)
Free provider route for experiments, demos, and cost-sensitive chat workloads
262,144 ctx
reasoning
Laguna S 2.1
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
1,048,576 ctx
reasoning
Laguna S 2.1 (free)
Free provider route for experiments, demos, and cost-sensitive chat workloads
262,144 ctx
reasoning
Laguna XS 2.1
Agentic coding model from Poolside in the XS size class for local deployment
262,144 ctx
reasoning
Laguna XS 2.1 (free)
Free provider route for experiments, demos, and cost-sensitive chat workloads
262,144 ctx
reasoning
Qwen2.5 72B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
32,768 ctx
Qwen2.5 7B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
32,768 ctx
Qwen2.5 Coder 32B Instruct
Qwen coding model for software agents, repository edits, and code reasoning
32,768 ctx
Qwen Plus
Qwen instruction model for multilingual chat, reasoning, and tool use
1,000,000 ctx
Qwen Plus 0728
Qwen instruction model for multilingual chat, reasoning, and tool use
1,000,000 ctx
Qwen Plus 0728 (thinking)
Qwen reasoning model for deliberate problem solving, math, and coding
1,000,000 ctx
reasoning
Qwen2.5 VL 72B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
128,000 ctx
Qwen3 14B
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
reasoning
Qwen3 235B-A22B
Large open Qwen MoE for multilingual reasoning, coding, and tool use
131,072 ctx
reasoning
Qwen3 235B A22B Instruct 2507
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
Qwen3 235B A22B Thinking 2507
Qwen reasoning model for deliberate problem solving, math, and coding
262,144 ctx
reasoning
Qwen3 30B A3B
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
reasoning
Qwen3 30B A3B Instruct 2507
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
Qwen3 30B A3B Thinking 2507
Qwen reasoning model for deliberate problem solving, math, and coding
81,920 ctx
reasoning
Qwen3 32B
Dense open Qwen model for self-hosted chat, reasoning, and coding
131,072 ctx
reasoning
Qwen3 8B
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
reasoning
Qwen3 Coder 480B A35B
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3-Coder 30B-A3B Instruct
Smaller Qwen coder for efficient local agents and repo-level fixes
262,144 ctx
Qwen3 Coder Flash
Qwen coding model for software agents, repository edits, and code reasoning
1,000,000 ctx
Qwen3 Coder Next
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3 Coder Plus
Hosted Qwen coder for software agents, repo edits, and long-context code
1,000,000 ctx
Qwen3 Max
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
262,144 ctx
Qwen3 Max Thinking
Qwen reasoning model for deliberate problem solving, math, and coding
262,144 ctx
reasoning
Qwen3-Next 80B-A3B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
Qwen3-Next 80B-A3B (Thinking)
Efficient Qwen thinking model for local reasoning, math, and coding agents
262,144 ctx
reasoning
Qwen3 VL 235B A22B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
Qwen3 VL 235B A22B Thinking
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
reasoning
Qwen3 VL 30B A3B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
Qwen3 VL 30B A3B Thinking
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3 VL 32B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
Qwen3 VL 8B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
Qwen3 VL 8B Thinking
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
reasoning
Qwen3.5 122B-A10B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 35B-A3B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 397B-A17B
Large open Qwen multimodal MoE for visual agents and long technical tasks
262,144 ctx
reasoning
Qwen3.5 9B
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
reasoning
Qwen3.5-Flash
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.5 Plus 2026-02-15
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.5 Plus 2026-04-20
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.6 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.6 35B-A3B
Open multimodal Qwen MoE for local agents that need vision, audio, and code
262,144 ctx
reasoning
Qwen3.6 Flash
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.6 Max Preview
Flagship Qwen model for complex reasoning, coding, and agentic workflows
262,144 ctx
reasoning
Qwen3.6 Plus
Earlier Qwen multimodal workhorse for million-token agent and document tasks
1,000,000 ctx
reasoning
Qwen3.7 Max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
1,000,000 ctx
reasoning
Qwen3.7 Plus
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
1,000,000 ctx
reasoning
Reka Edge
Multimodal model for analyzing text, images, documents, and rich media
16,384 ctx
Reka Flash 3
Efficient model for low-latency assistance, extraction, and routine automation
65,536 ctx
reasoning
Relace Apply 3
General-purpose chat model for instruction following, writing, and analysis
256,000 ctx
Relace Search
Tool-capable chat model for instruction following and agentic application workflows
256,000 ctx
Fugu Ultra
Quality-first multi-agent model for hard research, analysis, and competitions
1,000,000 ctx
reasoning
Llama 3 8B Lunaris
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
Llama 3.1 Euryale 70B v2.2
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Llama 3.3 Euryale 70B
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Step 3.5 Flash
StepFun flash lane for quick multimodal reasoning and coding assistance
262,144 ctx
reasoning
Step 3.7 Flash
Newer StepFun flash model for faster agents, coding, and multimodal prompts
262,144 ctx
reasoning
Hunyuan A13B Instruct
Tencent Hy reasoning model for coding, instruction following, and agent tasks
131,072 ctx
reasoning
Hy3
Tencent Hy reasoning model for coding, instruction following, and agent tasks
262,144 ctx
reasoning
Hy3 preview
Tencent Hy reasoning model for coding, instruction following, and agent tasks
262,144 ctx
reasoning
Cydonia 24B V4.1
Open-weight instruction model for adaptable chat and self-hosted production workloads
131,072 ctx
Rocinante 12B
Open-weight instruction model for adaptable chat and self-hosted production workloads
65,536 ctx
Skyfall 36B V2
Open-weight instruction model for adaptable chat and self-hosted production workloads
32,768 ctx
UnslopNemo 12B
Open-weight instruction model for adaptable chat and self-hosted production workloads
32,768 ctx
Inkling
Multimodal reasoning model for visual analysis, planning, and tool use
1,048,576 ctx
reasoning
ReMM SLERP 13B
Open-weight instruction model for adaptable chat and self-hosted production workloads
6,144 ctx
Solar Pro 3
Flagship model for demanding analysis, coding, and production agent workflows
128,000 ctx
reasoning
Palmyra X5
General-purpose chat model for instruction following, writing, and analysis
1,040,000 ctx
Grok 4.20
Grok model for agentic tool use, reasoning, coding, and live assistance
2,000,000 ctx
reasoning
Grok 4.20 Multi-Agent
Grok model for agentic tool use, reasoning, coding, and live assistance
2,000,000 ctx
reasoning
Grok 4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
1,000,000 ctx
reasoning
Grok 4.5
xAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk
500,000 ctx
reasoning
Grok Build 0.1
Fast Grok coding model tuned for agentic engineering and iterative edits
256,000 ctx
reasoning
MiMo-V2.5
Open MiMo model for multimodal coding agents and long-context automation
1,050,000 ctx
reasoning
MiMo-V2.5-Pro
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
1,050,000 ctx
reasoning
GLM-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
131,072 ctx
reasoning
GLM-4.5-Air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
131,072 ctx
reasoning
GLM-4.5V
GLM vision model for visual reasoning, documents, and multimodal agents
65,536 ctx
reasoning
GLM-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
204,800 ctx
reasoning
GLM-4.6V
GLM vision model for visual reasoning, documents, and multimodal agents
131,072 ctx
reasoning
GLM-4.7
Mature GLM model for dependable coding, reasoning, and structured agent tasks
204,800 ctx
reasoning
GLM-4.7-Flash
Budget GLM lane for fast coding help, routing, and everyday automation
202,752 ctx
reasoning
GLM-5
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
204,800 ctx
reasoning
GLM-5-Turbo
Faster GLM-5 lane for coding agents that need lower latency
202,752 ctx
reasoning
GLM-5.1
Strong GLM coding model for agentic engineering, terminals, and repository generation
204,800 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,048,576 ctx
reasoning
GLM-5V-Turbo
Fast GLM vision model for screenshots, documents, and multimodal agent tasks
202,752 ctx
reasoning
Claude Fable Latest
Claude model for creative writing, analysis, and controlled agent workflows
1,000,000 ctx
reasoning
Anthropic Claude Haiku Latest
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
reasoning
Claude Opus Latest
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Anthropic Claude Sonnet Latest
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Google Gemini Flash Latest
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Google Gemini Pro Latest
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,576 ctx
reasoning
MoonshotAI Kimi Latest
Kimi multimodal agent model for visual understanding, coding, and planning
1,048,576 ctx
reasoning
OpenAI GPT Latest
GPT model for general reasoning, writing, coding, and tool-assisted tasks
1,050,000 ctx
reasoning
OpenAI GPT Mini Latest
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
Grok Latest
Grok model for agentic tool use, reasoning, coding, and live assistance
500,000 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate