Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
OrcaRouter
Provider ID: orcarouter
Details
SDK package
@ai-sdk/openai-compatible
Models
81
API
https://api.orcarouter.ai/v1
Links
Documentation
Auth: ORCAROUTER_API_KEY
Models (81)
Claude Haiku 4.5 (latest)
Fast Claude lane for lightweight agents, office tasks, and responsive chat
200,000 ctx
reasoning
Claude Opus 4 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.1 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
High-end Claude for difficult coding, planning, and slower expert reasoning
1,000,000 ctx
reasoning
Claude Opus 4.7
Stronger Opus tier for advanced software work and high-stakes reasoning
1,000,000 ctx
reasoning
Claude Sonnet 4 (latest)
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.5 (latest)
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
DeepSeek Chat
DeepSeek chat model for instruction following, coding, and analysis
1,000,000 ctx
DeepSeek Reasoner
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
1,000,000 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,000,000 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,000,000 ctx
reasoning
Gemini 2.5 Flash
Fast Gemini workhorse for multimodal apps where latency and price matter
1,048,576 ctx
reasoning
Gemini 2.5 Flash-Lite
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
1,048,576 ctx
reasoning
Gemini 2.5 Pro
Google's proven reasoning model for coding, math, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3 Flash Preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
1,048,576 ctx
reasoning
Gemini 3 Pro Preview
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite Preview
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview Custom Tools
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,576 ctx
reasoning
Gemini Flash Latest
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemini Flash-Lite Latest
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemma 4 26B A4B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
Grok 4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
1,000,000 ctx
reasoning
Kimi K2.5
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
262,144 ctx
reasoning
Kimi K2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
262,144 ctx
reasoning
MiniMax-M2.5
Prior MiniMax coding model for agent workflows, office edits, and automation
204,800 ctx
reasoning
MiniMax-M2.5-highspeed
High-speed MiniMax model for low-latency coding and agent workflows
204,800 ctx
reasoning
MiniMax-M2.7
Open MiniMax flagship for coding agents, office automation, and complex environments
204,800 ctx
reasoning
MiniMax-M2.7-highspeed
Low-latency M2.7 variant for interactive coding plans and agent loops
204,800 ctx
reasoning
GPT-3.5-turbo
Compact GPT model for low-latency assistance and high-volume workloads
16,385 ctx
GPT-4
GPT model for general reasoning, writing, coding, and tool-assisted tasks
8,192 ctx
GPT-4 Turbo
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
GPT-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
128,000 ctx
GPT-4o (2024-05-13)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o (2024-08-06)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o (2024-11-20)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
128,000 ctx
GPT-5
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
400,000 ctx
reasoning
GPT-5 Chat (latest)
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
400,000 ctx
reasoning
GPT-5-Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5 Mini
Small GPT-5 for responsive agents, coding help, and everyday automation
400,000 ctx
reasoning
GPT-5 Nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
400,000 ctx
reasoning
GPT-5 Pro
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
400,000 ctx
reasoning
GPT-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.1 Chat
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
128,000 ctx
reasoning
GPT-5.1 Codex
Codex GPT for repository edits, code review, and practical software agents
400,000 ctx
reasoning
GPT-5.1 Codex Max
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.1 Codex mini
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
400,000 ctx
reasoning
GPT-5.2 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
reasoning
GPT-5.2 Codex
Code-specialist GPT for repository edits, reviews, and long-running software agents
400,000 ctx
reasoning
GPT-5.2 Pro
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
400,000 ctx
reasoning
GPT-5.3 Chat (latest)
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
1,050,000 ctx
reasoning
GPT-5.4 mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
400,000 ctx
reasoning
GPT-5.4 Pro
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
1,050,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
GPT-5.5 Pro
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
1,050,000 ctx
reasoning
OrcaRouter Auto
Automatic model router for matching prompts to suitable backends and budgets
128,000 ctx
Qwen3 Max
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
262,144 ctx
Qwen3.5 122B-A10B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 35B-A3B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 397B-A17B
Large open Qwen multimodal MoE for visual agents and long technical tasks
262,144 ctx
reasoning
Qwen3.5 Plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.6 35B-A3B
Open multimodal Qwen MoE for local agents that need vision, audio, and code
262,144 ctx
reasoning
Qwen3.6 Plus
Earlier Qwen multimodal workhorse for million-token agent and document tasks
1,000,000 ctx
reasoning
GLM-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
131,072 ctx
reasoning
GLM-4.5-Air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
131,072 ctx
reasoning
GLM-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
204,800 ctx
reasoning
GLM-4.7
Mature GLM model for dependable coding, reasoning, and structured agent tasks
204,800 ctx
reasoning
GLM-5
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
204,800 ctx
reasoning
GLM-5.1
Strong GLM coding model for agentic engineering, terminals, and repository generation
200,000 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate