Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Merge Gateway
Provider ID: merge-gateway
Details
SDK package
merge-gateway-ai-sdk-provider
Models
93
Links
Documentation
Auth: MERGE_GATEWAY_API_KEY
Models (93)
Qwen3.6 Plus
Earlier Qwen multimodal workhorse for million-token agent and document tasks
1,000,000 ctx
reasoning
Qwen3.7 Max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
1,000,000 ctx
reasoning
Claude Haiku 4.5
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
reasoning
Claude Opus 4.1
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
High-end Claude for difficult coding, planning, and slower expert reasoning
1,000,000 ctx
reasoning
Claude Opus 4.7
Stronger Opus tier for advanced software work and high-stakes reasoning
1,000,000 ctx
reasoning
Claude Opus 4.8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Sonnet 4
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.5
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
Command A
Cohere command model for multilingual enterprise agents, tools, and chat
256,000 ctx
Command R
Cohere retrieval model for long-context chat and enterprise RAG workflows
128,000 ctx
Command R+
Cohere's RAG workhorse for long-context enterprise search and tool use
128,000 ctx
Command R7B
Cohere retrieval model for long-context chat and enterprise RAG workflows
128,000 ctx
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,000,000 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,000,000 ctx
reasoning
Gemini 2.5 Flash
Fast Gemini workhorse for multimodal apps where latency and price matter
1,048,576 ctx
reasoning
Gemini 2.5 Flash-Lite
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
1,048,576 ctx
reasoning
Gemini 2.5 Pro
Google's proven reasoning model for coding, math, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3 Flash Preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
1,048,576 ctx
reasoning
Gemini 3 Pro Preview
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite Preview
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview Custom Tools
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemini Flash Latest
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemini Flash-Lite Latest
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemma 4 26B A4B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
MiniMax-M2
Efficient open MiniMax model built for coding agents and tool-heavy workflows
196,608 ctx
reasoning
MiniMax-M2.1
Earlier MiniMax agent model for practical coding and productivity tasks
204,800 ctx
reasoning
MiniMax-M2.5
Prior MiniMax coding model for agent workflows, office edits, and automation
204,800 ctx
reasoning
MiniMax-M2.5-highspeed
High-speed MiniMax model for low-latency coding and agent workflows
204,800 ctx
reasoning
MiniMax-M2.7
Open MiniMax flagship for coding agents, office automation, and complex environments
204,800 ctx
reasoning
MiniMax-M2.7-highspeed
Low-latency M2.7 variant for interactive coding plans and agent loops
204,800 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
512,000 ctx
reasoning
Codestral (latest)
Mistral code model for completions, refactors, and developer IDE workflows
256,000 ctx
Devstral 2
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
262,144 ctx
Devstral Medium
Mistral coding agent model for repository tasks and software engineering workflows
128,000 ctx
Devstral 2 (latest)
Mistral coding agent model for repository tasks and software engineering workflows
262,144 ctx
Devstral Small
Mistral coding agent model for repository tasks and software engineering workflows
128,000 ctx
Magistral Medium (latest)
Mistral reasoning model for transparent analysis, math, and complex decisions
128,000 ctx
reasoning
Mistral Large 2.1
Flagship Mistral model for advanced reasoning, coding, and multilingual work
131,072 ctx
Mistral Large 3
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
262,144 ctx
Mistral Large (latest)
Flagship Mistral model for advanced reasoning, coding, and multilingual work
262,144 ctx
Mistral Medium 3
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
131,072 ctx
Mistral Medium (latest)
Balanced Mistral model for enterprise assistants, multilingual work, and tools
262,144 ctx
reasoning
Mistral Small (latest)
Efficient Mistral model for fast chat, extraction, and production assistants
256,000 ctx
reasoning
Pixtral Large (latest)
Mistral's larger vision model for document-heavy image understanding and chat
128,000 ctx
Kimi K2 Thinking
Thinking Kimi model for slower research passes, planning, and hard technical questions
262,144 ctx
reasoning
Kimi K2.5
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
262,144 ctx
reasoning
Kimi K2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
262,144 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Kimi K2.7 Code Highspeed
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
262,144 ctx
reasoning
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
GPT-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
128,000 ctx
GPT-4o (2024-05-13)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o (2024-08-06)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o (2024-11-20)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
128,000 ctx
GPT-5
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
400,000 ctx
reasoning
GPT-5 Chat (latest)
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
400,000 ctx
reasoning
GPT-5 Mini
Small GPT-5 for responsive agents, coding help, and everyday automation
400,000 ctx
reasoning
GPT-5 Nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
400,000 ctx
reasoning
GPT-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.1 Chat
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
128,000 ctx
reasoning
GPT-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
400,000 ctx
reasoning
GPT-5.2 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
reasoning
GPT-5.3 Chat (latest)
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
1,050,000 ctx
reasoning
GPT-5.4 mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
400,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
o1
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o3
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
200,000 ctx
reasoning
o3-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
200,000 ctx
reasoning
o4-mini
Fast o-series model for compact reasoning, coding, and tool use
200,000 ctx
reasoning
Grok 4.20 (Reasoning)
Reasoning Grok for document-heavy analysis and long-horizon tool use
1,000,000 ctx
reasoning
Grok 4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
1,000,000 ctx
reasoning
GLM-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
131,072 ctx
reasoning
GLM-4.5-Air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
131,072 ctx
reasoning
GLM-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
204,800 ctx
reasoning
GLM-4.7
Mature GLM model for dependable coding, reasoning, and structured agent tasks
204,800 ctx
reasoning
GLM-4.7-FlashX
Efficient GLM model for fast reasoning, coding, and agent workflows
200,000 ctx
reasoning
GLM-5
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
204,800 ctx
reasoning
GLM-5-Turbo
Faster GLM-5 lane for coding agents that need lower latency
200,000 ctx
reasoning
GLM-5.1
Strong GLM coding model for agentic engineering, terminals, and repository generation
200,000 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,000,000 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate