Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
ZenMux
Provider ID: zenmux
Details
SDK package
@ai-sdk/openai-compatible
Models
120
API
https://zenmux.ai/api/v1
Links
Documentation
Auth: ZENMUX_API_KEY
Models (120)
Claude 3.5 Haiku
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
Claude 3.7 Sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Fable 5
Claude model for creative writing, analysis, and controlled agent workflows
1,000,000 ctx
reasoning
Claude Haiku 4.5
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
Claude Opus 4
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.1
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.7
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Sonnet 4
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 4.5
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 4.6
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 5
Everyday Claude agent model for coding, planning, browsing, and general work
1,000,000 ctx
reasoning
Claude Sonnet 5 (Free)
Everyday Claude agent model for coding, planning, browsing, and general work
1,000,000 ctx
reasoning
ERNIE 5.0
Multimodal reasoning model for visual analysis, planning, and tool use
128,000 ctx
reasoning
DeepSeek-V3.2 (Non-thinking Mode)
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
DeepSeek V3.2
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
DeepSeek-V3.2-Exp
DeepSeek chat model for instruction following, coding, and analysis
163,000 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,000,000 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,000,000 ctx
reasoning
Gemini 2.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,000 ctx
reasoning
Gemini 2.5 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,000 ctx
Gemini 2.5 Pro
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,000 ctx
reasoning
Gemini 3 Flash Preview
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,000 ctx
reasoning
Gemini 3.1 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite Preview
Low-latency Gemini model for high-volume multimodal and agent workloads
1,050,000 ctx
Gemini 3.1 Pro Preview
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
1,048,000 ctx
reasoning
Gemini 3.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Ling-1T
Tool-capable chat model for instruction following and agentic application workflows
128,000 ctx
Ring-1T
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
128,000 ctx
reasoning
inclusionAI: Ring-2.6-1T
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
262,000 ctx
reasoning
KAT-Coder-Pro-V2
Coding model for repository understanding, refactors, and agentic engineering tasks
256,000 ctx
MiniMax M2
MiniMax model for chat, coding, office work, and agentic tasks
204,000 ctx
reasoning
MiniMax M2.1
MiniMax model for chat, coding, office work, and agentic tasks
204,000 ctx
reasoning
MiniMax M2.5
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax M2.5 highspeed
High-speed MiniMax model for low-latency coding and agent workflows
204,800 ctx
reasoning
MiniMax M2.7
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax M2.7 highspeed
High-speed MiniMax model for low-latency coding and agent workflows
204,800 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
512,000 ctx
reasoning
Kimi K2 0905
Kimi model for long-context chat, coding, and agentic reasoning
262,000 ctx
Kimi K2 Thinking
Kimi reasoning model for long-horizon research, planning, and tool use
262,000 ctx
reasoning
Kimi K2 Thinking Turbo
Kimi reasoning model for long-horizon research, planning, and tool use
262,000 ctx
reasoning
Kimi K2.5
Kimi multimodal agent model for visual understanding, coding, and planning
262,000 ctx
reasoning
Kimi K2.6
Kimi multimodal agent model for visual understanding, coding, and planning
262,140 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Kimi K2.7 Code (Free)
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Kimi K3
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
1,048,576 ctx
reasoning
Kimi K3 (Free)
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
1,048,576 ctx
reasoning
GPT-5
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.1
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.1 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.1-Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.1-Codex-Mini
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.2-Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2-Pro
Frontier GPT model for professional reasoning, coding, and multimodal work
400,000 ctx
reasoning
GPT-5.3 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Frontier GPT model for professional reasoning, coding, and multimodal work
1,050,000 ctx
reasoning
GPT-5.4 Mini
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
GPT-5.4 Nano
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
GPT-5.4 Pro
Frontier GPT model for professional reasoning, coding, and multimodal work
1,050,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
GPT-5.5 Instant
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5.5 Pro
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
1,050,000 ctx
reasoning
GPT-5.6 Luna
Cost-efficient GPT-5.6 model for fast, high-volume workloads
1,050,000 ctx
reasoning
GPT-5.6 Sol
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
1,050,000 ctx
reasoning
GPT-5.6 Terra
Balanced GPT-5.6 model for capable, cost-efficient everyday work
1,050,000 ctx
reasoning
Qwen3-Coder-Plus
Qwen coding model for software agents, repository edits, and code reasoning
1,000,000 ctx
Qwen3-Max-Thinking
Qwen reasoning model for deliberate problem solving, math, and coding
256,000 ctx
reasoning
Qwen3.5 Flash
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,020,000 ctx
Qwen3.5 Plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.6-Plus
Qwen instruction model for multilingual chat, reasoning, and tool use
1,000,000 ctx
reasoning
Qwen3.7 Max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
1,000,000 ctx
reasoning
Qwen3.7 Plus
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
1,000,000 ctx
reasoning
Agnes 1.5 Lite
Efficient model for low-latency assistance, extraction, and routine automation
256,000 ctx
Agnes 1.5 Pro
Flagship model for demanding analysis, coding, and production agent workflows
256,000 ctx
reasoning
Step-3
StepFun flash model for efficient multimodal reasoning, coding, and tool use
65,536 ctx
reasoning
Step 3.5 Flash
StepFun flash model for efficient multimodal reasoning, coding, and tool use
256,000 ctx
Step 3.7 Flash
Newer StepFun flash model for faster agents, coding, and multimodal prompts
256,000 ctx
reasoning
Step 3.7 Flash (Free)
Newer StepFun flash model for faster agents, coding, and multimodal prompts
256,000 ctx
reasoning
Hy3 preview
Tencent Hy reasoning model for coding, instruction following, and agent tasks
256,000 ctx
reasoning
Doubao-Seed-1.8
Multimodal reasoning model for visual analysis, planning, and tool use
256,000 ctx
reasoning
Doubao Seed 2.0 Code
Coding model for repository understanding, refactors, and agentic engineering tasks
256,000 ctx
Doubao-Seed-2.0-lite
Multimodal reasoning model for visual analysis, planning, and tool use
256,000 ctx
reasoning
Doubao-Seed-2.0-mini
Multimodal reasoning model for visual analysis, planning, and tool use
256,000 ctx
reasoning
Doubao-Seed-2.0-pro
Multimodal reasoning model for visual analysis, planning, and tool use
256,000 ctx
reasoning
Doubao-Seed-Code
Coding model for repository understanding, refactors, and agentic engineering tasks
256,000 ctx
reasoning
Grok 4
Grok model for agentic tool use, reasoning, coding, and live assistance
256,000 ctx
reasoning
Grok 4 Fast
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
reasoning
Grok 4.1 Fast
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
reasoning
Grok 4.1 Fast Non Reasoning
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
Grok 4.2 Fast
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
reasoning
Grok 4.2 Fast Non Reasoning
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
Grok 4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
1,000,000 ctx
reasoning
Grok 4.5
xAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk
500,000 ctx
reasoning
Grok Build 0.1
Fast Grok coding model tuned for agentic engineering and iterative edits
256,000 ctx
reasoning
Grok Code Fast 1
Fast Grok model for responsive chat, reasoning, and tool-assisted work
256,000 ctx
reasoning
MiMo-V2-Flash
MiMo flash model for fast multimodal assistance and agent workflows
262,144 ctx
reasoning
MiMo V2 Omni
MiMo omni model for text, image, video, audio, and agents
265,000 ctx
reasoning
MiMo V2 Pro
Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks
1,000,000 ctx
reasoning
MiMo-V2.5
Open MiMo model for multimodal coding agents and long-context automation
1,048,576 ctx
reasoning
MiMo-V2.5-Pro
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
1,048,576 ctx
reasoning
GLM 4.5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
128,000 ctx
reasoning
GLM 4.5 Air
Efficient GLM model for fast reasoning, coding, and agent workflows
128,000 ctx
reasoning
GLM 4.6
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
200,000 ctx
reasoning
GLM 4.6V
GLM vision model for visual reasoning, documents, and multimodal agents
200,000 ctx
reasoning
GLM 4.6V FlashX
GLM vision model for visual reasoning, documents, and multimodal agents
200,000 ctx
reasoning
GLM 4.6V Flash (Free)
GLM vision model for visual reasoning, documents, and multimodal agents
200,000 ctx
reasoning
GLM 4.7
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
200,000 ctx
reasoning
GLM 4.7 Flash (Free)
Efficient GLM model for fast reasoning, coding, and agent workflows
200,000 ctx
reasoning
GLM 4.7 FlashX
Efficient GLM model for fast reasoning, coding, and agent workflows
200,000 ctx
reasoning
GLM 5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
200,000 ctx
reasoning
GLM 5 Turbo
Efficient GLM model for fast reasoning, coding, and agent workflows
200,000 ctx
reasoning
GLM-5.1
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
200,000 ctx
reasoning
GLM 5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,000,000 ctx
reasoning
GLM 5.2 (Free)
Open flagship GLM for long-horizon coding agents and million-token context work
1,000,000 ctx
reasoning
GLM 5V Turbo
GLM vision model for visual reasoning, documents, and multimodal agents
200,000 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate