Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Abacus
Provider ID: abacus
Details
SDK package
@ai-sdk/openai-compatible
Models
95
API
https://routellm.abacus.ai/v1
Links
Documentation
Auth: ABACUS_API_KEY
Models (95)
MiniMax-M2.7
Open MiniMax flagship for coding agents, office automation, and complex environments
204,800 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
1,000,000 ctx
reasoning
QwQ 32B
Qwen reasoning model for deliberate problem solving, math, and coding
32,768 ctx
reasoning
Qwen 2.5 72B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
128,000 ctx
Qwen3 235B A22B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
reasoning
Qwen3 32B
Dense open Qwen model for self-hosted chat, reasoning, and coding
131,072 ctx
reasoning
Qwen3-Coder 480B-A35B Instruct
Open Qwen coding heavyweight for repository reasoning and agentic engineering
262,144 ctx
Qwen3.6 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Claude Sonnet 3.7
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Fable 5
Claude model for creative writing, analysis, and controlled agent workflows
1,000,000 ctx
reasoning
Claude Haiku 4.5
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
reasoning
Claude Opus 4.1
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
High-end Claude for difficult coding, planning, and slower expert reasoning
1,000,000 ctx
reasoning
Claude Opus 4.7
Stronger Opus tier for advanced software work and high-stakes reasoning
1,000,000 ctx
reasoning
Claude Opus 4.8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Sonnet 4
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.5
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
Claude Sonnet 5
Everyday Claude agent model for coding, planning, browsing, and general work
1,000,000 ctx
reasoning
DeepSeek R1
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
128,000 ctx
reasoning
DeepSeek V3.1 Terminus
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
DeepSeek V3.2
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,000,000 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,000,000 ctx
reasoning
DeepSeek V3.1
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
Gemini 2.5 Flash
Fast Gemini workhorse for multimodal apps where latency and price matter
1,048,576 ctx
reasoning
Nano Banana
Nano Banana image model for fast generation, edits, and character-consistent assets
32,768 ctx
reasoning
Gemini 2.5 Pro
Google's proven reasoning model for coding, math, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3 Flash Preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
1,048,576 ctx
reasoning
Nano Banana Pro
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
65,536 ctx
reasoning
Nano Banana 2
Image model for prompt-driven generation, editing, and visual design workflows
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite Preview
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
1,048,576 ctx
reasoning
Gemini 3.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
GPT-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
128,000 ctx
GPT-4o (2024-11-20)
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
128,000 ctx
GPT-5
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
400,000 ctx
reasoning
GPT-5-Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5 Mini
Small GPT-5 for responsive agents, coding help, and everyday automation
400,000 ctx
reasoning
GPT-5 Nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
400,000 ctx
reasoning
GPT-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.1 Chat Latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
400,000 ctx
reasoning
GPT-5.1 Codex
Codex GPT for repository edits, code review, and practical software agents
400,000 ctx
reasoning
GPT-5.1 Codex Max
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
400,000 ctx
reasoning
GPT-5.2 Chat Latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
400,000 ctx
reasoning
GPT-5.2 Codex
Code-specialist GPT for repository edits, reviews, and long-running software agents
400,000 ctx
reasoning
GPT-5.3 Chat Latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
400,000 ctx
reasoning
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.3 Codex XHigh
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
400,000 ctx
reasoning
GPT-5.4 mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
400,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,000,000 ctx
reasoning
GPT-5.6 Luna
Cost-efficient GPT-5.6 model for fast, high-volume workloads
1,000,000 ctx
reasoning
GPT-5.6 Sol
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
1,000,000 ctx
reasoning
GPT-5.6 Terra
Balanced GPT-5.6 model for capable, cost-efficient everyday work
1,000,000 ctx
reasoning
Grok 4
Grok model for agentic tool use, reasoning, coding, and live assistance
256,000 ctx
reasoning
Grok 4.1 Fast (Non-Reasoning)
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
Grok 4 Fast (Non-Reasoning)
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
Grok 4.3
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
1,000,000 ctx
reasoning
Grok 4.5
xAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk
500,000 ctx
reasoning
Grok Code Fast 1
Fast Grok model for responsive chat, reasoning, and tool-assisted work
256,000 ctx
Kimi K2 Turbo Preview
Fast Kimi model for responsive chat, coding help, and agent loops
256,000 ctx
Kimi K2.5
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Llama 3.3 70B Versatile
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
Llama 4 Maverick 17B Instruct
Open multimodal Llama for strong reasoning with efficient everyday serving
1,048,576 ctx
Llama 3.1 405B Instruct Turbo
Compact Llama instruction model for fast chat and local deployment
128,000 ctx
Llama 3.1 8B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
Llama-3.3-70B-Instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
131,072 ctx
MiMo-V2-Pro
Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks
1,048,576 ctx
reasoning
Kimi K2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
262,144 ctx
reasoning
Muse Spark 1.1
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
1,000,000 ctx
reasoning
o3
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
200,000 ctx
reasoning
o3-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
200,000 ctx
reasoning
o3-pro
High-effort o3 tier for difficult technical reasoning and careful answers
200,000 ctx
reasoning
o4-mini
Fast o-series model for compact reasoning, coding, and tool use
200,000 ctx
reasoning
GPT OSS 120B
Open GPT reasoning model for self-hosted agents and controllable deployments
128,000 ctx
reasoning
Qwen 2.5 Coder 32B
Qwen coding model for software agents, repository edits, and code reasoning
128,000 ctx
Qwen3 Max
Flagship Qwen model for complex reasoning, coding, and agentic workflows
131,072 ctx
reasoning
RouteLLM
RouteLLM routes prompts to an appropriate Abacus-backed text-generation model
128,000 ctx
GLM-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
131,072 ctx
reasoning
GLM-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
202,752 ctx
reasoning
GLM-4.7
Mature GLM model for dependable coding, reasoning, and structured agent tasks
204,800 ctx
reasoning
GLM-5
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
204,800 ctx
reasoning
GLM-5.1
Strong GLM coding model for agentic engineering, terminals, and repository generation
204,800 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,048,576 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate