Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Pioneer
Provider ID: pioneer
Details
SDK package
@ai-sdk/openai-compatible
Models
76
API
https://api.pioneer.ai/v1
Links
Documentation
Auth: PIONEER_API_KEY
Models (76)
SmolLM3 3B Base
Tool-capable chat model for instruction following and agentic application workflows
32,768 ctx
reasoning
LFM2 24B A2B
Open-weight instruction model for adaptable chat and self-hosted production workloads
32,768 ctx
reasoning
MiniMax-M2.7
Open MiniMax flagship for coding agents, office automation, and complex environments
204,800 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
1,000,000 ctx
reasoning
Qwen3 1.7B Base
Qwen instruction model for multilingual chat, reasoning, and tool use
32,768 ctx
reasoning
Qwen3 32B
Dense open Qwen model for self-hosted chat, reasoning, and coding
131,072 ctx
reasoning
Qwen3 4B Base
Qwen instruction model for multilingual chat, reasoning, and tool use
32,768 ctx
reasoning
Qwen3 4B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
reasoning
Qwen3 8B
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
reasoning
Qwen3.5 9B
Qwen instruction model for multilingual chat, reasoning, and tool use
32,768 ctx
reasoning
Qwen3.6 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
32,768 ctx
reasoning
Qwen3.6 35B-A3B
Open multimodal Qwen MoE for local agents that need vision, audio, and code
262,144 ctx
reasoning
MiMo-V2.5
Open MiMo model for multimodal coding agents and long-context automation
1,050,000 ctx
reasoning
MiMo-V2.5-Pro
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
1,050,000 ctx
reasoning
Claude Haiku 4.5 (latest)
Fast Claude lane for lightweight agents, office tasks, and responsive chat
200,000 ctx
reasoning
Claude Opus 4.1 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5 (latest)
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
High-end Claude for difficult coding, planning, and slower expert reasoning
1,000,000 ctx
reasoning
Claude Opus 4.7
Stronger Opus tier for advanced software work and high-stakes reasoning
1,000,000 ctx
reasoning
Claude Opus 4.8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Sonnet 4.5 (latest)
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,000,000 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,000,000 ctx
reasoning
GLiGuard LLM Guardrails 300M
Tool-capable chat model for instruction following and agentic application workflows
8,192 ctx
GLiNER2 Base
Tool-capable chat model for instruction following and agentic application workflows
8,192 ctx
GLiNER2 Large
Flagship model for demanding analysis, coding, and production agent workflows
8,192 ctx
GLiNER2 Multi Large
Flagship model for demanding analysis, coding, and production agent workflows
8,192 ctx
GLiNER2 Multi
Tool-capable chat model for instruction following and agentic application workflows
8,192 ctx
GLiNER2 Privacy Filter PII (Multi)
Tool-capable chat model for instruction following and agentic application workflows
8,192 ctx
Gemini 3 Flash Preview
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
1,048,576 ctx
reasoning
Gemini 3.1 Pro Preview
Reasoning-first Gemini preview for agentic coding and complex problem solving
1,048,576 ctx
reasoning
Gemini 3.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
DiffusionGemma 26B-A4B IT
Gemini model for general assistance, reasoning, and multimodal workflows
262,144 ctx
reasoning
Gemma 3 4B (Pretrained)
Open Gemma instruction model for efficient chat and self-hosted deployments
32,768 ctx
reasoning
Gemma 4 12B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
32,768 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
32,768 ctx
reasoning
Gemma 4 E2B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
32,768 ctx
reasoning
Gemma 4 E4B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
32,768 ctx
reasoning
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
reasoning
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
reasoning
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
reasoning
GPT-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
128,000 ctx
reasoning
GPT-4o mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
128,000 ctx
reasoning
GPT-5 Mini
Small GPT-5 for responsive agents, coding help, and everyday automation
400,000 ctx
reasoning
GPT-5 Nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
400,000 ctx
reasoning
GPT-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
1,050,000 ctx
reasoning
GPT-5.4 mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
1,047,576 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
Llama 3.1 8B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
reasoning
Llama 3.2 1B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
reasoning
Llama 3.2 3B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
reasoning
Llama-3.3-70B-Instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
131,072 ctx
reasoning
Mistral Medium 3.5
Balanced Mistral model for enterprise assistants, multilingual work, and tools
262,144 ctx
reasoning
Mistral 7B Instruct v0.3
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
32,768 ctx
reasoning
Mistral Nemo
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
131,072 ctx
reasoning
Mistral Small 4
Fast Mistral production model for chat, extraction, and cost-sensitive agents
262,144 ctx
reasoning
Kimi K2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
262,144 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Nemotron 3 Nano 30B A3B
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
262,144 ctx
reasoning
Nemotron 3 Super 120B A12B
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
1,000,000 ctx
reasoning
Nemotron 3 Ultra 550B A55B
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
1,000,000 ctx
reasoning
GPT OSS 120B
Open GPT reasoning model for self-hosted agents and controllable deployments
131,072 ctx
reasoning
GPT OSS 20B
Open GPT reasoning model for self-hosted agents and controllable deployments
131,072 ctx
reasoning
Pioneer Auto
Automatic model router for matching prompts to suitable backends and budgets
1,048,576 ctx
reasoning
Qwen3.6 Flash
Qwen vision-language model for visual reasoning, documents, and agent tasks
1,000,000 ctx
reasoning
Qwen3.6 Max Preview
Flagship Qwen model for complex reasoning, coding, and agentic workflows
262,144 ctx
reasoning
Qwen3.6 Plus
Earlier Qwen multimodal workhorse for million-token agent and document tasks
1,000,000 ctx
reasoning
Qwen3.7 Max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
1,000,000 ctx
reasoning
Qwen3.7 Plus
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
1,000,000 ctx
reasoning
Fugu Ultra
Quality-first multi-agent model for hard research, analysis, and competitions
1,000,000 ctx
reasoning
GLM-5.1
Strong GLM coding model for agentic engineering, terminals, and repository generation
202,752 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,048,576 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate