Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Deep Infra
Provider ID: deepinfra
Details
SDK package
@ai-sdk/deepinfra
Models
40
Links
Documentation
Auth: DEEPINFRA_API_KEY
Models (40)
MiniMax M2.5
MiniMax model for chat, coding, office work, and agentic tasks
196,608 ctx
reasoning
MiniMax-M2.7
Open MiniMax flagship for coding agents, office automation, and complex environments
196,608 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
524,288 ctx
reasoning
Qwen3 32B
Dense open Qwen model for self-hosted chat, reasoning, and coding
40,960 ctx
reasoning
Qwen3 Coder 480B A35B Instruct Turbo
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3 Max
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
256,000 ctx
Qwen3-Next 80B-A3B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
Qwen3.5 122B-A10B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen 3.5 35B A3B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen 3.5 397B A17B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 9B
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
reasoning
Qwen3.6 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.6 35B A3B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.7 Max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
256,000 ctx
MiMo-V2.5
Open MiMo model for multimodal coding agents and long-context automation
262,144 ctx
reasoning
MiMo-V2.5-Pro
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
1,048,576 ctx
reasoning
DeepSeek-R1-0528
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
163,840 ctx
reasoning
DeepSeek-V3.2
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,048,576 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,048,576 ctx
reasoning
Gemma 4 26B A4B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
Llama 3.3 70B Turbo
Compact Llama instruction model for fast chat and local deployment
131,072 ctx
Llama 4 Maverick 17B FP8
Open multimodal Llama model for strong reasoning and fast responses
1,048,576 ctx
Llama 4 Scout 17B
Open multimodal Llama model for long-context analysis and efficient agents
327,680 ctx
Kimi K2.5
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Kimi K2.6
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Llama 3.3 Nemotron Super 49B v1.5
Nemotron model for efficient reasoning, coding, and specialized AI agents
131,072 ctx
reasoning
Nemotron 3 Nano 30B A3B
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
262,144 ctx
reasoning
Nemotron 3 Nano Omni 30B A3B Reasoning
Open Nemotron omni model combining reasoning with text, vision, and audio
262,144 ctx
reasoning
GPT OSS 120B
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
GPT OSS 20B
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
GLM-4.6
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-4.7
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-4.7-Flash
Efficient GLM model for fast reasoning, coding, and agent workflows
202,752 ctx
reasoning
GLM-5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-5.1
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,048,576 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate