Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Azure
Provider ID: azure
Details
SDK package
@ai-sdk/azure
Models
114
Links
Documentation
Auth: AZURE_RESOURCE_NAME, AZURE_API_KEY
Models (114)
Claude Fable 5
Claude model for creative writing, analysis, and controlled agent workflows
1,000,000 ctx
reasoning
Claude Haiku 4.5
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
reasoning
Claude Opus 4.1
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.8
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 5
Strongest Claude Opus model for coding, agents, and professional work
1,000,000 ctx
reasoning
Claude Sonnet 4.5
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
Claude Sonnet 5
Everyday Claude agent model for coding, planning, browsing, and general work
1,000,000 ctx
reasoning
Codestral 25.01
Mistral coding model for code completion, generation, and developer workflows
256,000 ctx
Codex Mini
Coding-optimized GPT model for repository edits, reviews, and agentic software work
200,000 ctx
reasoning
Command A
Cohere command model for multilingual enterprise agents, tools, and chat
131,072 ctx
reasoning
Command R
Cohere retrieval model for long-context chat and enterprise RAG workflows
128,000 ctx
Command R+
Cohere's RAG workhorse for long-context enterprise search and tool use
128,000 ctx
Embed v4
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
128,000 ctx
Embed v3 English
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
512 ctx
Embed v3 Multilingual
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
512 ctx
DeepSeek-R1
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
163,840 ctx
reasoning
DeepSeek-R1-0528
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
163,840 ctx
reasoning
DeepSeek-V3-0324
DeepSeek chat model for instruction following, coding, and analysis
131,072 ctx
DeepSeek-V3.1
DeepSeek chat model for instruction following, coding, and analysis
131,072 ctx
reasoning
DeepSeek-V3.2
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
DeepSeek-V3.2-Speciale
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
DeepSeek-V4-Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,000,000 ctx
reasoning
DeepSeek-V4-Pro
Open MoE flagship with million-token context for coding and long agent runs
1,000,000 ctx
reasoning
GPT-3.5 Turbo 0125
Compact GPT model for low-latency assistance and high-volume workloads
16,384 ctx
GPT-3.5 Turbo 0301
Compact GPT model for low-latency assistance and high-volume workloads
4,096 ctx
GPT-3.5 Turbo 0613
Compact GPT model for low-latency assistance and high-volume workloads
16,384 ctx
GPT-3.5 Turbo 1106
Compact GPT model for low-latency assistance and high-volume workloads
16,384 ctx
GPT-3.5 Turbo Instruct
Compact GPT model for low-latency assistance and high-volume workloads
4,096 ctx
GPT-4
GPT model for general reasoning, writing, coding, and tool-assisted tasks
8,192 ctx
GPT-4 32K
GPT model for general reasoning, writing, coding, and tool-assisted tasks
32,768 ctx
GPT-4 Turbo
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4 Turbo Vision
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
GPT-4o
Omni-era GPT for multimodal chat, practical coding, and general assistants
128,000 ctx
GPT-4o mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
128,000 ctx
GPT-5
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
reasoning
GPT-5-Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5 Mini
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5 Nano
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5 Pro
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
400,000 ctx
reasoning
GPT-5.1
Speech generation model for controllable voice, narration, and audio delivery
400,000 ctx
reasoning
GPT-5.1 Chat
Speech generation model for controllable voice, narration, and audio delivery
128,000 ctx
reasoning
GPT-5.1 Codex
Speech generation model for controllable voice, narration, and audio delivery
400,000 ctx
reasoning
GPT-5.1 Codex Max
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.1 Codex Mini
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.2 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
reasoning
GPT-5.2 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.3 Chat
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
128,000 ctx
reasoning
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
1,050,000 ctx
reasoning
GPT-5.4 Mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 Nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
400,000 ctx
reasoning
GPT-5.4 Pro
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
1,050,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
GPT-5.6 Luna
Cost-efficient GPT-5.6 model for fast, high-volume workloads
1,050,000 ctx
reasoning
GPT-5.6 Sol
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
1,050,000 ctx
reasoning
GPT-5.6 Terra
Balanced GPT-5.6 model for capable, cost-efficient everyday work
1,050,000 ctx
reasoning
GPT Chat Latest
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
reasoning
GPT-Image-1
OpenAI image model for production generation, edits, and brand-safe visual workflows
0
GPT-Image-1.5
Image model for prompt-driven generation, editing, and visual design workflows
0
GPT-Image-2
Image model for prompt-driven generation, editing, and visual design workflows
0
Grok 4.1 Fast (Non-Reasoning)
Fast Grok model for responsive chat, reasoning, and tool-assisted work
128,000 ctx
Grok 4.1 Fast (Reasoning)
Fast Grok model for responsive chat, reasoning, and tool-assisted work
128,000 ctx
reasoning
Grok 4.20 (Non-Reasoning)
Grok model for agentic tool use, reasoning, coding, and live assistance
262,000 ctx
Grok 4.20 (Reasoning)
Grok model for agentic tool use, reasoning, coding, and live assistance
262,000 ctx
reasoning
Grok 4 Fast (Reasoning)
Fast Grok model for responsive chat, reasoning, and tool-assisted work
2,000,000 ctx
reasoning
Kimi K2 Thinking
Kimi reasoning model for long-horizon research, planning, and tool use
262,144 ctx
reasoning
Kimi K2.5
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Kimi K2.6
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Llama-3.2-11B-Vision-Instruct
Open Llama multimodal model for image understanding and text reasoning
128,000 ctx
Llama-3.2-90B-Vision-Instruct
Open Llama multimodal model for image understanding and text reasoning
128,000 ctx
Llama-3.3-70B-Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
Llama 4 Maverick 17B 128E Instruct FP8
Open multimodal Llama model for strong reasoning and fast responses
1,000,000 ctx
Llama 4 Scout 17B 16E Instruct
Open multimodal Llama model for long-context analysis and efficient agents
128,000 ctx
Meta-Llama-3-70B-Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
Meta-Llama-3-8B-Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
Meta-Llama-3.1-405B-Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
Meta-Llama-3.1-70B-Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
Meta-Llama-3.1-8B-Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
Ministral 3B
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
128,000 ctx
Mistral Large 24.11
Flagship Mistral model for advanced reasoning, coding, and multilingual work
128,000 ctx
Mistral Medium 3
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
128,000 ctx
Mistral Nemo
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
128,000 ctx
Mistral Small 3.1
Efficient Mistral model for fast chat, extraction, and production assistants
128,000 ctx
Model Router
Automatic model router for matching prompts to suitable backends and budgets
200,000 ctx
o1
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o1-mini
O-series reasoning model for hard analysis, math, coding, and planning
128,000 ctx
reasoning
o3
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
200,000 ctx
reasoning
o3-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
200,000 ctx
reasoning
o4-mini
Fast o-series model for compact reasoning, coding, and tool use
200,000 ctx
reasoning
Phi-3-medium-instruct (128k)
Open-weight instruction model for adaptable chat and self-hosted production workloads
128,000 ctx
Phi-3-medium-instruct (4k)
Open-weight instruction model for adaptable chat and self-hosted production workloads
4,096 ctx
Phi-3-mini-instruct (128k)
Efficient model for low-latency assistance, extraction, and routine automation
128,000 ctx
Phi-3-mini-instruct (4k)
Efficient model for low-latency assistance, extraction, and routine automation
4,096 ctx
Phi-3-small-instruct (128k)
Efficient model for low-latency assistance, extraction, and routine automation
128,000 ctx
Phi-3-small-instruct (8k)
Efficient model for low-latency assistance, extraction, and routine automation
8,192 ctx
Phi-3.5-mini-instruct
Efficient model for low-latency assistance, extraction, and routine automation
128,000 ctx
Phi-3.5-MoE-instruct
Open-weight instruction model for adaptable chat and self-hosted production workloads
128,000 ctx
Phi-4
Open-weight instruction model for adaptable chat and self-hosted production workloads
128,000 ctx
Phi-4-mini
Efficient model for low-latency assistance, extraction, and routine automation
128,000 ctx
Phi-4-mini-reasoning
Efficient model for low-latency assistance, extraction, and routine automation
128,000 ctx
reasoning
Phi-4-multimodal
Multimodal model for analyzing text, images, documents, and rich media
128,000 ctx
Phi-4-reasoning
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
32,000 ctx
reasoning
Phi-4-reasoning-plus
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
32,000 ctx
reasoning
text-embedding-3-large
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
8,191 ctx
text-embedding-3-small
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
8,191 ctx
text-embedding-ada-002
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
8,192 ctx
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate