Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
DigitalOcean
Provider ID: digitalocean
Details
SDK package
@ai-sdk/openai-compatible
Models
82
API
https://inference.do-ai.run/v1
Links
Documentation
Auth: DIGITALOCEAN_ACCESS_TOKEN
Models (82)
Qwen3-32B
Qwen instruction model for multilingual chat, reasoning, and tool use
131,000 ctx
reasoning
All-MiniLM-L6-v2
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
256 ctx
Claude 3 Opus
Legacy model retained for compatibility with older integrations
200,000 ctx
Claude 3.5 Haiku
Legacy model retained for compatibility with older integrations
200,000 ctx
Claude 3.5 Sonnet
Legacy model retained for compatibility with older integrations
200,000 ctx
Claude 3.7 Sonnet
Legacy model retained for compatibility with older integrations
200,000 ctx
reasoning
Claude Opus 4.1
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Haiku 4.5
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
reasoning
Claude Sonnet 4.5
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Claude Sonnet 4.6
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Anthropic Claude Fable 5
Claude model for creative writing, analysis, and controlled agent workflows
1,000,000 ctx
reasoning
Claude Haiku 4.5
Fast Claude model for responsive assistance, classification, and lightweight agents
200,000 ctx
reasoning
Claude Opus 4
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.5
Flagship Claude model for deep reasoning, coding, and long-horizon agents
200,000 ctx
reasoning
Claude Opus 4.6
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.7
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Opus 4.8
Flagship Claude model for deep reasoning, coding, and long-horizon agents
1,000,000 ctx
reasoning
Claude Sonnet 4
Balanced Claude model for coding, analysis, agent workflows, and cost control
1,000,000 ctx
reasoning
Trinity Large Thinking
Flagship model for demanding analysis, coding, and production agent workflows
256,000 ctx
reasoning
BGE M3
Flagship model for demanding analysis, coding, and production agent workflows
8,192 ctx
BGE Reranker v2 M3
Reranking model for improving retrieval quality in search and recommendation systems
8,192 ctx
DeepSeek V3.2
DeepSeek chat model for instruction following, coding, and analysis
128,000 ctx
reasoning
Deepseek V4 Flash
Fast DeepSeek model for efficient chat, coding help, and agent loops
262,144 ctx
DeepSeek R1 Distill Llama 70B
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
131,072 ctx
reasoning
DeepSeek V3
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
DeepSeek V4 Pro
Flagship DeepSeek model for coding, reasoning, and agentic work
1,048,576 ctx
reasoning
E5 Large v2
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
512 ctx
ElevenLabs Multilingual TTS v2
Speech generation model for controllable voice, narration, and audio delivery
0
Fast SDXL
Image model for prompt-driven generation, editing, and visual design workflows
0
FLUX.1 [schnell]
Image model for prompt-driven generation, editing, and visual design workflows
0
Stable Audio 2.5 (Text-to-Audio)
Speech generation model for controllable voice, narration, and audio delivery
0
Gemma 4 31B
Open Gemma instruction model for efficient chat and self-hosted deployments
256,000 ctx
GLM 5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-5.1
Strong GLM coding model for agentic engineering, terminals, and repository generation
200,000 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,000,000 ctx
reasoning
GTE Large (v1.5)
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
8,192 ctx
Kimi K2.5
Kimi model for long-context chat, coding, and agentic reasoning
262,144 ctx
reasoning
Kimi K2.6
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Llama 4 Maverick 17B 128E Instruct
Open multimodal Llama model for strong reasoning and fast responses
1,000,000 ctx
Llama 3.1 Instruct (8B)
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Llama 3.3 Instruct 70B
Open Llama instruction model for multilingual chat, reasoning, and coding
128,000 ctx
MiniMax M2.5
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
Ministral 3 8B
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
262,144 ctx
Ministral 3 14B Instruct
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
262,144 ctx
Mistral 7B Instruct v0.3
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
32,768 ctx
Mistral Nemo Instruct
Legacy model retained for compatibility with older integrations
128,000 ctx
Multi-QA-mpnet-base-dot-v1
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
512 ctx
Nemotron 3 Nano 30B A3B
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
262,144 ctx
reasoning
Nemotron Nano 3 Omni
Open Nemotron omni model combining reasoning with text, vision, and audio
65,536 ctx
reasoning
Nemotron 3 Ultra
Flagship Nemotron model for high-throughput reasoning and complex agents
131,072 ctx
Nemotron Nano 12B v2 VL
Nemotron multimodal model for visual reasoning and agentic AI workflows
128,000 ctx
reasoning
Nemotron-3-Super-120B
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
256,000 ctx
reasoning
GPT-4.1
GPT model for general reasoning, writing, coding, and tool-assisted tasks
1,047,576 ctx
GPT-4o
GPT model for general reasoning, writing, coding, and tool-assisted tasks
128,000 ctx
GPT-4o mini
Compact GPT model for low-latency assistance and high-volume workloads
128,000 ctx
GPT-5
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5 mini
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5 nano
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5.1 Codex Max
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.2
GPT model for general reasoning, writing, coding, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.2 pro
Frontier GPT model for professional reasoning, coding, and multimodal work
400,000 ctx
reasoning
GPT-5.3 Codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
400,000 ctx
reasoning
GPT-5.4
Frontier GPT model for professional reasoning, coding, and multimodal work
1,000,000 ctx
reasoning
GPT-5.4 mini
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5.4 nano
Compact GPT model for low-latency assistance and high-volume workloads
400,000 ctx
reasoning
GPT-5.4 pro
Frontier GPT model for professional reasoning, coding, and multimodal work
400,000 ctx
reasoning
GPT-5.5
Frontier GPT model for professional reasoning, coding, and multimodal work
1,000,000 ctx
reasoning
GPT Image 1
Image model for prompt-driven generation, editing, and visual design workflows
0
GPT Image 1.5
Image model for prompt-driven generation, editing, and visual design workflows
0
GPT Image 2
Image model for prompt-driven generation, editing, and visual design workflows
0
gpt-oss-120b
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
gpt-oss-20b
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
o1
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o3
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
o3-mini
O-series reasoning model for hard analysis, math, coding, and planning
200,000 ctx
reasoning
Qwen 2.5 14B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
Qwen3 Coder Flash
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3 Embedding 0.6B
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
8,000 ctx
Qwen3 TTS VoiceDesign
Speech generation model for controllable voice, narration, and audio delivery
32,768 ctx
Qwen 3.5 397B A17B
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
reasoning
Stable Diffusion 3.5 Large
Image model for prompt-driven generation, editing, and visual design workflows
256 ctx
Wan2.2-T2V-A14B
Video model for prompt-guided generation, editing, and motion workflows
100 ctx
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate