Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Weights & Biases
Provider ID: wandb
Details
SDK package
@ai-sdk/openai-compatible
Models
30
API
https://api.inference.wandb.ai/v1
Links
Documentation
Auth: WANDB_API_KEY
Models (30)
Mellum2 12B A2.5B
Mellum2-12B-A2.5B-Instruct is a fast MoE model with 131K context built for coding, tool use, and low-latency AI workflows.
131,072 ctx
MiniMax M2.5
MoE model with a highly sparse architecture designed for high-throughput and low latency with strong coding capabilities.
196,608 ctx
reasoning
MiniMax M3
MiniMax M3 is a multimodal MoE model with 23B active parameters optimized for coding and agentic workflows.
262,144 ctx
reasoning
Qwen3 14B Instruct
An efficient multilingual, dense, instruction-tuned model, optimized by OpenPipe for building agents with finetuning.
32,768 ctx
Qwen3 235B A22B-2507
Efficient multilingual, Mixture-of-Experts, instruction-tuned model, optimized for logical reasoning.
262,144 ctx
Qwen3 235B A22B Thinking-2507
High-performance Mixture-of-Experts model optimized for structured reasoning, math, and long-form generation.
262,144 ctx
reasoning
Qwen3 30B A3B Instruct 2507
Qwen3-30B-A3B-Instruct-2507 is a 30.5B MoE instruction-tuned model with enhanced reasoning, coding, and long-context understanding.
262,144 ctx
Qwen3 Coder 480B A35B
Mixture-of-Experts model optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning.
262,144 ctx
Qwen3.5-27B
Qwen3.5-27B is a dense model from the Qwen3.5 family built for high performance across a large range of benchmarks.
262,144 ctx
reasoning
Qwen3.5-35B-A3B
Qwen3.5-35B-A3B is an open-weights multimodal MoE model built for efficient, high-throughput inference across chat, reasoning, and agentic tasks.
262,144 ctx
reasoning
Qwen3.6 27B
Qwen3.6-27B is a 27B dense multimodal model with 262K context built for flagship-level agentic coding.
262,144 ctx
reasoning
Qwen3.6 35B A3B
Qwen3.6-35B-A3B is an MoE multimodal model with 262K context optimized for agentic coding workflows.
262,144 ctx
reasoning
DeepSeek V3.1
A large hybrid model that supports both thinking and non-thinking modes via prompt templates.
161,000 ctx
DeepSeek V4 Flash
DeepSeek V4-Flash is an MoE model with 1M context length great for coding, reasoning, and agentic workloads.
1,048,576 ctx
reasoning
DeepSeek V4 Pro
DeepSeek V4-Pro is a 1.6T-parameter MoE model with 49B active parameters excelling at advanced reasoning, coding, and complex agentic workloads.
1,048,576 ctx
reasoning
Gemma 4 31B
Gemma 4 31B Dense is designed for advanced reasoning, agentic workflows, and longer context and is natively trained on 140+ languages.
262,144 ctx
reasoning
Granite 4.1 8B
Granite 4.1 8B is a long-context instruct model capable of enhanced tool calling, instruction following, and chat capabilities.
131,072 ctx
Llama 3.1 70B
Efficient conversational model optimized for responsive multilingual chatbot interactions.
128,000 ctx
Llama 3.1 8B
Efficient conversational model optimized for responsive multilingual chatbot interactions.
128,000 ctx
Llama 3.3 70B
Multilingual model excelling in conversational tasks, detailed instruction-following, and coding.
128,000 ctx
Phi 4 Mini 3.8B
Compact, efficient model ideal for fast responses in resource-constrained environments.
128,000 ctx
Kimi K2.5
Kimi K2.5 is a multimodal Mixture-of-Experts language model featuring 32 billion activated parameters and a total of 1 trillion parameters.
262,144 ctx
reasoning
Kimi K2.6
Kimi K2.6 is a multimodal Mixture-of-Experts language model featuring 32 billion activated parameters and a total of 1 trillion parameters.
262,144 ctx
reasoning
Kimi K2.7 Code
Kimi K2.7 Code is a 1T-parameter MoE model with 32B active parameters purpose-built for long-horizon agentic coding and software engineering.
262,144 ctx
reasoning
Nemotron 3 Super
Nemotron 3 is a LatentMoE model designed to deliver strong agentic, reasoning, and conversational capabilities.
262,144 ctx
reasoning
Nemotron 3 Ultra
Nemotron 3 Ultra is a powerful MoE model designed for long-running agents across coding, deep research, and enterprise automation.
262,144 ctx
reasoning
gpt-oss-120b
Efficient Mixture-of-Experts model designed for high-reasoning, agentic and general-purpose use cases.
131,072 ctx
reasoning
gpt-oss-20b
Lower latency Mixture-of-Experts model trained on OpenAI's Harmony response format with reasoning capabilities.
131,072 ctx
reasoning
GLM 5.1
Powerful MoE model for long-horizon agentic engineering and advanced reasoning.
202,752 ctx
reasoning
GLM 5.2
GLM-5.2 is a Mixture-of-Experts language model featuring 40 billion activated parameters and a total of 744 billion parameters.
262,144 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate