Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
Hugging Face
Provider ID: huggingface
Details
SDK package
@ai-sdk/openai-compatible
Models
52
API
https://router.huggingface.co/v1
Links
Documentation
Auth: HF_TOKEN
Models (52)
MiniMax-M2
Efficient open MiniMax model built for coding agents and tool-heavy workflows
204,800 ctx
reasoning
MiniMax-M2.1
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax-M2.5
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax-M2.7
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax-M3
MiniMax multimodal model for long-context coding, perception, and agent planning
524,288 ctx
reasoning
Qwen3 235B-A22B
Large open Qwen MoE for multilingual reasoning, coding, and tool use
40,960 ctx
reasoning
Qwen3-235B-A22B-Thinking-2507
Qwen reasoning model for deliberate problem solving, math, and coding
262,144 ctx
reasoning
Qwen3 32B
Dense open Qwen model for self-hosted chat, reasoning, and coding
131,072 ctx
reasoning
Qwen3-Coder 30B-A3B Instruct
Smaller Qwen coder for efficient local agents and repo-level fixes
262,144 ctx
Qwen3-Coder-480B-A35B-Instruct
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3-Coder-Next
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen 3 Embedding 4B
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
32,000 ctx
Qwen 3 Embedding 8B
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
32,000 ctx
Qwen3-Next-80B-A3B-Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
Qwen3-Next-80B-A3B-Thinking
Qwen reasoning model for deliberate problem solving, math, and coding
262,144 ctx
Qwen3.5 122B-A10B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 35B-A3B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5-397B-A17B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5 9B
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
reasoning
Qwen3.6 27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.6 35B-A3B
Open multimodal Qwen MoE for local agents that need vision, audio, and code
262,144 ctx
reasoning
MiMo-V2-Flash
MiMo flash model for fast multimodal assistance and agent workflows
262,144 ctx
reasoning
MiMo-V2.5
MiMo model for long-context reasoning, perception, and agentic tasks
262,144 ctx
reasoning
MiMo-V2.5-Pro
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
1,048,576 ctx
reasoning
DeepSeek-R1
Classic open reasoning model for transparent math, coding, and deliberate problem solving
64,000 ctx
reasoning
DeepSeek-R1-0528
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
163,840 ctx
reasoning
DeepSeek-V3.2
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,048,576 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,048,576 ctx
reasoning
Gemma 4 26B A4B IT
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
Llama-3.3-70B-Instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
131,072 ctx
Kimi-K2-Instruct
Kimi model for long-context chat, coding, and agentic reasoning
131,072 ctx
Kimi-K2-Instruct-0905
Kimi model for long-context chat, coding, and agentic reasoning
262,144 ctx
Kimi-K2-Thinking
Kimi reasoning model for long-horizon research, planning, and tool use
262,144 ctx
reasoning
Kimi-K2.5
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Kimi-K2.6
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
GPT OSS 120B
Open GPT reasoning model for self-hosted agents and controllable deployments
131,072 ctx
reasoning
GPT OSS 20B
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
Step 3.5 Flash
StepFun flash lane for quick multimodal reasoning and coding assistance
262,144 ctx
reasoning
Step 3.7 Flash
Newer StepFun flash model for faster agents, coding, and multimodal prompts
262,144 ctx
reasoning
GLM-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
131,072 ctx
reasoning
GLM-4.5-Air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
131,072 ctx
reasoning
GLM-4.5V
GLM vision model for visual reasoning, documents, and multimodal agents
65,536 ctx
reasoning
GLM-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
204,800 ctx
reasoning
GLM-4.7
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
204,800 ctx
reasoning
GLM-4.7-Flash
Efficient GLM model for fast reasoning, coding, and agent workflows
200,000 ctx
reasoning
GLM-5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-5.1
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
262,144 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate