Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
NovitaAI
Provider ID: novita-ai
Details
SDK package
@ai-sdk/openai-compatible
Models
107
API
https://api.novita.ai/openai
Links
Documentation
Auth: NOVITA_API_KEY
Models (107)
baichuan-m2-32b
Open-weight instruction model for adaptable chat and self-hosted production workloads
131,072 ctx
ERNIE 4.5 21B A3B
Open-weight instruction model for adaptable chat and self-hosted production workloads
120,000 ctx
ERNIE-4.5-21B-A3B-Thinking
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
131,072 ctx
reasoning
ERNIE 4.5 300B A47B
Open-weight instruction model for adaptable chat and self-hosted production workloads
123,000 ctx
ERNIE 4.5 VL 28B A3B
Multimodal reasoning model for visual analysis, planning, and tool use
30,000 ctx
reasoning
ERNIE-4.5-VL-28B-A3B-Thinking
Multimodal reasoning model for visual analysis, planning, and tool use
131,072 ctx
reasoning
ERNIE 4.5 VL 424B A47B
Multimodal reasoning model for visual analysis, planning, and tool use
123,000 ctx
reasoning
DeepSeek-OCR
OCR model for extracting structured text from documents and screenshots
8,192 ctx
deepseek/deepseek-ocr-2
OCR model for extracting structured text from documents and screenshots
8,192 ctx
Deepseek Prover V2 671B
Flagship DeepSeek model for coding, reasoning, and agentic work
160,000 ctx
DeepSeek R1 0528
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
163,840 ctx
reasoning
DeepSeek R1 0528 Qwen3 8B
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
128,000 ctx
reasoning
DeepSeek R1 Distill LLama 70B
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
8,192 ctx
reasoning
DeepSeek R1 Distill Qwen 14B
Qwen reasoning model for deliberate problem solving, math, and coding
32,768 ctx
DeepSeek R1 Distill Qwen 32B
Qwen reasoning model for deliberate problem solving, math, and coding
64,000 ctx
DeepSeek R1 (Turbo)
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
64,000 ctx
reasoning
DeepSeek V3 0324
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
DeepSeek V3 (Turbo)
Fast DeepSeek model for efficient chat, coding help, and agent loops
64,000 ctx
DeepSeek V3.1
DeepSeek chat model for instruction following, coding, and analysis
131,072 ctx
reasoning
Deepseek V3.1 Terminus
DeepSeek chat model for instruction following, coding, and analysis
131,072 ctx
reasoning
Deepseek V3.2
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
Deepseek V3.2 Exp
DeepSeek chat model for instruction following, coding, and analysis
163,840 ctx
reasoning
DeepSeek V4 Flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
1,048,576 ctx
reasoning
DeepSeek V4 Pro
Open MoE flagship with million-token context for coding and long agent runs
1,048,576 ctx
reasoning
Gemma 3 12B
Open Gemma instruction model for efficient chat and self-hosted deployments
131,072 ctx
Gemma 3 27B
Open Gemma instruction model for efficient chat and self-hosted deployments
98,304 ctx
Gemma 4 26B A4B
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Gemma 4 31B
Open Gemma instruction model for efficient chat and self-hosted deployments
262,144 ctx
reasoning
Mythomax L2 13B
Open-weight instruction model for adaptable chat and self-hosted production workloads
4,096 ctx
Ling-2.6-1T
Open-weight instruction model for adaptable chat and self-hosted production workloads
262,144 ctx
Ling-2.6-flash
Efficient model for low-latency assistance, extraction, and routine automation
262,144 ctx
Ring-2.6-1T
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
262,144 ctx
reasoning
Kat Coder Pro
Coding model for repository understanding, refactors, and agentic engineering tasks
256,000 ctx
Llama3 70B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
Llama 3 8B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
Llama 3.1 8B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
16,384 ctx
Llama 3.2 3B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
32,768 ctx
Llama 3.3 70B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
131,072 ctx
Llama 4 Maverick Instruct
Open multimodal Llama model for strong reasoning and fast responses
1,048,576 ctx
Llama 4 Scout Instruct
Open multimodal Llama model for long-context analysis and efficient agents
131,072 ctx
Wizardlm 2 8x22B
Open-weight instruction model for adaptable chat and self-hosted production workloads
65,535 ctx
MiniMax-M2
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
Minimax M2.1
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
MiniMax M2.5
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax M2.5 Highspeed
High-speed MiniMax model for low-latency coding and agent workflows
204,800 ctx
reasoning
MiniMax M2.7
MiniMax model for chat, coding, office work, and agentic tasks
204,800 ctx
reasoning
MiniMax-M2.7-highspeed
Low-latency M2.7 variant for interactive coding plans and agent loops
204,800 ctx
reasoning
MiniMax M1
MiniMax model for chat, coding, office work, and agentic tasks
1,000,000 ctx
reasoning
Mistral Nemo
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
60,288 ctx
Kimi K2 0905
Kimi model for long-context chat, coding, and agentic reasoning
262,144 ctx
Kimi K2 Instruct
Kimi model for long-context chat, coding, and agentic reasoning
131,072 ctx
Kimi K2 Thinking
Kimi reasoning model for long-horizon research, planning, and tool use
262,144 ctx
reasoning
Kimi K2.5
Kimi multimodal agent model for visual understanding, coding, and planning
262,144 ctx
reasoning
Kimi K2.6
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
262,144 ctx
reasoning
Kimi K2.7 Code
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
262,144 ctx
reasoning
Kimi K3
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
1,048,576 ctx
reasoning
Hermes 2 Pro Llama 3 8B
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
OpenAI GPT OSS 120B
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
OpenAI: GPT OSS 20B
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
131,072 ctx
reasoning
PaddleOCR-VL
Multimodal model for analyzing text, images, documents, and rich media
16,384 ctx
Qwen 2.5 72B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
32,000 ctx
Qwen MT Plus
Translation model for multilingual conversion, localization, and cross-language workflows
16,384 ctx
Qwen2.5 7B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
32,000 ctx
Qwen2.5 VL 72B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
32,768 ctx
Qwen3 235B A22B
Qwen instruction model for multilingual chat, reasoning, and tool use
40,960 ctx
reasoning
Qwen3 235B A22B Instruct 2507
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
Qwen3 235B A22b Thinking 2507
Qwen reasoning model for deliberate problem solving, math, and coding
131,072 ctx
reasoning
Qwen3 30B A3B
Qwen instruction model for multilingual chat, reasoning, and tool use
40,960 ctx
reasoning
Qwen3 32B
Qwen instruction model for multilingual chat, reasoning, and tool use
40,960 ctx
reasoning
Qwen3 4B
Qwen instruction model for multilingual chat, reasoning, and tool use
128,000 ctx
reasoning
Qwen3 8B
Qwen instruction model for multilingual chat, reasoning, and tool use
128,000 ctx
reasoning
Qwen3 Coder 30b A3B Instruct
Qwen coding model for software agents, repository edits, and code reasoning
160,000 ctx
Qwen3 Coder 480B A35B Instruct
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3 Coder Next
Qwen coding model for software agents, repository edits, and code reasoning
262,144 ctx
Qwen3 Max
Flagship Qwen model for complex reasoning, coding, and agentic workflows
262,144 ctx
Qwen3 Next 80B A3B Instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
131,072 ctx
Qwen3 Next 80B A3B Thinking
Qwen reasoning model for deliberate problem solving, math, and coding
131,072 ctx
reasoning
Qwen3 Omni 30B A3B Instruct
Qwen omni model for text, vision, audio, and multimodal agent tasks
65,536 ctx
Qwen3 Omni 30B A3B Thinking
Qwen omni model for text, vision, audio, and multimodal agent tasks
65,536 ctx
reasoning
Qwen3 VL 235B A22B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
Qwen3 VL 235B A22B Thinking
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
reasoning
qwen/qwen3-vl-30b-a3b-instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
qwen/qwen3-vl-30b-a3b-thinking
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
qwen/qwen3-vl-8b-instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
Qwen3.5-122B-A10B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5-27B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5-35B-A3B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.5-397B-A17B
Qwen vision-language model for visual reasoning, documents, and agent tasks
262,144 ctx
reasoning
Qwen3.7-Max
Flagship Qwen model for complex reasoning, coding, and agentic workflows
1,000,000 ctx
reasoning
L3 8B Stheno V3.2
Open Llama instruction model for multilingual chat, reasoning, and coding
8,192 ctx
L3 70B Euryale V2.1
Open-weight instruction model for adaptable chat and self-hosted production workloads
8,192 ctx
Sao10k L3 8B Lunaris
Open-weight instruction model for adaptable chat and self-hosted production workloads
8,192 ctx
L31 70B Euryale V2.2
Open-weight instruction model for adaptable chat and self-hosted production workloads
8,192 ctx
XiaomiMiMo/MiMo-V2-Flash
MiMo flash model for fast multimodal assistance and agent workflows
262,144 ctx
reasoning
MiMo-V2-Pro
Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks
1,048,576 ctx
reasoning
MiMo-V2.5-Pro
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
1,048,576 ctx
reasoning
AutoGLM-Phone-9B-Multilingual
GLM vision model for visual reasoning, documents, and multimodal agents
65,536 ctx
GLM-4.5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
131,072 ctx
reasoning
GLM 4.5 Air
Efficient GLM model for fast reasoning, coding, and agent workflows
131,072 ctx
reasoning
GLM 4.5V
GLM vision model for visual reasoning, documents, and multimodal agents
65,536 ctx
reasoning
GLM 4.6
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
204,800 ctx
reasoning
GLM 4.6V
GLM vision model for visual reasoning, documents, and multimodal agents
131,072 ctx
reasoning
GLM-4.7
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
204,800 ctx
reasoning
GLM-4.7-Flash
Efficient GLM model for fast reasoning, coding, and agent workflows
200,000 ctx
reasoning
GLM-5
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,800 ctx
reasoning
GLM-5.1
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
204,800 ctx
reasoning
GLM-5.2
Open flagship GLM for long-horizon coding agents and million-token context work
1,048,576 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate