Viberank
Trending models
Popular Labs
Hottest providers
Top Agents
Login
Sign up
← All providers
NEAR AI Cloud
Provider ID: nearai
Details
SDK package
@ai-sdk/openai-compatible
Models
37
API
https://cloud-api.near.ai/v1
Links
Documentation
Auth: NEARAI_API_KEY
Models (37)
Qwen3 30B-A3B Instruct 2507
Qwen instruction model for multilingual chat, reasoning, and tool use
262,144 ctx
Qwen3 Embedding 0.6B
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
40,960 ctx
Qwen3 Reranker 0.6B
Reranking model for improving retrieval quality in search and recommendation systems
40,960 ctx
Qwen3-VL 30B-A3B Instruct
Qwen vision-language model for visual reasoning, documents, and agent tasks
256,000 ctx
Qwen3.5 122B-A10B
Qwen vision-language model for visual reasoning, documents, and agent tasks
131,072 ctx
reasoning
Qwen 3.6 35B A3B FP8
Open multimodal Qwen MoE for local agents that need vision, audio, and code
262,144 ctx
reasoning
Claude Haiku 4.5 (latest)
Fast Claude lane for lightweight agents, office tasks, and responsive chat
200,000 ctx
reasoning
Claude Opus 4.6
High-end Claude for difficult coding, planning, and slower expert reasoning
200,000 ctx
reasoning
Claude Opus 4.7
Stronger Opus tier for advanced software work and high-stakes reasoning
1,000,000 ctx
reasoning
Claude Sonnet 4.5 (latest)
Balanced Claude model for coding, analysis, agent workflows, and cost control
200,000 ctx
reasoning
Claude Sonnet 4.6
Claude workhorse for coding agents, careful analysis, and production cost control
1,000,000 ctx
reasoning
FLUX.2 Klein 4B
Image model for prompt-driven generation, editing, and visual design workflows
128,000 ctx
Gemini 2.5 Flash
Fast Gemini workhorse for multimodal apps where latency and price matter
1,048,576 ctx
reasoning
Gemini 2.5 Flash-Lite
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
1,048,576 ctx
reasoning
Gemini 2.5 Pro
Google's proven reasoning model for coding, math, and multimodal analysis
1,048,576 ctx
reasoning
Gemini 3 Pro Preview
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
1,048,576 ctx
reasoning
Gemini 3.1 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
1,048,576 ctx
reasoning
Gemini 3.5 Flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
1,048,576 ctx
reasoning
Gemma 4 31B IT
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
262,144 ctx
reasoning
GPT-4.1
Long-lived GPT workhorse for coding, instruction following, and production apps
1,047,576 ctx
GPT-4.1 mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
1,047,576 ctx
GPT-4.1 nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
1,047,576 ctx
GPT-5
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
400,000 ctx
reasoning
GPT-5 Mini
Small GPT-5 for responsive agents, coding help, and everyday automation
400,000 ctx
reasoning
GPT-5 Nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
400,000 ctx
reasoning
GPT-5.1
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
400,000 ctx
reasoning
GPT-5.2
Reliable GPT generation for broad coding, writing, and tool-assisted product work
400,000 ctx
reasoning
GPT-5.4
Agent-ready GPT for coding and computer-use workflows at a lower cost
1,050,000 ctx
reasoning
GPT-5.4 mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
400,000 ctx
reasoning
GPT-5.4 nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
400,000 ctx
reasoning
GPT-5.5
Default frontier GPT for coding, computer use, research, and knowledge work
1,050,000 ctx
reasoning
GPT-OSS 120B
Open GPT reasoning model for self-hosted agents and controllable deployments
131,000 ctx
reasoning
o3
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
200,000 ctx
reasoning
o3-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
200,000 ctx
reasoning
o4-mini
Fast o-series model for compact reasoning, coding, and tool use
200,000 ctx
reasoning
Whisper Large v3
Speech transcription model for accurate audio-to-text and captioning workflows
448 ctx
GLM-5.1 FP8
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
202,752 ctx
reasoning
Comments (0)
Sign in
to leave a comment
No comments yet.
Community Rating
—
No ratings yet
Sign in
to rate