Early build · scores computed by the deterministic engine from real, sourced developments (30 days) · weights v1.0-prior
Updated 06:00 UTC
next refresh in
Vikshy Models · the catalogue

Every model, plainly compared

140 models from 13 labs, with what each is for, how you reach it, its context window, and current API pricing, each linked to the official docs. Prices are per million tokens and can change, so the linked pricing page is always the source of truth. For how the providers rank, see The Board. Updated 2026-07-21.

ModelContextInput /MOutput /MModalityAccess
GPT-5.6 SolFrontier
Teams tackling the hardest agentic and coding work.
1M
$5.00
$30.00
Multimodal
API + ChatGPT
GPT-5.5Flagship
Broad production use across chat and tools.
1M
$5.00
$30.00
Multimodal
API + ChatGPT
GPT-5.5 ProReasoning
Research-grade analysis where accuracy beats cost.
1M
$30.00
$180.00
Multimodal
API + ChatGPT
o3Reasoning
Math, logic, and analysis that need deeper thinking.
200K
$2.00
$8.00
Multimodal
API + ChatGPT
GPT-5.6 TerraBalanced
Everyday production apps that need strong quality on a budget.
1M
$2.50
$15.00
Multimodal
API + ChatGPT
GPT-Realtime-2Balanced
Voice agents and live conversation. Prices are per 1M audio tokens.
32K
$32.00
$64.00
Audio
API
GPT-image-2Balanced
Product imagery, design, and creative work. Text in, image tokens out.
N/A
$5.00
$30.00
Image
API
GPT-5.6 LunaEfficient
Chat, classification, and other high-throughput jobs.
1M
$1.00
$6.00
Multimodal
API + ChatGPT
GPT-5.4 miniEfficient
Lightweight tasks and cost-sensitive apps.
1M
$0.75
$4.50
Multimodal
API + ChatGPT
GPT-5.4 nanoEfficient
Classification, extraction, and routing at scale.
1M
$0.20
$1.25
Multimodal
API
gpt-oss-120bOpen
Teams that want to self-host a capable model.
128K
Open
Open
Text
Open weights
gpt-oss-20bOpen
On-device and low-footprint deployments.
128K
Open
Open
Text
Open weights
GPT-5Legacy
Apps not yet moved to the 5.5 or 5.6 line.
400K
$1.25
$10.00
Multimodal
API + ChatGPT
GPT-4.1Legacy
Legacy integrations still built on GPT-4.
1M
$2.00
$8.00
Multimodal
API + ChatGPT
ModelContextInput /MOutput /MModalityAccess
Claude Fable 5Frontier
Ambitious, long-horizon agent work.
1M
$10.00
$50.00
Multimodal
API + Claude apps
Claude Opus 4.8Flagship
Complex agentic coding and enterprise work.
1M
$5.00
$25.00
Multimodal
API + Claude apps
Claude Sonnet 5Balanced
Most production workloads. Intro pricing of $2 in, $10 out runs through Aug 31, 2026.
1M
$3.00
$15.00
Multimodal
API + Claude apps
Claude Haiku 4.5Efficient
Latency-sensitive and high-volume tasks.
200K
$1.00
$5.00
Multimodal
API + Claude apps
Claude Mythos 5Enterprise
Approved security teams in Project Glasswing. Invitation only.
1M
$10.00
$50.00
Multimodal
Limited availability
Claude Opus 4.7Legacy
Apps not yet moved to Opus 4.8.
1M
$5.00
$25.00
Multimodal
API + Claude apps
Claude Opus 4.6Legacy
Older Opus integrations.
1M
$5.00
$25.00
Multimodal
API + Claude apps
Claude Opus 4.5Legacy
Legacy Opus workloads.
200K
$5.00
$25.00
Multimodal
API + Claude apps
Claude Sonnet 4.6Legacy
Apps not yet moved to Sonnet 5.
1M
$3.00
$15.00
Multimodal
API + Claude apps
Claude Sonnet 4.5Legacy
Older Sonnet integrations.
200K
$3.00
$15.00
Multimodal
API + Claude apps
Claude Opus 4.1Legacy
Migrate to Opus 4.8 before retirement.
200K
$15.00
$75.00
Multimodal
API + Claude apps
Claude Haiku 3.5Legacy
Legacy high-volume apps.
200K
$0.80
$4.00
Multimodal
API + Claude apps
ModelContextInput /MOutput /MModalityAccess
Gemini 3.1 ProFrontier
Complex analysis over very long documents. Rates rise above 200K tokens.
2M
$2.00
$12.00
Multimodal
API + Gemini app
Gemini 3.6 FlashBalanced
Fast, affordable agentic and coding work.
1M
$1.50
$7.50
Multimodal
API + Gemini app
Gemini 3.5 FlashBalanced
General production workloads.
1M
$1.50
$9.00
Multimodal
API + Gemini app
Gemini 3.5 Flash-LiteEfficient
Large-scale, latency-sensitive tasks.
1M
$0.30
$2.50
Multimodal
API
Gemini 3.5 Flash CyberEnterprise
Security teams. No public API pricing or self-serve access.
1M
Multimodal
Enterprise
Gemini 3.1 Flash-LiteLegacy
Simple, high-throughput jobs.
1M
$0.125
$0.75
Multimodal
API
Gemini 3 FlashLegacy
Apps not yet moved to 3.5 or 3.6.
1M
$0.50
$3.00
Multimodal
API + Gemini app
Gemini 2.5 ProLegacy
Older integrations still on the 2.5 line.
1M
$1.25
$10.00
Multimodal
API + Gemini app
Gemini 2.5 Flash-LiteLegacy
Cost-first, high-volume workloads.
1M
$0.10
$0.40
Multimodal
API
ModelContextInput /MOutput /MModalityAccess
DeepSeek-V4-ProFlagship
Hard reasoning, long documents and agentic tasks.
1M
$1.74
$3.48
Text
API + open weights
DeepSeek-V4-FlashBalanced
Everyday chat, coding and high volume production use.
1M
$0.14
$0.28
Text
API + open weights
deepseek-chatLegacy
Existing code that still calls the old chat name.
1M
$0.14
$0.28
Text
API
deepseek-reasonerLegacy
Existing code that still calls the old reasoner name.
1M
$0.14
$0.28
Text
API
DeepSeek-R1Legacy
Self hosted reasoning and research on open weights.
128K
Open
Open
Text
Open weights
DeepSeek-V3.1Legacy
Self hosting the earlier V3 line.
128K
Open
Open
Text
Open weights
ModelContextInput /MOutput /MModalityAccess
Qwen3-MaxFlagship
The hardest reasoning and long context workloads on Model Studio.
256K
$0.86
$3.44
Text
API
Qwen-PlusBalanced
General purpose production traffic.
1M
Text
API
Qwen-FlashEfficient
Simple tasks that need fast cheap responses.
1M
$0.05
$0.40
Text
API
Qwen3-Coder-480B-A35BOpen
Agentic coding, tool use and browser automation.
256K
Open
Open
Text
API + open weights
Qwen3-Coder-NextOpen
Lighter self hosted coding agents.
256K
Open
Open
Text
API + open weights
Qwen3-VLOpen
Multimodal document and UI understanding.
256K
Open
Open
Multimodal
API + open weights
Qwen3.6-27BOpen
Strong coding on a single accessible open model.
256K
Open
Open
Text
Open weights
Qwen3-235B-A22BOpen
Self hosted general reasoning at scale.
256K
Open
Open
Text
API + open weights
Qwen-TurboLegacy
Existing integrations that have not yet migrated to Flash.
1M
Text
API
Qwen2.5-MaxLegacy
Integrations pinned to the Qwen2.5 line.
32K
Text
API
ModelContextInput /MOutput /MModalityAccess
Llama 4 MaverickFlagship
General assistant and multimodal workloads run through third-party hosts or self-hosting.
1M
Open
Open
Multimodal
Open weights + third-party API
Llama 4 ScoutEfficient
Very long context tasks and efficient on device or single GPU deployment.
10M
Open
Open
Multimodal
Open weights + third-party API
Llama 3.3 70BLegacy
Teams wanting a proven mid size open text model.
128K
Open
Open
Text
Open weights + third-party API
Llama 3.2 90B VisionLegacy
Multimodal apps that need self hosted image plus text.
128K
Open
Open
Vision
Open weights + third-party API
Llama 3.2 11B VisionLegacy
Lighter multimodal workloads and edge image tasks.
128K
Open
Open
Vision
Open weights + third-party API
Llama 3.2 3BLegacy
Edge and mobile assistants where footprint matters.
128K
Open
Open
Text
Open weights + third-party API
Llama 3.1 405BLegacy
Research and high end self hosted deployments.
128K
Open
Open
Text
Open weights + third-party API
Llama 3.1 8BLegacy
Cost sensitive text tasks and fine tuning starters.
128K
Open
Open
Text
Open weights + third-party API
ModelContextInput /MOutput /MModalityAccess
Mistral Medium 3.5Frontier
Complex agent pipelines and premium coding assistants.
256K
$1.50
$7.50
Multimodal
API
Mistral Large 3Flagship
Teams wanting a top open weight generalist with a hosted option.
256K
$0.50
$1.50
Multimodal
API + open weights
Magistral MediumReasoning
Math, logic, and multi step problem solving.
128K
Text
API
Magistral SmallReasoning
Local reasoning and cost controlled deployments.
128K
Open
Open
Text
API + open weights
Mistral Small 4Balanced
High throughput budget tasks that still need long context.
256K
Text
API + open weights
Devstral MediumBalanced
Autonomous coding agents and repo scale edits.
256K
$0.40
$2.00
Text
API
CodestralEfficient
IDE autocomplete and high frequency code generation.
256K
$0.30
$0.90
Text
API
Ministral 3 14BEfficient
On device and latency sensitive text tasks.
128K
Open
Open
Text
API + open weights
Voxtral MiniEfficient
Voice interfaces and transcription pipelines.
32K
Open
Open
Audio
API + open weights
Devstral SmallOpen
Self hosted coding agents on a budget.
256K
$0.10
$0.30
Text
API + open weights
Pixtral 12BOpen
Image plus text apps that want open weights.
128K
$0.15
$0.15
Multimodal
API + open weights
Mistral Large 2Legacy
Workloads pinned to the older Large line.
128K
$2.00
$6.00
Text
API + open weights
ModelContextInput /MOutput /MModalityAccess
Grok 4.5Frontier
Hard coding and agent tasks. Cached input is $0.50 per 1M.
500K
$2.00
$6.00
Multimodal
API + Grok app
Grok 4.3Flagship
General production use. Canonical chat-model alias on the docs.
1M
$1.25
$2.50
Multimodal
API + Grok app
Grok 4.20Reasoning
Multi-agent and long-horizon jobs.
2M
$1.25
$2.50
Multimodal
API
Grok Build 0.1Balanced
Codebase-scale software work.
256K
$1.00
$2.00
Text
API + Grok app
Grok ImagineBalanced
Creative and media work.
N/A
Image
API + Grok app
Grok 4.1 FastEfficient
High-volume, cost-sensitive tasks.
2M
$0.20
$0.50
Multimodal
API
Grok 4Legacy
Older integrations still pointing at Grok 4.
256K
$3.00
$15.00
Multimodal
API
Grok 3Legacy
Legacy apps built on Grok 3.
131K
$3.00
$15.00
Text
API
ModelContextInput /MOutput /MModalityAccess
Command A+Frontier
Enterprises that need top quality with self hosting options.
128K
Text
API + open weights
Command AFlagship
Enterprise RAG, agents and general chat.
256K
$2.50
$10.00
Text
API + open weights
Command A ReasoningReasoning
Agentic workflows that need deliberate reasoning.
256K
Text
API
Command R7BEfficient
High volume lightweight generation.
128K
$0.0375
$0.15
Text
API + open weights
Command A VisionEnterprise
Enterprise multimodal document analysis.
128K
Vision
API
Rerank 3.5Enterprise
Reranking retrieval results in search pipelines.
N/A
$2.00 per 1K searches
N/A
Text
API
Embed 4Enterprise
Enterprise semantic search and multimodal retrieval.
128K
$0.12
N/A
Multimodal
API
Command R+Legacy
Existing RAG stacks built on the R+ line.
128K
$2.50
$10.00
Text
API + open weights
Command RLegacy
Cost sensitive RAG and chat.
128K
$0.15
$0.60
Text
API + open weights
KMoonshot AI (Kimi)12 models
ModelContextInput /MOutput /MModalityAccess
Kimi K3Frontier
Complex multi-step agents, deep research, and large-codebase work where the top tier matters.
1M
$3.00
$15.00
Text and vision
API (model id: kimi-k3) at api.moonshot.ai, plus kimi.com and the Kimi apps. Open weights planned on Hugging Face.
Kimi K2.6Flagship
Teams wanting a capable open-weight agentic model with generous context at a moderate price.
256K
$0.95
$4.00
Text and vision
API (model id: kimi-k2.6) plus open weights on Hugging Face (moonshotai/Kimi-K2.6).
Kimi K2 ThinkingReasoning
Math, planning, and step-by-step problems that benefit from deliberate reasoning.
256K
Text
API (model id: kimi-k2-thinking) plus open weights on Hugging Face (moonshotai/Kimi-K2-Thinking).
kimi-thinking-previewReasoning
Harder multimodal reasoning tasks where a visible thinking pass helps.
128K
Text and vision
API only (model id: kimi-thinking-preview). Hosted preview.
Kimi K2.5Balanced
Everyday agentic and coding work where you want good value on a proven open model.
256K
$0.60
$3.00
Text
API (model id: kimi-k2.5) plus open weights on Hugging Face (moonshotai/Kimi-K2.5).
kimi-latestBalanced
Apps that want the newest Kimi without pinning a version, including vision inputs.
128K
Text and vision
API only (model id: kimi-latest). Hosted, no open weights for the alias.
Kimi K2.7-CodeOpen
Agentic coding, refactors, and tool-driven development inside IDEs and CLI agents.
256K
$0.72
$3.50
Text
API (model id: kimi-k2.7-code) plus open weights on Hugging Face under a Modified MIT license.
Kimi K2 (Instruct)Open
Self-hosting and research on a widely adopted open MoE base.
128K
Open
Open
Text
API (model id: kimi-k2-0711-preview) plus open weights on Hugging Face (moonshotai/Kimi-K2-Instruct).
moonshot-v1-128kLegacy
Legacy integrations still calling the pre-K2 endpoints for long documents.
128K
$2.00
$5.00
Text
API only (model id: moonshot-v1-128k). Hosted.
moonshot-v1-32kLegacy
Older apps needing moderate context on the legacy endpoint.
32K
$1.00
$3.00
Text
API only (model id: moonshot-v1-32k). Hosted.
moonshot-v1-8kLegacy
Simple, cheap legacy calls that do not need long context.
8K
$0.20
$2.00
Text
API only (model id: moonshot-v1-8k). Hosted.
moonshot-v1-128k-vision-previewLegacy
Older multimodal integrations built before the Kimi K vision models.
128K
$2.00
$5.00
Text and vision
API only (model id: moonshot-v1-128k-vision-preview). Hosted preview. Also offered in 8k and 32k vision variants.
ZZhipu AI (GLM)17 models
ModelContextInput /MOutput /MModalityAccess
GLM-5.2Frontier
Frontier agentic coding and long-context reasoning at a fraction of closed-model cost.
1M
$1.40
$4.40
Text
API (model id: glm-5.2) via Z.ai, plus the GLM Coding Plan and open weights on Hugging Face.
GLM-4.5-XFlagship
Cases that need GLM-4.5's strongest, fastest responses and will pay for it.
131K
$2.20
$8.90
Text
API (model id: glm-4.5-x) via Z.ai.
GLM-5.1Reasoning
Reasoning and coding work on a recent GLM-5 checkpoint.
1M
Text
API (model id: glm-5.1) via Z.ai, plus open weights on Hugging Face.
GLM-5-TurboBalanced
Agent workflows that need speed and strong tool calling without the full flagship price.
200K
$1.20
$4.00
Text
API (model id: glm-5-turbo) via Z.ai and the GLM Coding Plan.
GLM-5V-TurboBalanced
Vision-driven coding and UI tasks in agent tools like OpenClaw and Claude Code.
200K
$1.20
$4.00
Text and vision
API (model id: glm-5v-turbo) via Z.ai.
GLM-4.6VBalanced
Multimodal reasoning and document or UI understanding with tool use.
128K
$0.30
$0.90
Text and vision
API (model id: glm-4.6v) via Z.ai, plus open weights on Hugging Face.
GLM-4.7-FlashEfficient
High-volume, latency-sensitive tasks and prototyping at no token cost.
128K
Free
Free
Text
API (model id: glm-4.7-flash) via Z.ai, free tier.
GLM-4.6V-FlashEfficient
On-device or high-volume multimodal tasks at no token cost.
128K
Free
Free
Text and vision
API (model id: glm-4.6v-flash) via Z.ai, free tier, plus open weights on Hugging Face.
GLM-4.5-AirEfficient
Budget-sensitive apps that still want GLM-4.5-class output.
131K
Text
API (model id: glm-4.5-air) via Z.ai, plus open weights on Hugging Face.
GLM-4.5-FlashEfficient
Prototyping and high-volume light tasks at no token cost.
131K
Free
Free
Text
API (model id: glm-4.5-flash) via Z.ai, free tier.
GLM-5Open
Self-hosting and research on Zhipu's frontier open base.
1M
Open
Open
Text
Open weights on Hugging Face (zai-org) plus API via Z.ai.
GLM-4.7Open
Cost-conscious coding and agent tasks that still want deep reasoning.
200K
$0.60
$2.20
Text
API (model id: glm-4.7) via Z.ai, plus open weights on Hugging Face (zai-org/GLM-4.7).
GLM-4.6Legacy
Solid general coding and reasoning where you want a proven, cheaper checkpoint.
200K
$0.43
$1.74
Text
API (model id: glm-4.6) via Z.ai, plus open weights on Hugging Face.
GLM-4.5Legacy
General coding and reasoning on a stable, widely supported checkpoint.
131K
$0.60
$2.20
Text
API (model id: glm-4.5) via Z.ai, plus open weights on Hugging Face (zai-org/GLM-4.5).
GLM-4.5VLegacy
Multimodal understanding on a proven, lower-cost checkpoint.
131K
$0.60
$1.20
Text and vision
API (model id: glm-4.5v) via Z.ai, plus open weights on Hugging Face.
GLM-4-PlusLegacy
Older integrations still pinned to the GLM-4 series.
128K
Text
API (model id: glm-4-plus) via the Zhipu open platform (bigmodel.cn).
GLM-4-9BLegacy
Self-hosting a compact, permissively usable GLM base.
128K
Open
Open
Text
Open weights on Hugging Face (zai-org/glm-4-9b).
MMiniMax11 models
ModelContextInput /MOutput /MModalityAccess
MiniMax-M3Frontier
Frontier coding, agentic execution, and multimodal reasoning at low cost.
1M
$0.30
$1.20
Text, image, and video
API (model id: MiniMax-M3) via platform.minimax.io, plus open weights on Hugging Face.
MiniMax-M1Reasoning
Long-context reasoning and software tasks that need a big open thinking model.
1M
$0.40
$2.20
Text
API (model id: MiniMax-M1) via platform.minimax.io, plus open weights on GitHub and Hugging Face (Apache 2.0).
MiniMax-M2.5Balanced
High-volume coding and agent work where value per token is the priority.
204K
$0.15
$0.90
Text
API (model id: MiniMax-M2.5) via platform.minimax.io.
MiniMax-M2.1Open
Self-hosting or API use where you want an updated open M2-class model.
204K
$0.30
$1.20
Text
API (model id: MiniMax-M2.1) via platform.minimax.io, plus open weights on Hugging Face (MiniMaxAI/MiniMax-M2.1).
MiniMax-M2Open
Cost-efficient agentic coding at scale on an Apache 2.0 open model.
204K
$0.30
$1.20
Text
API (model id: MiniMax-M2) via platform.minimax.io, plus open weights on Hugging Face and GitHub (Apache 2.0).
MiniMax-Text-01Open
Extreme long-context research and self-hosting on an open linear-attention model.
4M
Open
Open
Text
API (model id: MiniMax-Text-01) via platform.minimax.io, plus open weights on Hugging Face and GitHub.
MiniMax-VL-01Open
Open multimodal research and self-hosting alongside MiniMax-Text-01.
1M
Open
Open
Text and vision
Open weights on Hugging Face and GitHub (MiniMax-AI/MiniMax-01).
Hailuo 02Enterprise
Text-to-video and image-to-video creation for creators and products.
N/A
Video generation
API (Hailuo video models) via platform.minimax.io and the MiniMax MCP server. Billed per generation.
Speech 2.5Enterprise
Voiceover, agents, and apps needing natural multilingual speech.
N/A
Text to speech
API (Speech 2.5 and Speech-02 series) via platform.minimax.io and the MiniMax MCP server. Billed per character or generation.
Music 2.6Enterprise
Music creation for content, games, and prototyping.
N/A
Text to music
API (Music series) via platform.minimax.io and the MiniMax MCP server. Billed per generation.
abab6.5sLegacy
Older MiniMax integrations still on the abab chat endpoints.
200K
Text
API (model id: abab6.5s-chat) via platform.minimax.io.
TTencent (Hunyuan)12 models
ModelContextInput /MOutput /MModalityAccess
Hunyuan Hy3Frontier
Cost-efficient frontier reasoning and agent work, positioned against other top open models.
256K
$0.14
$0.58
Text
API (model id: hy3) via Tencent Cloud TokenHub, plus open weights on Hugging Face and GitHub. Also in Yuanbao, CodeBuddy, and ima.
Hunyuan-TurboSFlagship
High-throughput general chat and assistant tasks that need speed and price efficiency.
256K
$0.11
$0.28
Text
API (model id: hunyuan-turbos) via Tencent Cloud.
Hunyuan-T1Reasoning
Deliberate reasoning tasks in math, code, and analysis.
64K
Text
API (model id: hunyuan-t1) via Tencent Cloud.
Hunyuan-StandardBalanced
General assistant and content work that needs long context at a moderate price.
256K
Text
API (model ids: hunyuan-standard, hunyuan-standard-256K) via Tencent Cloud.
Hunyuan-VisionBalanced
Multimodal understanding of documents, charts, and photos.
32K
Text and vision
API (model id: hunyuan-vision) via Tencent Cloud.
Hunyuan-LiteEfficient
Prototyping and high-volume light workloads at no token cost.
256K
Free
Free
Text
API (model id: hunyuan-lite) via Tencent Cloud, free tier.
Hunyuan-LargeOpen
Self-hosting and research on a large open Tencent base.
256K
Open
Open
Text
Open weights on Hugging Face (tencent/Tencent-Hunyuan-Large) and GitHub, plus API via Tencent Cloud.
Hunyuan-A13BOpen
Local deployment and cost-sensitive inference on an open model.
256K
Open
Open
Text
Open weights on Hugging Face and GitHub (Tencent-Hunyuan), plus API via Tencent Cloud.
HunyuanImage 3.0Open
Open text-to-image generation and research.
N/A
Open
Open
Image generation
Open weights on GitHub and Hugging Face (Tencent-Hunyuan/HunyuanImage-3.0), plus Tencent Cloud image APIs.
HunyuanVideoOpen
Open text-to-video and image-to-video generation and research.
N/A
Open
Open
Video generation
Open weights on Hugging Face (tencent/HunyuanVideo) and GitHub, plus Tencent Cloud video APIs.
Hunyuan3D 2.5Open
Game, product, and design workflows that need open 3D asset generation.
N/A
Open
Open
Text and image to 3D
Open weights on Hugging Face and GitHub (Tencent-Hunyuan), plus Tencent Cloud 3D APIs.
Hunyuan-TurboLegacy
Older integrations still calling the earlier Turbo endpoint.
256K
Text
API (model id: hunyuan-turbo) via Tencent Cloud.