Technical specifications, pricing, context window, and capabilities of 771 AI models from 108 companies. Sheets updated weekly with manufacturer data.
For performance ranking, visit the AI Benchmark.
771
Models
108
Companies
131
Open Source
219
Multimodal
| Model | OS |
|---|---|
| LFM2.5-2.6B 0 | — |
| Model | OS |
|---|---|
| AI21: Jamba Large 1.7 256K tokens·$2.00/1M | ✅ |
| Jamba 1.5 Large $2.00/1M | — |
| Jamba 1.5 Mini $0.20/1M | — |
| Jamba 1.6 Large $2.00/1M | — |
| Jamba 1.6 Mini $0.20/1M | — |
| Jamba 1.7 Mini 0 | — |
| Jamba Reasoning 3B 0 | — |
| Model | OS |
|---|---|
| G9v3-39A5B 0 | — |
| G9v3-3B 0 | — |
| Model | OS |
|---|---|
| Adobe Image5 | — |
| Model | OS |
|---|---|
| AionLabs: Aion-1.0 131K tokens·$4.00/1M | — |
| AionLabs: Aion-2.0 131K tokens·$0.80/1M | — |
| AionLabs: Aion-RP 1.0 (8B) 33K tokens·$0.80/1M | — |
| Model | OS |
|---|---|
| AlfredPros: CodeLLaMa 7B Instruct Solidity 4K tokens·$0.80/1M | ✅ |
| Model | OS |
|---|---|
| Llama 3.1 Tulu3 405B 0 | — |
| Molmo 7B-D 0 | — |
| Molmo2-8B 0 | — |
| OLMo 2 32B 0 | — |
| OLMo 2 7B 0 | — |
| Olmo 3 7B Instruct $0.10/1M | — |
| Olmo 3 7B Think 0 | — |
| Olmo 3.1 32B Think 0 | — |
| Model | OS |
|---|---|
| Olmo 3 32B Think 66K tokens00 | ✅ |
| Olmo 3.1 32B Instruct 66K tokens00 | ✅ |
| Model | OS |
|---|---|
| Amazon: Nova 2 Lite 1.0M tokens·$0.30/1M | — |
| Amazon: Nova Lite 1.0 300K tokens·$0.06/1M | — |
| Amazon: Nova Micro 1.0 128K tokens·$0.04/1M | — |
| Amazon: Nova Pro 1.0 300K tokens·$0.80/1M | — |
| Nova 2.0 Lite (high) $0.30/1M | — |
| Nova 2.0 Omni (low) $0.30/1M | — |
| Nova 2.0 Omni (medium) $0.30/1M | — |
| Nova 2.0 Omni (Non-reasoning) $0.30/1M | — |
| Nova 2.0 Pro Preview (medium) $1.25/1M | — |
| Nova Premier $2.50/1M | — |
| Model | OS |
|---|---|
| Ling-3.0-flash $0.07/1M | — |
| Model | OS |
|---|---|
| Apodex 1.1 $0.30/1M | — |
| Model | OS |
|---|---|
| Arcee AI: Coder Large 33K tokens·$0.50/1M | — |
| Arcee AI: Maestro Reasoning 131K tokens·$0.90/1M | — |
| Arcee AI: Spotlight 131K tokens·$0.18/1M | — |
| Arcee AI: Trinity Large Thinking 262K tokens·$0.23/1M | ✅ |
| Arcee AI: Trinity Mini 131K tokens·$0.04/1M | ✅ |
| Arcee AI: Virtuoso Large 131K tokens·$0.75/1M | — |
| Model | OS |
|---|---|
| Baidu: ERNIE 4.5 21B A3B Thinking 131K tokens·$0.07/1M | ✅ |
| Baidu: ERNIE 4.5 300B A47B 123K tokens·$0.28/1M | ✅ |
| Baidu: ERNIE 4.5 VL 28B A3B 30K tokens·$0.14/1M | ✅ |
| Baidu: ERNIE 4.5 VL 424B A47B 123K tokens·$0.42/1M | ✅ |
| ERNIE 5.0 Thinking Preview 0 | — |
| Model | OS |
|---|---|
| FLUX1.1 [pro] | — |
| Model | OS |
|---|---|
| ByteDance: UI-TARS 7B 128K tokens·$0.10/1M | ✅ |
| Doubao Seed Code 0 | — |
| Model | OS |
|---|---|
| ByteDance Seed: Seed 1.6 Flash 262K tokens·$0.07/1M | — |
| ByteDance Seed: Seed-2.0-Lite 262K tokens·$0.25/1M | — |
| Seed-OSS-36B-Instruct $0.21/1M | — |
| Model | OS |
|---|---|
| JT-35B-Flash 0 | — |
| JT-4.1 Flash 236B A21B 0 | — |
| JT-MINI 0 | — |
| Model | OS |
|---|---|
| Devin | — |
| Model | OS |
|---|---|
| Cohere: Command A 256K tokens·$2.50/1M | — |
| Cohere: Command R+ (08-2024) 128K tokens·$2.50/1M | — |
| command | — |
| Command A+ 0 | — |
| Command-R (Mar '24) $0.50/1M | — |
| Command-R+ (Apr '24) $3.00/1M | — |
| North Mini Code 0 | — |
| Tiny Aya Global 0 | — |
| Model | OS |
|---|---|
| DBRX Instruct 0 | — |
| Model | OS |
|---|---|
| Cogito v2.1 (Reasoning) $1.25/1M | — |
| Deep Cogito: Cogito v2.1 671B 128K tokens·$1.25/1M | — |
| Model | OS |
|---|---|
| Eleven v3 | — |
| ElevenAgents | — |
| Model | OS |
|---|---|
| EssentialAI: Rnj 1 Instruct 33K tokens·$0.15/1M | ✅ |
| Model | OS |
|---|---|
| Goliath 120B 6K tokens·$3.75/1M | ✅ |
| Model | OS |
|---|---|
| Granite 3.3 8B (Non-reasoning) $0.03/1M | — |
| Granite 4.0 1B 0 | — |
| Granite 4.0 350M 0 | — |
| Granite 4.0 H 1B 0 | — |
| Granite 4.0 H 350M 0 | — |
| Granite 4.0 H Small $0.06/1M | — |
| Granite 4.0 Micro 131K tokens00 | ✅ |
| Granite 4.1 30B 0 | — |
| Granite 4.1 3B 0 | — |
| Granite 4.1 8B $0.05/1M | — |
| Granite 4.2 30B $0.16/1M | — |
| Granite 4.2 3B $0.03/1M | — |
| Granite 4.2 8B $0.06/1M | — |
| Model | OS |
|---|---|
| Ideogram 4.0 | ✅ |
| Model | OS |
|---|---|
| Inception: Mercury 2 128K tokens·$0.25/1M | — |
| Model | OS |
|---|---|
| Ling 2.6 Flash $0.10/1M | — |
| Ling-2.6-1T $0.30/1M | — |
| Model | OS |
|---|---|
| Ling 3.0 Tiny 0 | — |
| Ling-1T 0 | — |
| Ling-flash-2.0 $0.14/1M | — |
| Ling-mini-2.0 0 | — |
| Ring-1T 0 | — |
| Ring-2.6-1T $0.30/1M | — |
| Ring-flash-2.0 $0.14/1M | — |
| Model | OS |
|---|---|
| Inflection: Inflection 3 Pi 8K tokens·$2.50/1M | — |
| Inflection: Inflection 3 Productivity 8K tokens·$2.50/1M | — |
| Model | OS |
|---|---|
| Kimi K2 Thinking 262K tokens·$0.60/1M | — |
| Kimi K2.7 Code $0.95/1M | — |
| Kimi Linear 48B A3B Instruct 0 | — |
| Model | OS |
|---|---|
| Kling 2.1 | — |
| Model | OS |
|---|---|
| Mi:dm K 2.5 Pro 0 | — |
| Mi:dm K 2.5 Pro Preview 0 | — |
| Model | OS |
|---|---|
| Kling AI 2.0 | — |
| Model | OS |
|---|---|
| KAT-Coder-Pro V1 0 | — |
| Model | OS |
|---|---|
| Kwaipilot: KAT-Coder-Pro V2 256K tokens·$0.30/1M | — |
| Model | OS |
|---|---|
| EXAONE 4.5 33B 0 | — |
| Model | OS |
|---|---|
| LFM2-24B-A2B 33K tokens00 | ✅ |
| Model | OS |
|---|---|
| LongCat 2.0 $0.75/1M | — |
| LongCat Flash Lite 0 | — |
| Model | OS |
|---|---|
| Luma Dream Machine 1.6 | — |
| Luma Ray3.2 | — |
| Model | OS |
|---|---|
| K2 Think V2 0 | — |
| K2-V2 (medium) 0 | — |
| Model | OS |
|---|---|
| Magnum v4 72B 16K tokens·$3.00/1M | ✅ |
| Model | OS |
|---|---|
| Mancer: Weaver (alpha) 8K tokens·$0.75/1M | — |
| Model | OS |
|---|---|
| Llama 2 Chat 13B 0 | — |
| Llama 2 Chat 70B 0 | — |
| Llama 2 Chat 7B $0.05/1M | — |
| Llama 3 70B Instruct 8K tokens·$0.65/1M | ✅ |
| Llama 3 8B Instruct 8K tokens·$0.04/1M | ✅ |
| Llama 3.1 70B Instruct 131K tokens·$0.56/1M | ✅ |
| Llama 3.1 8B Instruct 16K tokens·$0.02/1M | ✅ |
| Llama 3.1 Instruct 405B $2.50/1M | — |
| Llama 3.2 11B Vision Instruct 131K tokens·$0.34/1M | ✅ |
| Llama 3.2 1B Instruct 60K tokens00 | ✅ |
| Llama 3.2 3B Instruct 80K tokens00 | ✅ |
| Llama 3.2 Instruct 90B (Vision) 0 | — |
| Llama 3.3 70B Instruct 131K tokens·$0.66/1M | ✅ |
| Llama 4 Maverick 1.0M tokens·$0.26/1M | ✅ |
| Llama 4 Scout 1.3M tokens·$0.18/1M | ✅ |
| Llama 65B 0 | — |
| Llama Guard 3 8B 131K tokens·$0.48/1M | ✅ |
| Llama Guard 4 12B 164K tokens·$0.18/1M | ✅ |
| Muse Glimmer (high) $0.32/1M | — |
| Muse Spark 0 | — |
| Muse Spark 1.1 $1.25/1M | — |
| Muse Spark 1.1 (xhigh) $1.25/1M | — |
| Muse Spark 1.2 (xhigh) $1.25/1M | — |
| Model | OS |
|---|---|
| Microsoft: Phi 4 16K tokens·$0.13/1M | ✅ |
| Phi-3 Mini Instruct 3.8B 0 | — |
| Phi-4 Mini Instruct 0 | — |
| Phi-4 Multimodal Instruct 0 | — |
| WizardLM-2 8x22B 66K tokens·$0.62/1M | ✅ |
| Model | OS |
|---|---|
| Midjourney V8.1 | — |
| Model | OS |
|---|---|
| Hailuo 2.3 | — |
| Hailuo MiniMax Video-01 | — |
| MiniMax M1 40k 0 | — |
| MiniMax M1 80k $0.55/1M | — |
| MiniMax-M2 205K tokens·$0.30/1M | — |
| MiniMax-M3 1.0M tokens·$0.30/1M | — |
| MiniMax: MiniMax M1 1.0M tokens·$0.40/1M | — |
| MiniMax: MiniMax M2-her 66K tokens·$0.30/1M | — |
| MiniMax: MiniMax M2.1 197K tokens·$0.30/1M | ✅ |
| MiniMax: MiniMax M2.5 197K tokens·$0.30/1M | ✅ |
| MiniMax: MiniMax M2.7 197K tokens·$0.30/1M | ✅ |
| MiniMax: MiniMax-01 1.0M tokens·$0.20/1M | ✅ |
| Music 2.6 | — |
| Speech 2.8 | — |
| Model | OS |
|---|---|
| Inkling-Small 524K tokens·$0.30/1M | — |
| Model | OS |
|---|---|
| Devstral 2 0 | — |
| Devstral Small (Jul '25) 131K tokens00 | — |
| Devstral Small (May '25) 0 | — |
| Devstral Small 2 0 | — |
| Magistral Medium 1 0 | — |
| Magistral Small 1 0 | — |
| Magistral Small 1.2 $0.50/1M | — |
| Ministral 3 14B $0.20/1M | — |
| Ministral 3 3B $0.10/1M | — |
| Ministral 3 8B $0.15/1M | — |
| Mistral 7B Instruct $0.25/1M | — |
| Mistral Large 2 (Jul '24) 131K tokens·$2.00/1M | — |
| Mistral Large 2 (Nov '24) $4.00/1M | — |
| Mistral Large 3 $0.50/1M | — |
| Mistral Medium $1.50/1M | — |
| Mistral Small (Feb '24) $0.15/1M | — |
| Mistral Small (Sep '24) $0.20/1M | — |
| Mistral Small 3 $0.10/1M | — |
| Mistral Small 3.1 $0.10/1M | — |
| Mistral Small 3.2 $0.10/1M | — |
| Mixtral 8x22B Instruct 0 | — |
| Model | OS |
|---|---|
| Magistral Medium 1.2 $2.00/1M | — |
| Mistral | — |
| Mistral Large 128K tokens·$2.00/1M | ✅ |
| Mistral: Codestral 2508 256K tokens·$0.30/1M | — |
| Mistral: Codestral 2508 (batch) 256K tokens·$0.30/1M | ✅ |
| Mistral: Devstral Medium 131K tokens00 | ✅ |
| Mistral: Ministral 3 8B 2512 (batch) 262K tokens·$0.15/1M | ✅ |
| Mistral: Ministral 8B 128K tokens·$0.11/1M | ✅ |
| Mistral: Mistral Large 3 2512 (batch) 262K tokens·$0.50/1M | ✅ |
| Mistral: Mistral Medium 3 131K tokens·$0.40/1M | ✅ |
| Mistral: Mistral Medium 3.1 131K tokens·$0.40/1M | ✅ |
| Mistral: Mistral Medium 3.1 (batch) 131K tokens·$0.40/1M | ✅ |
| Mistral: Mistral Medium 3.5 262K tokens·$1.50/1M | ✅ |
| Mistral: Mistral Medium 3.5 (batch) 262K tokens·$0.75/1M | ✅ |
| Mistral: Mistral Nemo 131K tokens·$0.02/1M | ✅ |
| Mistral: Mistral Small 4 262K tokens·$0.15/1M | ✅ |
| Mistral: Mistral Small 4 (batch) 262K tokens·$0.15/1M | ✅ |
| Mistral: Mixtral 8x22B Instruct 66K tokens·$2.00/1M | ✅ |
| Mistral: Mixtral 8x7B Instruct 33K tokens·$0.45/1M | ✅ |
| Mistral: Pixtral Large 2411 131K tokens00 | — |
| Mistral: Saba 33K tokens·$0.20/1M | ✅ |
| Mistral: Voxtral Small 24B 2507 32K tokens·$0.10/1M | ✅ |
| Model | OS |
|---|---|
| MoonshotAI: Kimi K2 0711 131K tokens·$0.57/1M | ✅ |
| MoonshotAI: Kimi K2 0905 262K tokens·$0.60/1M | ✅ |
| MoonshotAI: Kimi K2.5 262K tokens·$0.60/1M | ✅ |
| MoonshotAI: Kimi K2.6 262K tokens·$0.95/1M | ✅ |
| Model | OS |
|---|---|
| Morph: Morph V3 Fast 82K tokens·$0.80/1M | — |
| Morph: Morph V3 Large 262K tokens·$0.90/1M | — |
| Model | OS |
|---|---|
| Motif 3 | — |
| Motif 3 (Beta) 0 | — |
| Motif-2-12.7B-Reasoning 0 | — |
| Model | OS |
|---|---|
| HyperNova 60B 2605 $0.04/1M | — |
| Model | OS |
|---|---|
| MythoMax 13B 4K tokens·$0.06/1M | ✅ |
| Model | OS |
|---|---|
| Nemotron 3.5 Lightning 262K tokens·$0.07/1M | — |
| Model | OS |
|---|---|
| Nanbeige4.1-3B 0 | — |
| Model | OS |
|---|---|
| HyperCLOVA X SEED Think (32B) 0 | — |
| Model | OS |
|---|---|
| Nex AGI: DeepSeek V3.1 Nex N1 131K tokens·$0.14/1M | ✅ |
| Nex-N2-Pro 262K tokens·$0.50/1M | — |
| Model | OS |
|---|---|
| Nous: Hermes 3 405B Instruct 131K tokens·$1.00/1M | ✅ |
| Nous: Hermes 3 70B Instruct 131K tokens·$0.30/1M | ✅ |
| Nous: Hermes 4 405B 131K tokens·$1.00/1M | ✅ |
| Nous: Hermes 4 70B 131K tokens·$0.13/1M | ✅ |
| Model | OS |
|---|---|
| NousResearch: Hermes 2 Pro - Llama-3 8B 8K tokens·$0.14/1M | ✅ |
| Model | OS |
|---|---|
| OpenChat 3.5 (1210) 0 | — |
| Model | OS |
|---|---|
| Perplexity: Sonar Deep Research 128K tokens·$2.00/1M | — |
| Perplexity: Sonar Pro Search 200K tokens·$3.00/1M | — |
| R1 1776 0 | — |
| Sonar 127K tokens00 | — |
| Sonar Reasoning 127K tokens00 | — |
| Sonar Reasoning Pro 128K tokens00 | — |
| Model | OS |
|---|---|
| Pika 2.5 | — |
| Model | OS |
|---|---|
| Pika 2.1 | — |
| Model | OS |
|---|---|
| INTELLECT-3 131K tokens00 | ✅ |
| Model | OS |
|---|---|
| ReMM SLERP 13B 6K tokens·$0.45/1M | ✅ |
| Model | OS |
|---|---|
| Recraft V4.1 | — |
| Model | OS |
|---|---|
| Reka Edge 16K tokens·$0.10/1M | ✅ |
| Model | OS |
|---|---|
| Reka Flash 3 66K tokens·$0.20/1M | ✅ |
| Model | OS |
|---|---|
| Relace: Relace Apply 3 256K tokens·$0.85/1M | — |
| Relace: Relace Search 256K tokens·$1.00/1M | — |
| Model | OS |
|---|---|
| Runway Gen-3 Alpha | — |
| Runway Gen-4.5 | — |
| Model | OS |
|---|---|
| A.X-K2 | — |
| Model | OS |
|---|---|
| Sao10K: Llama 3 8B Lunaris 8K tokens·$0.04/1M | ✅ |
| Sao10k: Llama 3 Euryale 70B v2.1 8K tokens·$1.48/1M | ✅ |
| Sao10K: Llama 3.1 70B Hanami x1 16K tokens·$3.00/1M | ✅ |
| Sao10K: Llama 3.1 Euryale 70B v2.2 131K tokens·$0.85/1M | ✅ |
| Sao10K: Llama 3.3 Euryale 70B 131K tokens·$0.65/1M | ✅ |
| Model | OS |
|---|---|
| Agnes 2.5 Pro Alpha $0.45/1M | — |
| Agnes 2.5 Pro Beta $0.10/1M | — |
| Model | OS |
|---|---|
| Sarvam 105B (high) $0.04/1M | — |
| Sarvam 30B $0.03/1M | — |
| Sarvam M (Reasoning) 0 | — |
| Model | OS |
|---|---|
| Arctic Instruct 0 | — |
| Model | OS |
|---|---|
| Grok 4.5 $2.00/1M | — |
| Model | OS |
|---|---|
| Step 3.5 Flash 262K tokens·$0.10/1M | ✅ |
| Step 3.7 Flash $0.20/1M | — |
| Step3 VL 10B 0 | — |
| Model | OS |
|---|---|
| Suno v4.5 | — |
| Model | OS |
|---|---|
| Apertus 70B Instruct $0.82/1M | — |
| Apertus 8B Instruct $0.10/1M | — |
| Model | OS |
|---|---|
| Falcon-H1R-7B 0 | — |
| Model | OS |
|---|---|
| Hy3-preview (Non-reasoning) 262K tokens·$0.06/1M | — |
| Hy3-preview (Reasoning) 262K tokens·$0.06/1M | — |
| Tencent: Hunyuan A13B Instruct 131K tokens·$0.14/1M | ✅ |
| Model | OS |
|---|---|
| TheDrummer: Cydonia 24B V4.1 131K tokens·$0.30/1M | ✅ |
| TheDrummer: Rocinante 12B 33K tokens·$0.17/1M | ✅ |
| TheDrummer: Skyfall 36B V2 33K tokens·$0.55/1M | ✅ |
| TheDrummer: UnslopNemo 12B 33K tokens·$0.40/1M | ✅ |
| Model | OS |
|---|---|
| Inkling $1.00/1M | — |
| Inkling Small $0.30/1M | — |
| Model | OS |
|---|---|
| Inkling 1.0M tokens·$1.00/1M | — |
| Model | OS |
|---|---|
| Tongyi DeepResearch 30B A3B 131K tokens·$0.09/1M | ✅ |
| Model | OS |
|---|---|
| Tri-21B-Think 0 | — |
| Tri-21B-think Preview 0 | — |
| Model | OS |
|---|---|
| Solar Mini $0.15/1M | — |
| Solar Open 100B (Reasoning) 0 | — |
| Solar Open2 250B | — |
| Solar Pro 2 (Non-reasoning) 0 | — |
| Solar Pro 2 (Preview) (Non-reasoning) 0 | — |
| Solar Pro 2 (Preview) (Reasoning) 0 | — |
| Solar Pro 3 128K tokens·$0.15/1M | — |
| Solar Pro 4 524K tokens·$0.30/1M | — |
| Model | OS |
|---|---|
| Writer: Palmyra X5 1.0M tokens·$0.60/1M | — |
| Model | OS |
|---|---|
| MiMo-V2-Flash (Feb 2026) 0 | — |
| MiMo-V2-Flash (Reasoning) 262K tokens·$0.10/1M | — |
| MiMo-V2-Omni-0327 0 | — |
| MiMo-V2.5 $0.14/1M | — |
| Xiaomi: MiMo-V2-Omni 262K tokens00 | — |
| Xiaomi: MiMo-V2-Pro 1.0M tokens00 | — |
| Xiaomi: MiMo-V2.5-Pro 1.0M tokens·$0.43/1M | — |
| Model | OS |
|---|---|
| GLM-5.3-Flash $0.15/1M | — |
| Model | OS |
|---|---|
| GLM-4.5 (Reasoning) 131K tokens00 | — |
| GLM-4.5V (Reasoning) $0.60/1M | — |
| GLM-4.6 (Reasoning) $0.57/1M | — |
| GLM-4.6V (Reasoning) $0.30/1M | — |
| GLM-4.7 (Reasoning) $0.60/1M | — |
| GLM-5 (Non-reasoning) 205K tokens·$1.00/1M | — |
| GLM-5-Turbo 203K tokens00 | — |
| GLM-5.1 (Non-reasoning) $1.38/1M | — |
| GLM-5.2 (max) $1.40/1M | — |
| GLM-5.3 $1.40/1M | — |
| Z.ai: GLM 4 32B 128K tokens·$0.10/1M | — |
| Z.ai: GLM 4.5 Air 131K tokens·$0.17/1M | ✅ |
| Z.ai: GLM 4.7 Flash 203K tokens·$0.07/1M | ✅ |
| Z.ai: GLM 5.1 203K tokens·$1.28/1M | ✅ |
| Z.ai: GLM 5V Turbo 203K tokens·$1.20/1M | — |
| Model | OS |
|---|---|
| Grok 2 (Dec '24) 0 | — |
| Grok 3 131K tokens·$4.00/1M | — |
| Grok 3 Beta 131K tokens·$3.00/1M | — |
| Grok 3 Mini 131K tokens·$0.30/1M | — |
| Grok 3 Mini Beta 131K tokens·$0.30/1M | — |
| Grok 4 256K tokens·$3.00/1M | — |
| Grok 4 Fast 2.0M tokens·$0.20/1M | — |
| Grok 4.1 Fast 2.0M tokens00 | — |
| Grok 4.20 0309 (Reasoning) $2.00/1M | — |
| Grok 4.20 Multi-Agent 2.0M tokens·$1.25/1M | — |
| Grok Beta 0 | — |
| Grok Build 0.1 0616 $1.00/1M | — |
| Grok Code Fast 1 256K tokens00 | — |
| Grok-1 0 | — |
| SpaceXAI: Grok 4.20 2.0M tokens·$1.25/1M | — |
| SpaceXAI: Grok 4.20 Multi-Agent 2.0M tokens·$1.25/1M | — |
| SpaceXAI: Grok 4.3 1.0M tokens·$1.25/1M | — |
| SpaceXAI: Grok 4.5 500K tokens·$2.00/1M | — |
| SpaceXAI: Grok 4.6 500K tokens·$2.00/1M | — |
| SpaceXAI: Grok Build 0.1 256K tokens·$1.00/1M | — |
| xAI: Grok Build 0.1 256K tokens·$1.00/1M | — |
The AI model ecosystem in 2026 is dominated by four major families: GPT from OpenAI, Claude from Anthropic, Gemini from Google, and Llama from Meta. Each family has models of different sizes and specializations, with varying prices and capabilities for different use cases.
OpenAI offers the GPT-4o line as its main model, with variants at different costs and speeds. GPT-4o-mini is the most affordable option with excellent cost-effectiveness. The OpenAI API is the most widely supported by third-party tools and integrations, making it the default choice for many applications.
Anthropic positions Claude with a focus on safety and following complex instructions. Claude Opus is the most capable model in the lineup, with a 200K token context window — ideal for analyzing long documents. Claude Haiku is the fastest and cheapest option. Anthropic has a strong presence in enterprise and compliance-sensitive use cases.
Gemini is notable for its 1 million token context window — the largest among commercial models — and native integration with the Google ecosystem (Search, Workspace, Cloud). Gemini Flash is the most affordable option with exceptional speed.
The open source segment has advanced significantly. Meta AI released Llama 4 with competitive performance in certain tasks. Alibaba maintains the Qwen family with a focus on multilingual support. DeepSeek surprised with frontier performance at substantially lower cost than equivalent proprietary models.
GPT-4o from OpenAI and Claude Opus from Anthropic are both frontier models with similar capabilities. GPT-4o has better speed and integration with the OpenAI ecosystem. Claude Opus excels at tasks with long context and complex reasoning.
Context window is the maximum amount of text the model can process in a single request, measured in tokens (approximately 4 characters per token in English). Models with larger context windows can analyze complete documents and extensive codebases.
Open source models include Llama (Meta), Qwen (Alibaba), Mistral, DeepSeek, and Gemma (Google). They are available under licenses that allow use, modification, and self-deployment, without depending on paid APIs.
LLMs charge per tokens processed — separated by input tokens (what you send) and output tokens (what the model generates). Prices are in USD per 1 million tokens. Output tokens typically cost 3-5x more than input tokens.