All Things AI
Beginner

Compare AI Models

Filter and sort 44 AI models across 11 providers side-by-side - pricing, context window, type, and what each is best for.

Covers OpenAI, Anthropic, Google, Meta, DeepSeek, Microsoft, Amazon, Mistral, xAI, Cohere, and Qwen. Via-API pricing per 1M tokens as of Sep 2026. Always verify current pricing at the official provider pricing page before production use. Subscription plan pricing differs from API pricing.

Comparison Table

Provider
Type
Price
Modality

Showing 44 of 44 models

ModelProviderTypeModalitiesParamsContext โ†•Input /1M โ†‘Output /1M โ†•TierOWBest for
Gemma 4 1BGoogleGeneral
TxtImgAud
1B128K$0.010$0.020Budgetโœ“Extreme edge, mobile inference
Llama 3.2 1BMetaGeneral
Txt
1B128K$0.010$0.010Budgetโœ“Smallest open-weight, edge & mobile
Mistral NemoMistralGeneral
Txt
12B128K$0.020$0.040Budgetโœ“Apache 2.0, ultra-budget open-weight
Gemma 4 4BGoogleGeneral
TxtImgAud
4B128K$0.030$0.060Budgetโœ“On-device and edge deployments
Phi-4-miniMicrosoftGeneral
Txt
3.8B16K$0.030$0.070Budgetโœ“3.8B ultra-small, on-device deployment
Nova MicroAmazonGeneral
Txt
-128K$0.035$0.14Budget-Text-only, lowest latency on AWS
Gemma 4 12BGoogleGeneral
TxtImgAud
12B128K$0.040$0.13Budgetโœ“Efficient open-weight text, vision & audio
Llama 3.2 11B VisionMetaGeneral
TxtImg
11B128K$0.050$0.050Budgetโœ“Open-weight vision & image understanding
Qwen3 8BQwenGeneral
Txt
8B128K$0.050$0.20Budgetโœ“Small open-weight, local inference
Nova LiteAmazonGeneral
TxtImgVid
-300K$0.060$0.24Budget-Fast multimodal, very low cost on AWS
Phi-4MicrosoftGeneral
Txt
14B16K$0.065$0.14Budgetโœ“14B - outperforms many larger text models
Phi-4-multimodalMicrosoftGeneral
TxtImgAud
5.6B128K$0.070$0.14Budgetโœ“Audio + image + text, small multimodal model
Gemma 4 27BGoogleGeneral
TxtImgAud
27B256K$0.080$0.16Budgetโœ“Open-weight multimodal + audio, self-hostable
Phi-4 Mini ReasoningMicrosoftReasoning
Txt
3.8B131K$0.080$0.32Budgetโœ“Small open-weight reasoning, 131K context
GPT-6 LunaOpenAIGeneral
TxtImg
-1.05M$0.10$0.50Budget-Cheapest GPT-6 generation model, high-volume tasks
Llama 4 ScoutMetaGeneral
TxtImg
109B (MoE)10M$0.10$0.35Budgetโœ“Ultra-long 10M context window, multimodal
Llama 3.3 70BMetaGeneral
Txt
70B128K$0.10$0.32Budgetโœ“Open-weight text workhorse, widely deployed
Mistral Small 3.1MistralGeneral
TxtImg
24B128K$0.10$0.30Budget-SOTA small model with vision, multilingual
DeepSeek V4 FlashDeepSeekGeneral
Txt
-163K$0.14$0.28Budgetโœ“Ultra-low cost open-weight inference
Command RCohereGeneral
Txt
35B128K$0.15$0.60Budget-Efficient RAG for lighter workloads
Qwen3 32BQwenGeneral
Txt
32B128K$0.15$0.60Budgetโœ“Strong open-weight, self-hostable
Llama 4 MaverickMetaGeneral
TxtImg
400B (MoE)1M$0.20$0.85Budgetโœ“Open-weight multimodal quality at low cost
Qwen3 CoderQwenGeneral
Txt
480B (MoE)128K$0.22$0.90Budgetโœ“Alibaba coding specialist, open-weight
DeepSeek V4 ProDeepSeekGeneral
Txt
1.6T (MoE)163K$0.27$1.10Budgetโœ“MIT license, frontier open-weight
Gemini 3.5 Flash-LiteGoogleGeneral
TxtImg
-1M$0.30$2.50Budget-Cheapest current-generation Flash tier
CodestralMistralGeneral
Txt
22B256K$0.30$0.90Budget-Code specialist - 80+ programming languages
DeepSeek R1DeepSeekReasoning
Txt
671B163K$0.55$2.19Budgetโœ“Open-weight reasoning, matches frontier reasoning models
Gemini 3.8 FlashGoogleGeneral
TxtImgVidAud
-1M$0.75$3.75Budget-Latest Flash, GA Sept 2026 - intro pricing through Dec 2026
Nova ProAmazonGeneral
TxtImgVid
-300K$0.80$3.20Budget-Multimodal agentic workflows on AWS
Claude Haiku 4.5AnthropicGeneral
TxtImgDoc
-200K$1.00$5.00Budget-High-volume extraction & classification
Grok 4.3xAIBoth
TxtImg
-131K$1.25$2.50Mid-Previous-gen Grok, still widely available and cheaper
Mistral Medium 3.5MistralGeneral
Txt
-128K$1.50$7.50Mid-Current flagship-tier, multilingual tasks
Grok 4.7xAIBoth
TxtImg
-131K$1.60$4.80Mid-Latest Grok (Sept 2026), frontier reasoning, X/Twitter knowledge
GPT-6 AstraOpenAIBoth
TxtImgDoc
-1.05M$2.00$10.00Mid-Flagship agentic model - operates software through screens, fills forms, edits spreadsheets
GPT-5.6 TerraOpenAIGeneral
TxtImgDoc
-1M$2.00$12.00Mid-Previous-generation mid-tier, still broadly deployed
Claude Sonnet 5AnthropicBoth
TxtImgDoc
-200K$2.00$10.00Mid-Most agentic Sonnet yet, everyday production balance
Gemini 3.1 ProGoogleBoth
TxtImgVidAudDoc
-1M$2.00$12.00Mid-Frontier tier, top coding benchmarks (rises to $4/$18 above 200K tokens)
Pixtral LargeMistralGeneral
TxtImg
-128K$2.00$6.00Mid-Multimodal flagship - vision + text
Nova PremierAmazonBoth
TxtImgVidAud
-1M$2.50$12.50Mid-Amazon flagship, extended thinking, 1M context
Command ACohereGeneral
Txt
-256K$2.50$10.00Mid-Enterprise agentic workflows
Command R+CohereGeneral
Txt
104B128K$2.50$10.00Mid-Enterprise RAG, grounded tool use
Claude Opus 5.5AnthropicBoth
TxtImgDoc
-1M$4.00$20.00Mid-Near-Fable quality at a fraction of the cost, agentic pipelines
Qwen3.8 MaxQwenBoth
Txt
-1M$4.00$12.00Midโœ“Alibaba flagship (Sept 2026), extended 1M context
Claude Fable 5.1AnthropicBoth
TxtImgDoc
-1M$10.00$50.00Frontier-Most capable Claude - coding and knowledge work, always-on adaptive thinking
Modalities:TxtText inputImgImage inputVidVideo inputAudAudio inputDocFiles input| Params: - = undisclosed ยท (MoE) = Mixture-of-Experts total

OW = Open-weight (publicly released model weights). Click column headers to sort. Prices approximate via API (May 2026); always verify at the official provider pricing page before production use.

Reading This Table

Budget (โ‰ค $1/1M input)

Best for high-volume, cost-sensitive workloads: classification, extraction, summarization at scale.

Mid ($1โ€“$5/1M input)

Production workhorses. Good quality/cost ratio for most use cases: coding, analysis, chat.

Frontier (> $5/1M input)

Reserve for hard reasoning, complex agentic tasks, or where output quality is the primary constraint.

General vs Reasoning vs Both

General = standard chat/completion. Reasoning = uses internal thinking steps (DeepSeek R1). Both = supports both modes - e.g., Sonnet 5 with extended thinking on/off.

Open-weight

Open-weight models have publicly released weights. You can self-host them or access via third-party APIs (Groq, Together, Fireworks). Prices shown are typical API rates - self-hosting costs vary by GPU.