Source audit: August 11, 2026

Compare AI Models: Hosted, Open Source, and Open Weight

Compare current hosted and open-model families with official sources, license context, standard API pricing where published, local deployment reality, and workflow guidance. Rankings are editorial judgments, not benchmark claims.

New open-model library

Eleven reviewed open-model families, with the details people actually search for.

Each page separates model access from ownership, open source from open weight, API price from infrastructure cost, and impressive context length from practical deployment. Start with the job you need the model to do, then open the detailed source record.

What is included

11 model pages9 providers covered1M context models explainedPrimary sources only
Provider

Open model data

Reuse the source-dated comparison

Download the current model, pricing, workflow-fit, and official-source fields. Licensed under CC BY 4.0 with attribution to T-Minus AI.

Best overall default

ChatGPT

Best if you want one assistant that can cover writing, research, image work, and agent-style workflows.

Best for writing

Claude

Best for long-form writing, nuanced tone, and careful document-heavy work.

Best for Google users

Gemini

Best if your work already runs through Gmail, Docs, Drive, Meet, and the wider Google stack.

Best for live social context

Grok

Best when X-native signals, trending context, and real-time commentary matter.

Best low-cost API option

DeepSeek

Best for teams optimizing hard around cost-per-token and API efficiency.

Best open-weight option

Llama

Best for builders who want self-hosting, open weights, and more infrastructure control.

Fast answer

ChatGPT vs Claude vs Gemini vs Grok vs DeepSeek vs Llama

If you only want the short version: ChatGPT is the best all-purpose default, Claude is the best writing model, Gemini is the strongest fit for Google-native work, Grok is best for live social context, DeepSeek is one of the best cost-efficient API choices, and Llama is the best-known open-weight family for teams that want more control.

ChatGPT

Broad all-purpose work

Strongest default if you want one assistant across research, writing, image tools, and agent-style workflows.

Claude

Writing and document work

The cleaner fit for long-form writing, document analysis, and careful structured output.

Gemini

Google-native workflows

A strong fit if your work happens inside Gmail, Docs, Sheets, Meet, and the Google ecosystem.

Grok

Live social and trend context

Most relevant when real-time X context and social signal tracking are central to the workflow.

DeepSeek

Low-cost API deployment

One of the most cost-efficient ways to access capable models for coding and budget-sensitive developer workflows.

Llama

Open-weight control

The best-known open-weight family when you want self-hosting, privacy control, or infrastructure flexibility.

Side-by-side specs

Compare AI Models, head-to-head.

ModelInput /1MOutput /1MAccessBest For
GPT-5.6 SolOpenAI$5$30ChatGPT PlusFrontier coding, knowledge work, and agentic tasks
GPT-5.6 TerraOpenAI$2.50$15ChatGPT Free (limited)Balanced everyday execution and coding
GPT-5.6 LunaOpenAI$1$6ChatGPT PlusCost-sensitive, high-throughput API work
Claude Sonnet 5Anthropic$2$10ClaudeBalanced coding, agents, and document-heavy work
Claude Opus 5Anthropic$5$25API and cloud platformsPremium reasoning and complex tool work
Claude Fable 5Anthropic$10$50API and cloud platformsMost demanding reasoning and long-horizon agents
Gemini 3.6 FlashGoogle$1.50$7.50GeminiMultimodal and agentic API work
Gemini 3.5 Flash-LiteGoogle$0.30$2.50GeminiHigh-volume, low-cost API tasks
Grok 4.5xAI$2$6GrokGeneral work and API experimentation
Kimi K3Moonshot AI¥20¥100Kimi API and certified inference partnersLong-horizon coding agents
MiniMax M3MiniMax$0.30$1.20MiniMax API and Token PlanAgentic coding
Qwen3.6Qwen / AlibabaCheck providerCheck providerQwen Studio and Alibaba Cloud Model StudioLocal coding assistance
GLM-5.2Z.aiModel-specific price not listedModel-specific price not listedZ.ai general API and GLM Coding PlanLong-horizon coding agents
DeepSeek V4DeepSeek$0.14$0.28DeepSeek APICost-sensitive API workflows
Mistral Small 4Mistral AI$0.15$0.60Mistral API and Mistral AI StudioMultimodal document analysis
Mistral Large 3Mistral AICheck providerCheck providerMistral AI Studio and selected cloud providersEnterprise custom models
Gemma 4GoogleNo universal Gemma API rateNo universal Gemma API rateDownloadable model family and Google AI ecosystemRight-sized open-model experiments
gpt-oss-120bOpenAINo OpenAI API priceNo OpenAI API priceSelf-hosted or third-party hosting providersPrivate-cloud reasoning services
gpt-oss-20bOpenAINo OpenAI API priceNo OpenAI API priceSelf-hosted or third-party hosting providersLocal text reasoning
Nemotron 3 SuperNVIDIACheck providerCheck providerDownloadable weights, NVIDIA services, and selected hosting providersEnterprise agent systems

Standard API costs per 1M tokens where the provider publishes a comparable USD rate. Open-model rows link to detail pages that separate API pricing from local operating cost. Last verified August 11, 2026.

Claude Sonnet 5 introductory API pricing is $2 input and $10 output per million tokens through August 31, 2026. Standard $3/$15 pricing begins September 1, 2026.

Model access in consumer chat plans varies by provider, region, and plan. Check the linked provider page before purchasing.

AI Model Finder

Find Your Perfect AI Model

Answer a few quick questions and we'll recommend the best AI model for your needs.

Question 1 of 333% Complete

What will you primarily use AI for?

Quick Reference Guide

Which AI for Which Task?

Not sure which model to use? This matrix maps common use cases to the best AI models with clear reasoning.

Writing long-form content

Creators
Best:Claude Sonnet 5

A practical current Claude option for document-heavy drafting and revision workflows.

Alt:GPT-5.6 Terra

Quick research with sources

Research
Best:Perplexity Pro

Fast web discovery with citations and strong source traceability.

Alt:Gemini 3.6 Flash

Advanced reasoning & math

Developers
Best:GPT-5.6 Sol

OpenAI positions its flagship current model for difficult knowledge work and agentic tasks.

Alt:Claude Opus 5

Code generation & debugging

Developers
Best:Claude Sonnet 5

A current Claude option for coding and agent workflows.

Alt:GPT-5.6 Sol

Excel/Word automation

Corporate
Best:Microsoft 365 Copilot

Native integration in Word, Excel, Outlook, and Teams workflows.

Alt:ChatGPT (GPT-5.6)

Real-time social signals

Research
Best:Grok 4.5

Strong fit when live X context is central to the workflow.

Alt:Perplexity Pro

Cost-effective API usage

Developers
Best:DeepSeek V3.2 Reasoner

Low-cost API with separate reasoning mode for budget-sensitive deployments.

Alt:Gemini 3.5 Flash-Lite

Privacy-focused browsing

Corporate
Best:Brave Leo

Browser-native assistant with privacy-first defaults and BYOM flexibility.

Alt:Local LLMs

Multi-language translation

Corporate
Best:Gemini 3.6 Flash

Strong multilingual handling with multimodal capabilities.

Alt:GPT-5.6 Terra

Technical documentation

Developers
Best:Claude Sonnet 5

A current Claude option for structured technical documentation workflows.

Alt:GPT-5.6 Sol

Brainstorming sessions

Creators
Best:GPT-5.6 Terra

A balanced current OpenAI option for quick iteration and idea generation.

Alt:Gemini 3.6 Flash

Image-generation briefs

Creators
Best:ChatGPT (GPT-5.6 + image tools)

Strong prompt iteration and multimodal planning for visual direction.

Alt:Gemini 3.6 Flash

This matrix is reviewed monthly. Last verified on August 11, 2026 using official provider documentation.

1

ChatGPT

OpenAI
Position Most adopted all-purpose assistant
Models GPT-5.6 Sol · GPT-5.6 Terra · GPT-5.6 Luna
"OpenAI positions the GPT-5.6 family across flagship, balanced, and cost-efficient tiers for coding, knowledge work, and agentic tasks."
Official source
2

Microsoft Copilot

Microsoft
Position Best native fit for Microsoft 365 workflows
Models Workspace-centric Copilot experiences across Microsoft 365
"Deep integration with Outlook, Word, Excel, Teams, and enterprise controls. Best fit when your AI layer must sit inside existing Microsoft workflows."
Official source
3

Google Gemini

Google
Position Best multimodal coverage + price-performance range
Models Gemini 3.6 Flash · Gemini 3.5 Flash · Gemini 3.5 Flash-Lite
"Google lists Gemini 3.6 Flash as its most balanced model for speed and intelligence, with 3.5 Flash and Flash-Lite covering sustained agentic and lower-cost workloads."
Official source
4

Perplexity

Perplexity
Position Fastest research-first assistant for source-backed discovery
Models Standard · Pro · Max · Enterprise Pro
"Perplexity's official plan guide makes the product easier to slot into a stack: free discovery at Standard, deeper research and advanced model access on Pro, and higher-end throughput on Max and Enterprise tiers."
Official source
5

Claude

Anthropic
Position High-trust long-context and agent workflows
Models Claude Fable 5 · Claude Opus 5 · Claude Sonnet 5 · Claude Haiku 4.5
"Claude Fable 5 is Anthropic's highest-capability generally available model for demanding reasoning and long-horizon agents. Opus 5 is the current premium Opus model, Sonnet 5 is the balanced coding and agent option, and Haiku 4.5 is the lower-cost tier."
Official source
6

Grok

xAI
Position Current xAI model and API offering
Models Grok 4.5
"xAI lists Grok 4.5 at $2 per million input tokens and $6 per million output tokens in its official release update."
Official source
7

DeepSeek

DeepSeek
Position Most cost-efficient API for coding and reasoning
Models DeepSeek V4 · DeepSeek V3 · DeepSeek R1
"DeepSeek's transparency center now lists V4 with an April 24, 2026 release date. The main draw remains cost pressure for API-heavy coding, reasoning, and builder workflows."
Official source
8

Brave Leo

Brave
Position Browser-native assistant with privacy focus
Models Model menu varies by free, premium, and BYOM options
"Built into Brave with privacy-centric defaults and model flexibility."
Official source
9

Meta Llama

Meta
Position Best open-source multimodal AI — free to self-host
Models Llama 4 Scout (17B) · Llama 4 Maverick (17B MoE) · Behemoth (training)
"Llama 4 Maverick beats GPT-4o and Gemini 2.0 Flash on major benchmarks while being fully open-weight. Multimodal (text + image), multilingual (12 languages), runs on a single H100."
Official source
10

Mistral AI

Mistral
Position European open-weight frontier AI with enterprise focus
Models Mistral 3 · Mistral Large 3 · Devstral 2
"Mistral 3 gives builders another serious open multimodal and multilingual model family, especially useful for privacy, EU data-residency, and on-premise deployment decisions."
Official source
Pricing Guide

AI Model Pricing Comparison

Compare pricing across major AI models. Most offer free tiers with paid upgrades for power users.

ChatGPT

$20/mo
✓ Free tier available

GPT-5.6 Luna text chat plus limited uploads, image creation, deep research, memory, context, Codex, and Work features.

Paid Tier: Plus

GPT-5.6 advanced reasoning, plus projects, tasks, custom GPTs, and higher limits across product features.

Pro: $200/mo

GPT-5.6 Sol Pro and the highest listed limits for supported ChatGPT capabilities.

Best for:

General execution and teams that want the broadest ChatGPT product surface

Claude

$20/mo
✓ Free tier available

Core Claude access with memory across conversations, web search, and limited advanced features

Paid Tier: Pro

Higher limits with stronger access to research, code execution, projects, and Claude's premium feature surface

Max 5×: $100/mo · Max 20×: $200/mo

Heavy-usage access for professionals who hit Pro limits regularly

Best for:

Long-form writing, technical docs, coding, and agent workflows

Google Gemini

Region-dependent
✓ Free tier available

Gemini app free tier with standard model access

Paid Tier: Google AI Plus / Pro / Ultra

More Gemini access and plan benefits vary by country and plan. Check the official plan page before buying.

Google AI Ultra: region-dependent

Highest limits and early access to advanced Gemini capabilities

Best for:

Google ecosystem power users, high-volume API workloads, multimodal tasks

Perplexity

$20/mo
✓ Free tier available

Practically unlimited basic searches with very limited Pro searches and basic file uploads

Paid Tier: Pro

Extended Pro Search, advanced model access, image and video generation, and higher upload limits

Max: $200/mo

Highest individual tier for users who need more throughput than Pro

Best for:

Source-backed research, market scanning, and multi-step agentic work

Microsoft Copilot

From $20/mo consumer; business plans vary
✓ Free tier available

Basic Copilot chat and web-grounded responses

Paid Tier: Copilot plans

Microsoft 365 app integration and work-grounded productivity features

Enterprise and agent add-ons available

Advanced controls, governance, and agent capabilities

Best for:

Office-heavy organizations and enterprise rollout on Microsoft stack

Pricing snapshot last verified on August 11, 2026. Plan names, limits, and regional prices can change quickly.

Live references

Check benchmark dashboards and official provider updates

Links open in a new tab. Snapshot last verified on August 11, 2026.

Frequently Asked Questions

AI Models — Common Questions

What is the best AI model in 2026?

There is no single best model for every workflow. This page was source-audited on August 11, 2026. ChatGPT is a broad all-purpose choice, Claude suits document-heavy work, Gemini fits Google-centric workflows, and each provider has different plan limits and API pricing.

What is the difference between ChatGPT and Claude in 2026?

ChatGPT is the broader product surface: more plan tiers, more built-in tools, deeper research and agent features, and stronger general-purpose coverage. Claude is the steadier choice for long documents, nuanced writing, Projects, Research, and focused document-heavy workflows.

How do Grok, DeepSeek, and Llama compare with ChatGPT and Claude?

Grok matters most when live X context is central to the workflow, DeepSeek is attractive when low API cost is the deciding factor, and Llama is the clearest open-weight option for teams that want self-hosting and infrastructure control. ChatGPT and Claude remain the strongest defaults for most general-purpose business and writing workflows.

Which AI model is best for coding in 2026?

For coding, the safest recommendation is to separate chat models from coding products. Claude remains a strong coding assistant, ChatGPT is useful for broad debugging and tool-heavy workflows, and GitHub Copilot is the best fit if you want coding help directly inside your IDE with team controls.

What is the best free AI model in 2026?

The best free AI model depends on the job. ChatGPT Free is the broadest starting point, Claude Free is strong for writing and documents, Gemini Free works well for Google-native workflows, and Llama or DeepSeek can be better if you care more about open weights or low-cost experimentation than polished consumer apps.

Should I pay for ChatGPT, Claude, Google AI, or Perplexity first?

Most people should pay for ChatGPT first if they want one broad default assistant. Choose Claude first if long-form writing, analysis, and document-heavy work matter more. Choose Google AI first if your workflow already lives inside Google. Choose Perplexity first if research with citations is the real bottleneck.