AI // Model Catalog

Every major AI assistant, the company behind it, and whether each model is frontier (closed, cloud only), hybrid (cloud and your own hardware) or local (runs on your own machine). New to the terms? Start with the glossary.

26 companies  |  frontier, hybrid, local  |  lineups checked 13 Sep 2026

Logos and screenshots belong to their owners and are here so each product can be recognized.

App name to company

The name on the app is usually not the name of the company. This table maps one to the other, and shows what kind of model sits behind each app.

You know it asMade byTop model todayKindGood to know
ChatGPT OpenAI GPT-6 Astra Frontier Astra needs a paid plan. On Plus it sits in Work and Codex; regular chat tops out at GPT-5.6 Sol.
Claude Anthropic Claude Fable 5.1 Frontier Also Opus 5, Sonnet 5 and Haiku 4.5. Claude Code is the coding tool.
Gemini Google (DeepMind lab) Gemini 3.1 Pro Frontier Called Bard until February 2024. Gemini 3.5 Pro is announced but not public yet.
Copilot Microsoft A mix: OpenAI GPT, Anthropic Claude, Microsoft MAI Frontier Built into Windows, Edge and Microsoft 365. Microsoft does not make GPT; it invests in OpenAI.
GitHub Copilot Microsoft (owns GitHub) Your pick: GPT, Claude, Gemini or MAI-Code-1-Flash Frontier Coding assistant that lives inside code editors.
Grok SpaceXAI (was xAI) Grok 4.6 Frontier Elon Musk's AI, built into X (Twitter). xAI became part of SpaceX in 2026.
Meta AI Meta Muse Spark Frontier Inside WhatsApp, Instagram, Facebook, Messenger and Ray-Ban Meta glasses. Muse replaced Llama as Meta's top model.
Siri Apple Apple Foundation Models Hybrid Small requests run on the phone itself, bigger ones in Apple's cloud. Built with Google's Gemini.
Alexa+ Amazon Amazon Nova and Anthropic Claude Cloud only Echo speakers and the Alexa app.
Le Chat Mistral AI Mistral Large 3 Hybrid French lab. Use it in Le Chat, or download the model.
DeepSeek DeepSeek DeepSeek V4-Pro Hybrid Chinese. Free app, and the model is free to download under the MIT license.
Qwen Alibaba Qwen3.8-Max Hybrid Chinese. Comes in many sizes; the small ones run locally.
Kimi Moonshot AI Kimi K3 Hybrid Chinese. 2.8 trillion parameters, billed as the largest downloadable model.
Z.ai Zhipu AI GLM-5.3 Hybrid Chinese. The chat app used to be called ChatGLM.
MiniMax MiniMax MiniMax M3 Hybrid Chinese. Hailuo is its video generator.
Doubao ByteDance Seed models Cloud only Made by TikTok's parent company.
ERNIE Bot Baidu ERNIE 5.1 Cloud only Assistant from China's biggest search engine. The older ERNIE 4.5 is downloadable.
Perplexity Perplexity Sonar, plus GPT, Claude and Gemini Cloud only AI search engine that rents other labs' models.
Midjourney Midjourney Midjourney V8.1 Cloud only Images, and short video, from a text prompt.
Sora OpenAI Sora Frontier Video generator.
Veo Google Veo 3.1 Frontier Video generator with sound, inside Gemini.
Gemma, Llama, Phi Google, Meta, Microsoft Gemma 4, Llama, Phi Local No app of their own. You download them and run them offline with a tool like Ollama or LM Studio.

Frontier, hybrid or local

Every model on this page carries one of these labels. The difference is where the model actually runs, and whether you can ever get the model file.

Frontier

Closed. Cloud only.

The best models from the five frontier labs. You reach them only through the company's app or API, and the model file is never released. Needs an internet connection and usually a subscription; your prompts go to the company's servers.

Claude Fable 5.1, GPT-6 Astra, Gemini 3.1 Pro, Grok 4.6, Muse Spark

Hybrid

Both. Cloud and your own hardware.

The maker runs the model in its own app and API, and the model file is also free to download, or the work is split between your device and the company's cloud (Apple). The big ones need a multi-GPU server to run yourself, not a home PC.

DeepSeek V4-Pro, Kimi K3, Qwen3.8-Max, Mistral Large 3, Apple Foundation Models

Local

Runs on your own PC, laptop or phone.

Small downloadable models you run offline. Nothing leaves your machine and there is no usage fee, but they sit well behind the frontier. Good for private documents, offline use and tinkering.

Gemma 4, gpt-oss-20b, Qwen3.8-27B, Ministral 3, Llama, Phi, Stable Diffusion

Cloud only  is used for closed models from companies outside the five frontier labs, such as Amazon Nova, Microsoft MAI, ERNIE 5.1 and Midjourney. They work the same way as frontier models (online, never downloadable) but are not at the very top.

FrontierHybridLocal
Where it runsThe company's data centersThe company's cloud, or your own server or deviceYour own PC, laptop or phone
Download the modelNoYes (Apple: built into the device)Yes
Works offlineNoOnly when you run it yourselfYes
PrivacyPrompts go to the companyYour choiceNothing leaves the machine
CostSubscription or pay per useFree file; pay for hardware or for the maker's APIFree once you own the hardware
CapabilityThe best availableClose behind the frontierFine for everyday tasks, well behind

Running models locally

Local models are downloaded from Hugging Face or a model library and run with a free tool: Ollama (simplest, one command per model), LM Studio (desktop app with a chat window) or llama.cpp (the engine under many of these). What fits depends mostly on memory: graphics card memory on a PC, unified memory on a Mac.

Model sizeMemory needed (rough, 4-bit)Runs onExamplesKind
1B to 4B2 to 4 GBPhones and any recent laptopGemma 4 E2B and E4B, Ministral 3 3B, Gemini NanoLocal
7B to 14B6 to 12 GBA laptop or mid-range gaming GPUMinistral 3 8B and 14B, Llama 3.1 8B, PhiLocal
20B to 32B16 to 24 GBA high-end gaming GPU or a 32 GB Macgpt-oss-20b, Gemma 4 26B and 31B, Qwen3.8-27BLocal
100B to 130B64 to 80 GBOne data-center GPU, or a workstation with 128 GBgpt-oss-120b, Mistral Medium 3.5Hybrid
300B and upHundreds of GBMulti-GPU serversDeepSeek V4, Kimi K3, Qwen3.8-Max, GLM-5.3, MiniMax M3, Mistral Large 3Hybrid

"B" is billions of parameters. For mixture-of-experts models the total size decides the memory needed, even though only part of the model works on each word.

How they connect

The companies are tangled together. These are the links that explain most of the confusion.

WhoConnected toHow
MicrosoftOpenAIOpenAI's biggest outside investor. Copilot runs OpenAI models, and also Anthropic's Claude and Microsoft's own MAI models.
AmazonAnthropicMajor investor. AWS is a main home for Claude, and Alexa+ runs on Amazon Nova plus Claude.
GoogleAnthropicAlso an investor in Anthropic, while competing with its own Gemini.
AppleGoogleThe Apple Foundation Models behind the new Siri were built with Google's Gemini (deal announced 12 Jan 2026).
SpaceXSpaceXAIBought xAI on 2 Feb 2026 and renamed it SpaceXAI on 6 Jul 2026. Grok and X (Twitter) came with it.
NVIDIAEveryoneSells the GPU chips nearly every lab here trains and runs its models on.
Perplexity, GitHub CopilotMany labsApps that let you pick another company's model instead of only building their own.

Glossary

LLM (large language model)
The engine inside a chatbot. It is trained on a huge amount of text to predict what comes next, which turns out to be enough to write, code and reason.
Model vs. app
ChatGPT is the app; GPT-6 Astra is the model running inside it. One app can offer several models, and one model can power many apps. Claude, for example, runs inside Copilot and Alexa+.
Frontier model
The most capable models that exist at a given moment, from the few labs that can afford to train them. On this page the label also means closed and cloud only.
Hybrid model
A model you can use both ways: in the maker's cloud, or on your own hardware. Usually that means downloadable weights that the maker also hosts.
Local model
A model small enough to download and run on your own PC, laptop or phone, with no internet connection.
Open weights
The model file itself is published, so anyone can download it. Every hybrid and local model on this page has open weights, except Apple's models and Gemini Nano, which ship built into devices.
Reasoning model
Spends extra time working through a problem before answering. Slower and more expensive, but better at math, code and planning.
Parameters
The size of a model, counted in billions (B) or trillions (T). "1.6T total / 49B active" describes a mixture-of-experts model, which switches on only part of itself for each word.
Quantization
Storing a model's numbers with fewer bits (for example 4-bit) so it fits in less memory, at a small cost in quality. It is how 20B to 30B models fit on one gaming GPU.
Context window
How much text a model can take in at once, measured in tokens. A token is about three quarters of a word, so 1M tokens is roughly 750,000 words.
API
The paid developer connection that lets other companies build a model into their own products.

Frontier labs

The five companies building the most capable general models. Nearly everything they sell is frontier (cloud only); a few small models are downloadable.

Anthropic logo

Anthropic makes Claude

Based in
San Francisco
Founded
2021 by Dario and Daniela Amodei and other former OpenAI researchers
Backed by
Amazon and Google
  • Claude
  • Claude Code
ModelWhat it isKind
Claude Fable 5.1TopAnthropic's most capable model, for coding, knowledge work and long agent tasks. Released 1 Sep 2026.Frontier
Claude Mythos 5.1The same model as Fable 5.1 with different safeguards. Restricted to vetted cybersecurity and life-sciences programs.Frontier
Claude Opus 5Close to Fable at about half the price, with a low, medium or high effort setting. Released 24 Jul 2026.Frontier
Claude Sonnet 5Mid-size everyday model.Frontier
Claude Haiku 4.5Small, fast and cheapest. Still cloud only.Frontier
OpenAI logo

OpenAI makes ChatGPT

Based in
San Francisco
Founded
2015; CEO Sam Altman
Backed by
Microsoft, its largest outside investor
  • ChatGPT
  • Codex
  • Sora
ModelWhat it isKind
GPT-6 AstraTopOpenAI's most capable model, built for computer use, browsing, coding and science. Released 3 Sep 2026. Pro, Business and Enterprise plans also get GPT-6 Astra Pro.Frontier
GPT-5.6 SolThe previous flagship. Still the top choice in regular chat on the Plus plan.Frontier
GPT-Live-1Real-time voice model that can listen and talk at the same time. Developer API, Sep 2026.Frontier
ChatGPT Images 2.5Image generation inside ChatGPT. Released 8 Sep 2026.Frontier
SoraVideo generation, as its own app and model.Frontier
gpt-oss-120bDownloadable model (Aug 2025). Needs one 80 GB data-center GPU; also hosted by cloud providers.Hybrid
gpt-oss-20bSmaller downloadable model (Aug 2025). Runs on a PC or laptop with 16 GB of memory.Local
Google DeepMind logo

Google DeepMind makes Gemini

Based in
London; part of Google (Alphabet)
Formed
2023, when Google Brain and DeepMind merged; CEO Demis Hassabis
Old name
The Gemini app was called Bard until February 2024
  • Gemini
  • Gemma
  • Veo
ModelWhat it isKind
Gemini 3.1 ProTopGoogle's top public model in the Gemini app and API.Frontier
Gemini 3.5 ProNot public yet. Announced at Google I/O on 19 May 2026 and used inside Google; the release has slipped.Frontier
Gemini 3.7 FlashNewest fast "workhorse" model, aimed at coding and agents.Frontier
Gemini 3.6 Flash, 3.5 Flash-LiteCheaper models for high-volume work (July 2026).Frontier
Gemini 3.5 Flash CyberFinds and patches software vulnerabilities. Restricted to governments and selected partners.Frontier
Veo 3.1Video generation with sound.Frontier
Gemma 4Downloadable family in four sizes (E2B, E4B, 26B, 31B), Apache 2.0, released 2 Apr 2026. The small ones run on phones.Local
Gemini NanoBuilt into Android phones and Chrome; works without an internet connection.Local
SpaceXAI logo

SpaceXAI makes Grok (formerly xAI)

Founded
2023 by Elon Musk, as xAI
Owned by
SpaceX, which bought xAI on 2 Feb 2026 and renamed it SpaceXAI on 6 Jul 2026
Lives in
X (Twitter), grok.com and Tesla cars
  • Grok
  • xAI (old logo)
ModelWhat it isKind
Grok 4.6TopNewest flagship, tuned for coding and agents, with a 500K-token context window. Released 12 Aug 2026.Frontier
Grok 4.5Previous flagship, released 8 Jul 2026.Frontier
Grok 4 HeavyMulti-agent version of Grok 4 on the most expensive plan.Frontier
Grok 5Reported to be in training.Frontier
Grok ImagineImage and video generation.Frontier
Meta logo

Meta makes Meta AI

Based in
Menlo Park, California
AI lab
Meta Superintelligence Labs
Lives in
WhatsApp, Instagram, Facebook, Messenger and Ray-Ban Meta glasses
  • Meta AI
ModelWhat it isKind
Muse SparkTopMeta's top model and the first from Meta Superintelligence Labs (Apr 2026), with point updates since. Closed, unlike Llama.Frontier
Muse ImageImage generation (Jul 2026).Frontier
LlamaDownloadable family. Llama 4 (Apr 2025) is a generation behind; the small Llama 3 sizes are still widely run on home PCs.Local

Big tech

Giants that mostly package other labs' models into their products, while building some of their own.

Microsoft logo

Microsoft makes Copilot

Based in
Redmond, Washington
AI chief
Mustafa Suleyman, Microsoft AI
Models used
OpenAI's GPT, Anthropic's Claude, and its own MAI models
  • Copilot
  • GitHub Copilot
ModelWhat it isKind
MAI-Thinking-1TopMicrosoft's first in-house reasoning model, launched at Build in June 2026.Cloud only
MAI-Code-1-FlashFast coding model rolling into GitHub Copilot.Cloud only
MAI-Image-2.5Image generation and editing.Cloud only
MAI-Voice-2Speech generation.Cloud only
PhiSmall downloadable models that run on ordinary PCs.Local
Apple logo

Apple makes Siri and Apple Intelligence

Based in
Cupertino, California
Partner
Google. The new models were built with Gemini (announced 12 Jan 2026)
Runs on
The device itself, plus Apple's Private Cloud Compute servers
  • Siri
ModelWhat it isKind
Apple Foundation ModelsTopA small model runs on the iPhone, iPad or Mac itself; bigger requests go to Apple's Private Cloud Compute servers. Powers Siri AI in iOS 27, previewed at WWDC on 8 Jun 2026.Hybrid
Amazon logo

Amazon makes Alexa+ and Nova

Based in
Seattle
Invests in
Anthropic
Cloud
AWS Bedrock rents out Claude, Llama, Mistral and others to businesses
  • Alexa+
  • Nova
  • AWS
ModelWhat it isKind
Amazon Nova 2TopAmazon's own family (Lite, Pro, Sonic) covering text, images and real-time speech.Cloud only
Nova ActAgent that operates a web browser for you.Cloud only
NVIDIA logo

NVIDIA makes the chips

Based in
Santa Clara, California
Why it matters
Nearly every lab in this catalog trains and runs its models on NVIDIA GPUs, and most local models run on its gaming cards
ModelWhat it isKind
NemotronDownloadable models in sizes from laptop-friendly Nano up to data-center Ultra. NVIDIA also hosts them.Hybrid

China

Most Chinese labs are Hybrid: they run their own chat apps and also publish the model files. The flagships are far too big for a home PC, but their small sizes run locally. Version numbers change almost monthly.

DeepSeek logo

DeepSeek

Based in
Hangzhou
Founded
2023 by Liang Wenfeng, funded by the hedge fund High-Flyer
Known for
The R1 release in January 2025 that knocked US AI stocks
  • DeepSeek app
ModelWhat it isKind
DeepSeek V4-ProTop1.6T total / 49B active parameters, MIT license. First released 24 Apr 2026; current checkpoint August 2026.Hybrid
DeepSeek V4-Flash284B total / 13B active. Cheaper and faster, still server-sized.Hybrid
Alibaba Group logo

Alibaba makes Qwen

Based in
Hangzhou
Known for
The widest range of downloadable model sizes of any lab
  • Qwen
ModelWhat it isKind
Qwen3.8-MaxTopMultimodal flagship. Its text weights were opened on 12 Aug 2026.Hybrid
Qwen3.8-27BSmall multimodal model, Apache 2.0 license. Fits on one high-end gaming GPU.Local
Moonshot AI logo

Moonshot AI makes Kimi

Based in
Beijing
Founded
2023 by Yang Zhilin
  • Kimi
ModelWhat it isKind
Kimi K3Top2.8 trillion parameters, billed as the largest downloadable model. Released 17 Jul 2026; weights published 27 Jul.Hybrid
Z.ai logo

Zhipu AI makes GLM and Z.ai

Based in
Beijing
Founded
2019, out of Tsinghua University
Old name
The chat app was ChatGLM; the international brand is now Z.ai
  • Z.ai
  • ChatGLM
ModelWhat it isKind
GLM-5.3TopCurrent flagship, focused on coding.Hybrid
GLM-5.3-Flash320B total / 18B active, multimodal.Hybrid
MiniMax logo

MiniMax makes MiniMax and Hailuo

Based in
Shanghai
Founded
2021
  • MiniMax
  • Hailuo
ModelWhat it isKind
MiniMax M3Top428B total / 23B active, about 1M-token context, under MiniMax's own license. Released 1 Jun 2026.Hybrid
HailuoVideo generation app.Cloud only
ByteDance logo

ByteDance makes Doubao

Based in
Beijing
Also owns
TikTok and Douyin
  • Doubao
ModelWhat it isKind
Seed modelsTopByteDance's in-house family behind the Doubao chatbot.Cloud only
Seedance 2.0Video generation.Cloud only
Baidu logo

Baidu makes ERNIE

Based in
Beijing
Known for
China's largest search engine
  • ERNIE Bot (Wenxin)
ModelWhat it isKind
ERNIE 5.1TopBaidu's newest model, released 8 May 2026. Hosted only; the weights are not released.Cloud only
ERNIE 4.5Previous generation, released as ten downloadable sizes in 2025, from 0.3B up to 424B.Hybrid

Europe and others

Mistral AI logo

Mistral AI makes Le Chat

Based in
Paris
Founded
2023 by Arthur Mensch, Guillaume Lample and Timothée Lacroix
Known for
Europe's leading lab; nearly every model is downloadable
  • Le Chat
ModelWhat it isKind
Mistral Large 3Top675B total / 41B active, Apache 2.0 license. Released 2 Dec 2025.Hybrid
Mistral Medium 3.5128B, downloadable under a modified MIT license; self-hosts on about four GPUs. Apr 2026.Hybrid
Ministral 3Small models in 3B, 8B and 14B sizes, Apache 2.0. Run on laptops.Local
Cohere logo

Cohere makes Command

Based in
Toronto
Founded
2019; co-founder Aidan Gomez co-wrote the 2017 Transformer paper
Sells to
Businesses and governments, not consumers
ModelWhat it isKind
Command A+Top218B total / 25B active, reads text and images, 48 languages, Apache 2.0. Runs on two data-center GPUs. Released 20 May 2026.Hybrid
Perplexity logo

Perplexity makes AI search

Based in
San Francisco
Founded
2022 by Aravind Srinivas and others
Models used
Its own Sonar models, plus GPT, Claude and Gemini
Kind
Cloud only
  • Perplexity

Image, video and voice

Labs that make pictures, video, voices and music rather than chat.

CompanyModelMakesKind
MidjourneyMidjourney V8.1Images and short video. San Francisco.Cloud only
Black Forest LabsFLUX.2Images. Pro runs in the cloud; dev and klein are downloadable, and klein runs on gaming GPUs. Freiburg, Germany.Hybrid
OpenAISora, ChatGPT Images 2.5Video and images.Frontier
GoogleVeo 3.1Video with sound.Frontier
SpaceXAIGrok ImagineImages and video.Frontier
RunwayGen-4.5Video. New York.Cloud only
KuaishouKling 3.0Video. Beijing.Cloud only
ByteDanceSeedance 2.0Video.Cloud only
MiniMaxHailuoVideo.Cloud only
ElevenLabsElevenLabs voice modelsVoices, dubbing and speech.Cloud only
SunoSunoSongs with vocals from a text prompt.Cloud only
Stability AIStable DiffusionImages. The classic model people run at home on a gaming GPU. London.Local

Logo wall

The original logos, so each one is recognizable on sight.

Companies

Anthropic
AnthropicClaude
OpenAI
OpenAIChatGPT, Sora
Google
GoogleGemini, Veo, Gemma
Google DeepMind
Google DeepMindGoogle's AI lab
SpaceXAI
SpaceXAIGrok, new logo Jul 2026
xAI
xAIOld logo, before Jul 2026
Meta
MetaMeta AI, Muse, Llama
Microsoft
MicrosoftCopilot, MAI, Phi
Apple
AppleSiri, Apple Intelligence
Amazon
AmazonAlexa+, Nova
NVIDIA
NVIDIAGPUs, Nemotron
DeepSeek
DeepSeekDeepSeek V4
Alibaba
AlibabaQwen
Moonshot AI
Moonshot AIKimi
Z.ai
Zhipu AI (Z.ai)GLM
MiniMax
MiniMaxMiniMax M3, Hailuo
ByteDance
ByteDanceDoubao, Seedance
Baidu
BaiduERNIE
Mistral AI
Mistral AILe Chat
Cohere
CohereCommand
Perplexity
PerplexityAI search
Midjourney
MidjourneyImages
Black Forest Labs
Black Forest LabsFLUX
Runway
RunwayVideo
Kuaishou
KuaishouKling
ElevenLabs
ElevenLabsVoice
Stability AI
Stability AIStable Diffusion

Apps and models

ChatGPT
ChatGPTOpenAI
Claude
ClaudeAnthropic
Claude Code
Claude CodeAnthropic
Codex
CodexOpenAI
Sora
SoraOpenAI
Gemini
GeminiGoogle
Gemma
GemmaGoogle, local
Copilot
CopilotMicrosoft
GitHub Copilot
GitHub CopilotMicrosoft
Grok
GrokSpaceXAI
Meta AI
Meta AIMeta
Siri
SiriApple
Alexa
Alexa+Amazon
Amazon Nova
NovaAmazon
Le Chat
Le ChatMistral AI
DeepSeek
DeepSeekDeepSeek
Qwen
QwenAlibaba
Kimi
KimiMoonshot AI
ChatGLM
ChatGLMZhipu AI
Hailuo
HailuoMiniMax
Doubao
DoubaoByteDance
ERNIE Bot
ERNIE BotBaidu
Perplexity
PerplexityPerplexity
FLUX
FLUXBlack Forest Labs
Kling
KlingKuaishou
Suno
SunoSuno

Screens

What some of these look like in use. Screenshots come from Wikipedia and Wikimedia Commons, so a few show older versions.

Google Gemini app screenshot
Gemini, Google, 2026Google_Gemini_Screenshot_(2026).png
Microsoft Copilot on Windows screenshot
Copilot on Windows 10 and 11, MicrosoftMicrosoft_Copilot_on_Windows_10,_11.png
Grok chatbot screenshot
Grok, SpaceXAIGrok_chatbot_example_screenshot.webp
Qwen chatbot screenshot
Qwen, Alibaba (Qwen 3 era)Qwen_3_chatbot_example_screenshot.webp
Kimi chatbot screenshot
Kimi, Moonshot AI (Kimi K2 era)Kimi_K2_chatbot_example_screenshot.webp
Siri with Apple Intelligence screenshot
Siri with Apple Intelligence, iOS 18.1 beta (2024)Apple_Intelligence_Siri_(iOS_18.1_Beta_4).png
Sample image generated by FLUX.2 Pro
FLUX.2 Pro sample output, Black Forest LabsThe_Path_to_the_Mountain_(FLUX.2_Pro).webp