| GPT-5.x | OpenAI | General purpose | GPT-5.6 Sol/Terra/Luna (Jul 9) is the current flagship API family; GPT-5.5 Instant is the ChatGPT default; next-gen Astra unreleased, flagged for possible critical cyber capability |
| Claude 5.x Family | Anthropic | Coding & reasoning | Claude Opus 5 (Jul 24) is the recommended starting model — near-Fable-5 intelligence at about half the price; Fable 5 remains the max-capability tier |
| Gemini 3.x | Google DeepMind | Multimodal | Gemini 3.7 Flash (Aug 13) is the new default workhorse; Gemini 3.5 Pro remains delayed with no confirmed launch date |
| Grok 4.x | xAI | Real-time info | Grok 4.6 (Aug 12) is the current flagship — 1.5T MoE, $2/$6 per 1M tokens, built for long-running agentic work |
| Llama 4 | Meta | Open-source | 10M token context, fully self-hostable; Behemoth shelved as Meta pivots to closed-weight Muse Spark |
| Muse Spark 1.2 | Meta Superintelligence Labs | Agentic coding & reasoning | Meta's first closed-weight frontier model (Aug 5) — agentic coding focus, ships with Muse Code terminal agent |
| DeepSeek V4 | DeepSeek | Cost efficiency | V4-Pro and V4-Flash-0731 — MIT license, 1M context, among the cheapest frontier-class models available |
| Mistral 3 Family | Mistral | EU compliance | Large 3, Medium 3.5, Small 4, Voxtral — enterprise-safe with EU data sovereignty |
| Qwen 3.8 | Alibaba | Multilingual | Qwen3.8-Max (Aug 2) is the new flagship; Qwen3.8-27B (Aug 13) is a separate, genuinely self-hostable Apache 2.0 companion |
| Microsoft MAI | Microsoft | Speech & media AI | MAI-Transcribe-1, MAI-Voice-1, MAI-Image-2, Phi-4-reasoning — Microsoft's own foundation stack on Foundry |
| Amazon Nova 2 | Amazon / AWS | AWS-native enterprise | Lite, Pro, Sonic, Omni — AWS-native family powering the Nova Act agent |
| MiniMax M3 | MiniMax | Cost-efficient coding | 1M context, 59.0% SWE-bench Pro at $0.15/$1.15 per 1M tokens |
| Command A+ | Cohere | Enterprise RAG | 218B MoE, Cohere's first fully Apache 2.0 frontier model, tuned for RAG |
| Gemma 4 | Google | On-device open-weight | 31B ranks #3 on Arena AI; E2B/E4B optimized for on-device Android |
| Kimi K3 | Moonshot AI | Agentic open-source | 2.8T MoE, open-weight — second only to Fable 5 and GPT-5.6 on most benchmarks |
| GLM-5.3 | Zhipu AI | Open frontier | MIT license, 1M context; GLM-5.3-Flash (Aug 26) adds native multimodal at ~1/10th the price, trained on Huawei Ascend chips |
| GPT-5.5-Cyber | OpenAI | Defensive cybersecurity | TAC-gated fine-tune for defensive cybersecurity — CrowdStrike, Cloudflare, Palo Alto, Cisco partners |
| Claude Mythos 5 | Anthropic | Security research | Cybersecurity research model via Project Glasswing — 150+ partner orgs, same weights as Fable 5 |
| Sakana Fugu Ultra | Sakana AI | Orchestration / routing | Meta-model that dynamically routes tasks across frontier models internally; 73.7% SWE-Bench Pro |
| Apple AFM 3 | Apple | On-device + cloud privacy | Five-model family spanning on-device and cloud — privacy-preserving inference for iOS/macOS |
| Sonar | Perplexity | Search & research | Search-grounded, citation-first answers at 1,200 tok/s on Cerebras inference |
| Composer 2.5 | Cursor | AI-native coding | Built on Kimi K2.5 with Cursor's post-training — matches Opus 4.7 at ~10× cheaper |
| SubQ | Subquadratic | Architecturally novel | First commercial subquadratic LLM — 12M token context at ~1/5th the compute cost of transformers |