AI Tool Tier List · Updated Daily
AI tools, ranked S to F.
Every tool tested, scored, and placed in its tier. We report the bugs, show the pricing traps, and tell you what actually works. No sponsored placements.
01 The Tier List
Top picks across 5 categories.
Last updated: Jul 22, 2026
02 Browse by Category
Find tools for exactly what you need.
03 Latest Reviews
Recently reviewed and updated.
Claude (Anthropic)
Anthropic's flagship LLM family. After a 19-day US-government export-control suspension (June 12-30), Claude Fable 5 -- the first publicly available Mythos-class model -- returned globally on July 1, 2026. New on June 30: Claude Sonnet 5, the 'most agentic Sonnet yet,' now the default on Free/Pro at $2/$10 per 1M (intro through Aug 31, then $3/$15). Opus 4.8 remains the top-end flagship at $5/$25 per 1M with a 1M-token context, effort control, and a cheap fast mode
Gemini (Google)
Google's LLM with deep Google Workspace integration, 2M token context window, and native code execution -- Gemini 3.6 Flash + 3.5 Flash-Lite GA 2026-07-21 (the 'upgraded Flash stopgap'; 3.6 Flash at $1.50/$7.50 per 1M, 17% fewer output tokens), Gemini 3.5 Pro STILL delayed and partner-testing-only (Bloomberg, 7/16 -- coding shortfalls, no ship date), Gemini 4 pre-training now underway
Mistral AI
European AI lab with open and commercial models -- Le Chat is now **Vibe** (May 28 2026): one agent across Work Mode + Code Mode with a VS Code extension and CLI, powered by Mistral Medium 3.5 (128B dense, 256k context, 77.6% SWE-Bench Verified). Earlier 2026 line: Small 4 (119B MoE Apache 2.0), Medium 3, Voxtral TTS
DeepSeek
DeepSeek V4 shipped 2026-04-24: V4-Pro (1.6T/49B active MoE) + V4-Flash (284B/13B active), 1M native context, Hybrid Attention Architecture, open-source on HF. Trails only Gemini 3.1 Pro on world knowledge
Qwen (Alibaba)
Alibaba's open-weights + API family -- Qwen3.8-Max flagship previewed at WAIC (Jul 19 2026: 2.4T sparse-MoE multimodal, closed preview, 'second only to Fable 5'), Qwen 3.7 Max GA (SWE-Bench Pro 60.6%, Terminal-Bench 69.7%, $2.50/$7.50 per 1M), Qwen3.6-27B dense Apache 2.0 (beats the 397B MoE on coding from one consumer GPU)
Kimi K3 (Moonshot)
Moonshot's 2.8T-parameter Kimi K3 (launched 2026-07-16/17) is the largest open-weight model ever announced -- 1M context, multimodal, $3/$15 per 1M via API, ranked best-available on Arena.AI at launch. Weights promised late July (press cites 7/27); K2.6/K2.7-Code remain the shipped-weights line
04 Reviews you can actually trust
Every review is based on hands-on testing, cross-referenced user sentiment from G2, Reddit, and Capterra, and real pricing data. We report known bugs. We don't do paid placements.