June
Full June 2026 report →
Major Release
Anthropic Claude Fable 5 โ 1M tokens context, reasoning model.
๐ฐ $10.00 in / $50.00 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
Cohere North Mini Code โ 256K tokens context, reasoning model.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Google Gemma 4 12B โ 131K tokens context, reasoning model.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text, image, video
๐ค Out: text
Major Release
Alibaba Qwen3.7 Plus โ 1M tokens context, reasoning model.
๐ฐ $0.40 in / $1.16 out per 1M tok
๐ In: text, image, video
๐ค Out: text
Major Release
MiniMax MiniMax-M3 โ 1M tokens context, reasoning model.
๐ฐ $0.30 in / $1.20 out per 1M tok
๐ In: text, image, video
๐ค Out: text
May
Full May 2026 report →
Major Release
StepFun Step 3.7 Flash โ 256K tokens context, reasoning model.
๐ฐ $0.20 in / $1.15 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
Anthropic Claude Opus 4.8 โ 1M tokens context, reasoning model.
๐ฐ $6.25 in / $25.00 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
Liquid AI LFM2.5-8B-A1B โ 32K tokens context, reasoning model.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Multiverse Computing HyperNova 60B 2605 โ 131K tokens context, reasoning model.
๐ฐ $0.04 in / $0.14 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
OpenBMB MiniCPM5-1B โ 128K tokens context.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Alibaba Qwen3.7 Max โ 1M tokens context, reasoning model.
๐ฐ $2.50 in / $7.50 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Google Gemini 3.5 Flash โ 1M tokens context, reasoning model.
๐ฐ $1.50 in / $9.00 out per 1M tok
๐ In: text, image, audio, video
๐ค Out: text
Major Release
China Mobile JT-35B-Flash โ 256K tokens context.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
OpenBMB MiniCPM-V 4.6 1.3B โ 262K tokens context.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text, image, video
๐ค Out: text
Major Release
InclusionAI Ring-2.6-1T โ 262K tokens context, reasoning model.
๐ฐ $0.30 in / $2.50 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
OpenAI GPT-5.5 Instant โ 400K tokens context, reasoning model.
๐ฐ $5.00 in / $30.00 out per 1M tok
๐ In: text, image
๐ค Out: text
April
Full April 2026 report →
Major Release
xAI Grok 4.3 โ 1M tokens context, reasoning model.
๐ฐ $1.25 in / $2.50 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
IBM Granite 4.1 30B โ 131K tokens context.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
IBM Granite 4.1 3B โ 131K tokens context.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
IBM Granite 4.1 8B โ 131K tokens context.
๐ฐ $0.05 in / $0.10 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Mistral Mistral Medium 3.5 โ 256K tokens context, reasoning model.
๐ฐ $1.50 in / $7.50 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
DeepSeek DeepSeek V4 Flash โ 1M tokens context, reasoning model.
๐ฐ $0.14 in / $0.28 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
DeepSeek DeepSeek V4 Pro โ 1M tokens context, reasoning model.
๐ฐ $1.74 in / $3.48 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
InclusionAI Ling-2.6-1T โ 262K tokens context.
๐ฐ $0.30 in / $2.50 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
OpenAI GPT-5.5 โ 922K tokens context, reasoning model.
๐ฐ $5.00 in / $30.00 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
Tencent Hy3-preview โ 256K tokens context, reasoning model.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Alibaba Qwen3.6 27B โ 262K tokens context, reasoning model.
๐ฐ $0.60 in / $3.60 out per 1M tok
๐ In: text, image, video
๐ค Out: text
Major Release
Xiaomi MiMo-V2.5 โ 1M tokens context, reasoning model.
๐ฐ $0.36 in / $1.80 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
Xiaomi MiMo-V2.5-Pro โ 1M tokens context, reasoning model.
๐ฐ $1.00 in / $3.00 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
InclusionAI Ling 2.6 Flash โ 262K tokens context.
๐ฐ $0.10 in / $0.30 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Alibaba Qwen3.6 Max Preview โ 256K tokens context, reasoning model.
๐ฐ $1.30 in / $7.80 out per 1M tok
๐ In: text
๐ค Out: text
Major Release
Kimi Kimi K2.6 โ 256K tokens context, reasoning model.
๐ฐ $0.95 in / $4.00 out per 1M tok
๐ In: text, image, video
๐ค Out: text
Major Release
Alibaba Qwen3.6 35B A3B โ 262K tokens context, reasoning model.
๐ฐ $0.25 in / $1.49 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
Anthropic Claude Opus 4.7 โ 1M tokens context, reasoning model.
๐ฐ $6.25 in / $25.00 out per 1M tok
๐ In: text, image
๐ค Out: text
Major Release
China Mobile JT-MINI โ 128K tokens context.
๐ฐ $0.00 in / $0.00 out per 1M tok
๐ In: text
๐ค Out: text
February
Full February 2026 report →
Major Release
Google's flagship reasoning model with a 2x jump on hard multi-step tasks. โ Google's latest flagship model with a major 2x jump in reasoning capabilities
๐ฐ $2.50 in / $10.00 out per 1M tok
๐ In: text, image, audio, video, PDF
๐ค Out: text
๐ Google AI Studio ยท Vertex AI ยท Gemini API
โจ Key Features
- 2x reasoning improvement
- ARC-AGI-2 score of 77.1%
- Enhanced multimodal understanding
- Deep Think mode
๐ What's new vs previous version
- 2x reasoning score on ARC-AGI-2 vs Gemini 3 Pro
- Context window expanded to 2M tokens
- Deep Think mode enabled by default on the Pro tier
- Lower latency on first-token despite larger context
Major Release
Near-Opus quality at a fraction of the cost, with Agent Teams orchestration. โ Anthropic's latest Sonnet with Agent Teams capability and near-Opus performance at a fraction of the cost
๐ฐ $3.00 in / $15.00 out per 1M tok
๐ In: text, image, PDF
๐ค Out: text
๐ Anthropic API ยท AWS Bedrock ยท Google Vertex AI
โจ Key Features
- Agent Teams: orchestrate 2-16 Claude instances
- Near-Opus performance at 1/5th cost
- 80.8% SWE-bench Verified
- Fast mode research preview
๐ What's new vs previous version
- Agent Teams: orchestrate 2โ16 Claude instances in parallel
- +8.5pt on SWE-bench Verified vs Sonnet 4
- 1/5 the cost of Opus 4.5 at ~95% of coding quality
- Fast mode research preview for lower-latency inference
Update
Open-weight MoE with a 1M+ token context window and strong coding. โ Major update with 10x context window expansion to over 1 million tokens
๐ฐ $0.27 in / $1.10 out per 1M tok
๐ In: text
๐ค Out: text
๐ DeepSeek API ยท Hugging Face ยท Together AI ยท Fireworks AI
โจ Key Features
- 1M+ token context window (10x expansion)
- Improved reasoning capabilities
- Open source release
- Cost-effective inference
๐ What's new vs previous version
- 10x context window expansion (128K โ 1M+ tokens)
- Sliding-window attention for long-context throughput
- Improved chain-of-thought reasoning
- Native FP8 inference support
Major Release
First frontier model trained entirely on Huawei Ascend silicon. โ First frontier AI model trained entirely without NVIDIA GPUs, using Huawei Ascend chips
๐ฐ $0.11 in / $0.28 out per 1M tok
๐ In: text, image
๐ค Out: text
๐ Zhipu BigModel API
โจ Key Features
- First frontier model trained on Huawei Ascend chips (no NVIDIA)
- #1 HLE score (50.4%)
- 1.2% hallucination rate via Slime RL
- 136x cheaper than Claude Opus 4.5
๐ What's new vs previous version
- Trained entirely on Huawei Ascend 910B clusters (no NVIDIA)
- Slime RL fine-tuning drops hallucination rate to 1.2%
- 136x cheaper than Claude Opus 4.5 at comparable quality
Major Release
Coding-specialized variant of GPT-5.3, tuned for agentic IDE workflows. โ OpenAI's specialized self-improving coding model with state-of-the-art software engineering performance
๐ฐ $1.25 in / $10.00 out per 1M tok
๐ In: text, image
๐ค Out: text
๐ OpenAI API ยท Azure OpenAI ยท GitHub Copilot
โจ Key Features
- Self-improving agentic coding
- 25% faster than GPT-5.2-Codex
- 1,000+ tokens/sec generation
- First OpenAI model flagged 'high' on cybersecurity framework
๐ What's new vs previous version
- +4pt on SWE-bench Verified vs GPT-5.2 Codex
- Native IDE tool-calling at reduced latency
- Extended max output to 100K for multi-file patches