News · tag
pricing
6 posts
llm.blog dataset refresh: 14 current models, AA Intelligence Index adopted
The tracked model set is rebuilt against Artificial Analysis data as of August 11, 2026: Claude Opus 5 leads at 63, Kimi K3 tops open weights, and DeepSeek V4 Flash sets the price floor.
Meta's Muse Spark 1.2 and Alibaba's Qwen3.8 Max land in the frontier top ten
Two releases in three days reshuffle the mid-frontier: Meta goes proprietary with full multimodal input, and Alibaba undercuts every flagship at $2/$6 per million tokens.
Google cuts Gemini 3 Flash input pricing 17% to $0.25 per million tokensDraft
The second Flash-tier price cut this year puts Google level with GPT-5 mini on input cost — with a 1M-token context window.
OpenAI halves GPT-5.1 cached-input pricing to $0.125 per million tokensDraft
Cached input now costs 10% of the base rate — a direct subsidy for long-system-prompt agents and high-frequency RAG.
DeepSeek extends off-peak API discount window to 12 hours dailyDraft
Half-price inference now covers the full UTC night — and batch schedulers are already migrating.
Claude Sonnet 4.5's 1M-token context window graduates from betaDraft
Long-context pricing kicks in above 200K tokens. The beta header is gone; the premium tier is now the story.