Skip to content
Tech News
← Back to articles

Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard

read original more articles
Why This Matters

Opus 5's top ranking on the Artificial Analysis Intelligence Leaderboard highlights its advanced capabilities, but the article underscores the importance of understanding different caching cost structures across providers. This impacts the overall cost-efficiency and scalability for businesses deploying AI models, making informed decisions crucial for optimizing expenses.

Key Takeaways

The blended cache price shown here uses cache hit price only. Other caching costs differ by provider:

Anthropic: charges a separate cache write fee, with different rates for 5-minute and 1-hour TTLs (1-hour TTL is more expensive).

charges a separate cache write fee, with different rates for 5-minute and 1-hour TTLs (1-hour TTL is more expensive). Google (Vertex/Gemini): charges a per-hour cache storage fee in addition to cache hit pricing. Some providers also use tiered pricing for prompts above 200K tokens.

charges a per-hour cache storage fee in addition to cache hit pricing. Some providers also use tiered pricing for prompts above 200K tokens. OpenAI, DeepSeek, others: typically charge only cache hit pricing with no write or storage fee.

See Prompt Caching for the full breakdown.