technspire
Home
TeamBlogGuides
← Back to Blog

Posts tagged with "LLM Cost Optimization"

Found 2 posts

AI & Machine Learning
August 31, 2026

Anthropic prompt caching pricing: write, read and TTL math

Anthropic prices prompt caching with three numbers: a 1.25x or 2x premium on cache writes depending on TTL, a 0.1x rate on cache reads, and the base input rate for everything after the last breakpoint. This deep-dive verifies every figure against the current official docs and covers per-model minimums, break-even math, batch stacking and how Claude on Azure converts it all into CCUs.

Prompt Caching
Anthropic
Claude
Claude API
LLM Cost Optimization
Microsoft Foundry
Azure
Batch API
By Falak Mahmood
AI & Machine Learning
July 13, 2026

GPT-5.6 Sol vs Terra vs Luna: an Azure routing playbook

OpenAI released the GPT-5.6 series on 9 July 2026 in three tiers: Sol for hard reasoning and long autonomous runs, Terra for everyday work and Luna for speed and cost, with same-day availability in Microsoft Foundry and a new preferred-model role in Microsoft 365 Copilot. This playbook maps Azure workloads to the right tier, works the token math in SEK and flags the Copilot subprocessor setting Swedish admins must review before 24 July.

GPT-5.6
OpenAI
Azure OpenAI
Microsoft Foundry
Model Routing
LLM Cost Optimization
Microsoft 365 Copilot
ChatGPT Work
EU Data Residency
By Falak Mahmood
technspire

Leading provider of AI services, cloud development, and digital transformation solutions for Swedish enterprises and government agencies.

Org.nr: 559022-9422
VAT: SE559022942201

Services

  • Azure OpenAI Integration
  • Next.js & React Development
  • TypeScript Modernization
  • Payment System Integration
  • On-Premise AI Solutions
  • Cloud Migration

Company

  • Solution Examples
  • Our Team
  • Blog
  • Guides
  • Contact

Contact

  • Markörvägen 1a
    Stockholm
    Sweden
  • hello@technspire.com
© 2026 Technspire AB. All rights reserved.
Privacy PolicyTerms of ServiceCookie Policy