Products & Services
Claude Sonnet 5.5 launches with more than 30% faster output
Claude Sonnet 5.5 speeds up output while keeping API rates unchanged. We compare performance and costs in evaluations from Artificial Analysis, Vals AI, Mercor, and others, and explain availability and API migration considerations.
Products & ServicesMeta launches Muse, a personal AI agent
Muse carries out research and operates services in a personal cloud environment, requesting approval when needed. We examine its rollout, the work it can handle, and how it manages data.
Products & ServicesClaude Opus 5.5 launches with roughly 60% lower cost per evaluation task
Claude Opus 5.5 improves its score in an independent evaluation while reducing cost per task by about 60%. We explain performance, pricing and practical caveats.
Products & ServicesGPT-6 Sol and Luna launch with API rates cut by half or more
In independent evaluation, cost per task falls about 47% for Sol and 61% for Luna. Overall performance is near the previous generation, with gains in avoiding wrong answers and business automation but weaker results in document creation and other work.
Products & ServicesGemini 3.8 Live launches, keeping conversation going while work runs
The two models differ in handling complex requests, acknowledgments and interruptions. We explain Japanese support, the cost of a whole conversation and availability across apps.
Products & ServicesApple releases iOS 27 and the English beta of the new Siri
Using models built in collaboration with Google's Gemini, the new Siri finds information in messages and email and operates apps. We explain supported iPhones and the planned timing for Japanese.
Products & ServicesGPT-6 Astra launches with improved computer-use performance
GPT-6 Astra scores higher on computer-use evaluations and completes tasks faster. We explain performance and pricing changes, availability and usage restrictions.
Products & ServicesAnthropic announces Claude Fable 5.1 and Mythos 5.1
Fable 5.1 outperforms its predecessor on coding and research evaluations, while cached-input reads cost one quarter as much. We explain when costs fall and how it differs from the more restricted Mythos.
Products & Services