Anthropic Launches Claude Sonnet 5, Closing In on Opus 4.8
Anthropic released Claude Sonnet 5, its most agentic mid-tier model, which it says approaches the pricier Opus 4.8 on many tasks at a lower per-token cost.
BLEU, or Bilingual Evaluation Understudy, is an automatic metric originally developed for machine translation. It compares short sequences of words in generated text with one or more reference translations, combines precision across several n-gram lengths, and applies a penalty when output is too short. BLEU is inexpensive and useful for comparing systems on the same dataset, but it does not directly measure factuality, fluency, meaning, or human preference. Valid alternative wording can receive a low score, while awkward text can match many reference phrases. Results depend on tokenization and implementation, so reporting should include the exact evaluation settings and, when possible, human judgment.
Anthropic released Claude Sonnet 5, its most agentic mid-tier model, which it says approaches the pricier Opus 4.8 on many tasks at a lower per-token cost.
The US Commerce Department lifted export controls on Anthropic’s Claude Fable 5 and Mythos 5, ending an 18-day standoff and restoring the models globally from July 1.
A new study of 22,000 US firms found that the heaviest AI spenders grew headcount faster than peers, including entry-level roles, complicating fears of an AI jobs apocalypse.
As Washington limits Anthropic’s and OpenAI’s top models on security grounds, critics warn the restrictions are handing momentum to Chinese open-weight rivals like GLM-5.2.
About $2.3 trillion was wiped from the Magnificent 7 in June as investors questioned Big Tech’s huge AI spending, even as chip and memory stocks kept climbing.
A record European heatwave is highlighting a growing threat to the AI boom: extreme weather that stresses data centers and the power grids they depend on at once.
Microsoft is heading for its worst month since 2008 as investors worry that its heavy AI spending may not pay off and that AI could erode demand for its core software.
South Korea announced about $880 billion in planned investment to expand chipmaking, AI data centers and robotics, aiming to double memory output and revive regions outside Seoul.
HP is scaling a strategic partnership with OpenAI, using the Frontier platform to roll out AI agents across customer support, software development and security after early pilots.
Google has limited Meta’s use of its Gemini AI models after Meta sought more computing capacity than Google could supply, delaying some of Meta’s internal projects.