IBM Shares Plunge 23% After Warning of a Quarterly Earnings Miss
IBM shares sank about 23% after it warned that preliminary second-quarter results fell short, as clients diverted spending toward memory and servers ahead of price increases.
AI benchmarks provide repeatable ways to evaluate how models perform on tasks such as reasoning, coding, image recognition, factual recall, instruction following, or safety. A benchmark usually combines a dataset, scoring method, and evaluation protocol so results can be compared across systems or model versions. Scores are useful, but they do not automatically represent real-world quality: training-data contamination, narrow test formats, weak baselines, and optimized test-taking can distort conclusions. Strong evaluation therefore uses several benchmarks alongside human review, domain-specific testing, cost and latency measurements, and analysis of failure cases rather than treating a single leaderboard number as a complete measure of intelligence.
IBM shares sank about 23% after it warned that preliminary second-quarter results fell short, as clients diverted spending toward memory and servers ahead of price increases.
Google DeepMind CEO Demis Hassabis proposed a US-led standards body to test the most powerful AI models before release and coordinate an industry slowdown if risks mount.
Google unveiled SensorFM, a health AI trained on over a trillion minutes of smartwatch data from 5 million people, that predicts 35 health measures from cardiovascular risk to depression.
Cloudflare launched Precursor, a bot-detection system that watches how visitors move, click and type across a whole session to tell humans from increasingly capable AI agents.
Tencent is in talks to become Manus’s largest shareholder as investors race to reverse Meta’s $2 billion purchase of the AI agent startup, after Beijing ordered the deal undone.
SpaceXAI’s Grok 4.5 tied GPT-5.6 Sol for first on the SWE-Atlas-QnA coding benchmark, a narrow win that highlights the model’s real edge: efficiency and low cost.
A South Korean developer says Anthropic’s billing system tried to charge him $16.6 million despite being on the free plan with zero API usage, one of several billing complaints.
Anthropic extended free Claude Fable 5 access to July 19, its second extension in a week, as OpenAI removed usage caps and made GPT-5.6 Sol cheaper to run.
Apple sued OpenAI for trade secret theft, alleging former employees now at OpenAI stole confidential hardware information to build its first AI device. OpenAI denies it.
SK Hynix raised $26.5 billion in its Nasdaq debut, the largest US listing ever by a foreign company, as investors rushed to buy into the AI memory chip boom.