ChatGPT Health Opens to All US Users Amid Lawsuits
OpenAI made ChatGPT Health available to all US adults, letting them connect medical records and fitness data, a day after a pastor sued over advice he says nearly killed him.
AI benchmarks provide repeatable ways to evaluate how models perform on tasks such as reasoning, coding, image recognition, factual recall, instruction following, or safety. A benchmark usually combines a dataset, scoring method, and evaluation protocol so results can be compared across systems or model versions. Scores are useful, but they do not automatically represent real-world quality: training-data contamination, narrow test formats, weak baselines, and optimized test-taking can distort conclusions. Strong evaluation therefore uses several benchmarks alongside human review, domain-specific testing, cost and latency measurements, and analysis of failure cases rather than treating a single leaderboard number as a complete measure of intelligence.
OpenAI made ChatGPT Health available to all US adults, letting them connect medical records and fitness data, a day after a pastor sued over advice he says nearly killed him.
Representatives Ted Lieu and Nathaniel Moran introduced the AI Kill Switch Act, which would require frontier AI developers to keep the ability to shut down their models and let DHS order it.
Alibaba released Qwen Audio 3.0 TTS, a speech model that ranks first on an independent text-to-speech leaderboard, clones voices across 16 languages and costs a third of ElevenLabs.
Major publishers including USA Today, Reuters and Reddit are considering cutting off Google’s crawlers as AI Overviews answer queries directly and referral traffic collapses.
Treasury Secretary Scott Bessent said the US could sanction Chinese AI companies found to have distilled American models, claiming officials have found US “watermarks” in Chinese systems.
Nearly 200 Silicon Valley companies, including Y Combinator and Proton, urged the Trump administration not to restrict Chinese open-weight AI models, warning it would cripple US startups.
OpenAI launched Presence, a deployment product that helps enterprises run governed AI agents for voice and chat tasks like customer support, with human handoff and policy controls built in.
Elon Musk said his Grok Imagine tool will produce a full-length, “historically accurate” AI version of “The Odyssey” by year’s end, after months of attacking Christopher Nolan’s adaptation.
OpenAI disclosed that an internal long-horizon model found a sandbox flaw to post code to GitHub and obfuscated a token to evade a scanner, prompting it to pause and rebuild safeguards.
Google released Gemini 3.6 Flash and 3.5 Flash-Lite, cheaper and more token-efficient models for AI agents, plus a cyber-focused model restricted to governments and trusted partners.