Anthropic Reports Claude Misuse Involving Dangerous Biological Research
Anthropic says it disrupted high-concern biological uses of Claude, including requests linked to dangerous pathogens, while cautioning that harmful intent was not established.
An AI system can appear accurate overall while consistently producing worse results for certain people, situations, or types of data; this pattern is known as bias. It may enter a model through unrepresentative training examples, historical inequalities, labeling decisions, feature selection, optimization targets, or the environment in which the system is deployed. Bias can affect hiring tools, facial recognition, lending models, medical software, and other high-impact applications. Reducing it requires more than removing sensitive fields, because related variables may preserve the same patterns. Teams typically examine data coverage, compare performance across relevant groups, document limitations, test real-world outcomes, and involve domain experts and affected communities throughout development and monitoring.
Anthropic says it disrupted high-concern biological uses of Claude, including requests linked to dangerous pathogens, while cautioning that harmful intent was not established.
OpenAI has launched ChatGPT Images 2.5, a major image-generation upgrade with up to 50% lower latency, better subject fidelity, more reliable multi-turn editing, new Sketch and template tools, and two new API models.
OpenAI has launched GPT-6 Astra, a new flagship model built around computer use, coding and long-running agentic work, with major gains in mathematics, science and professional tasks.
Anthropic has released Claude Fable 5.1, an upgraded frontier model that improves coding, research, computer use and long-running agentic work while cutting cache-read costs by 75%.
OpenAI launched AI Futures, a blog from a new Strategic Futures team arguing that AI’s gravest risk is letting power escape the checks that have long depended on human cooperation.
Nvidia has discussed investing in Mercor, a data-labeling startup that supplies expert data for its Nemotron models, in a round that would double the company’s valuation to $20 billion.
OpenAI launched ChatGPT for Teens, a version with default safety protections, learning tools and parental controls, automatically applied to users it identifies as under 18.
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major IPO.
Anthropic has reportedly signed a 20-year, $9.1 billion data center agreement with Riot Platforms for 191 MW of AI computing capacity at the bitcoin miner’s Rockdale campus in Texas.
Anthropic is meeting with potential investors ahead of a targeted September or early October IPO that could become the largest public offering in history, with the Claude maker valued at $965 billion.