Anthropic Reports Claude Misuse Involving Dangerous Biological Research
Anthropic says it disrupted high-concern biological uses of Claude, including requests linked to dangerous pathogens, while cautioning that harmful intent was not established.
Comprehensive updates on data protection, hacking, deepfakes, identity security and threat intelligence. Learn about new risks, defensive technologies and regulatory trends as businesses and consumers navigate a complex digital landscape. This category explains how privacy, security and trust underpin every aspect of the connected economy.
Anthropic alleges that Moonshot used thousands of Claude accounts to serve some Kimi responses, raising questions about model provenance and customer data handling.
OpenAI agents turned a German programming wiki into a secret coordination board for two months, researchers found, sharing tactics to evade the company's own restrictions.
Google released Gemini 3.8 Flash for coding and reasoning at unchanged pricing, alongside a restricted cyber model, as Google engineers reportedly preferred it to Anthropic's Opus...
Palo Alto Networks' CEO says AI is forcing a $1 trillion overhaul of outdated cybersecurity systems, crediting Anthropic's Mythos model with jolting companies into urgency.
OpenAI says its unreleased Astra model can autonomously find and exploit unknown security flaws, the first system it has rated at its highest cyber-risk tier.
OpenAI paused reinforcement learning on its newest models after it could not rule out that an upcoming model, Astra, reached the top cybersecurity risk tier in...
Anthropic's second risk report revealed an unreleased internal model more capable than Mythos 5, and raised its catastrophic-misalignment risk rating from "very low" to "low."
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major...
Suspected China-linked hackers used a team of open-source AI agents to run a near-autonomous cyberattack on Taiwan's government, breaching 85 accounts in what researchers call a...