Anthropic Doubles Public First Action Donation to $40M for AI Policy
Anthropic has committed an additional $20 million to Public First Action, bringing its total contribution to the nonpartisan AI policy organization to $40 million.
Explore AIstify's latest reporting, research, and expert analysis tagged with "ai security", collected in one continuously updated archive.
Anthropic has committed an additional $20 million to Public First Action, bringing its total contribution to the nonpartisan AI policy organization to $40 million.
Cloudflare launched Precursor, a bot-detection system that watches how visitors move, click and type across a whole session to tell humans from increasingly capable AI agents.
Anthropic published the cybersecurity rules behind its redeployed Claude Fable 5 model and proposed an industry framework for scoring how dangerous an AI jailbreak is.
OpenAI has proposed handing the US government a roughly 5% stake, worth about $43 billion, to seed a public wealth fund and ease mounting political pressure in Washington.
The US Commerce Department lifted export controls on Anthropic’s Claude Fable 5 and Mythos 5, ending an 18-day standoff and restoring the models globally from July 1.
GPT-5.6 Sol brings stronger coding and cyber capabilities with OpenAI’s most robust safeguards, but a government-approved rollout echoes the Anthropic model ban.
Anthropic told US senators that Alibaba-linked operators used about 25,000 fake accounts to extract Claude’s capabilities in what it calls its largest known distillation attack.
A new Anthropic privacy policy taking effect July 8 lets the company ask some flagged Claude users to upload government IDs and submit biometric selfies.
Anthropic CEO Dario Amodei published an essay urging binding, FAA-style safety testing for frontier AI models, marking a shift from the company’s earlier focus on transparency.
Prompt injection is an AI security attack that uses malicious instructions to override a model’s rules, expose data, or trigger unauthorized tool actions.
Anthropic released a new security plugin for Claude Code that reviews AI-generated code for vulnerabilities while it is being written. The system uses a separate Claude instance to scan changes in the background and automatically suggest fixes before code reaches pull requests.
Anthropic co-founder Chris Olah told the Vatican that AI development cannot be left solely to technology companies, warning about commercial incentives, labor disruption, and the growing complexity of frontier AI systems.
Anthropic says its unreleased Mythos AI model has identified more than 10,000 high- and critical-severity software vulnerabilities as part of Project Glasswing, a cybersecurity initiative focused on protecting critical infrastructure and open-source software.
A growing trend in China allows users to create AI replicas of former partners using personal data. The practice is raising concerns about privacy, emotional dependency, and relationships.
Attackers are using fake install guides for popular developer tools to trick users into running malicious commands. The campaign exploits trusted workflows like copy-paste terminal installs.