Hugging Face Says an Autonomous AI Agent Breached Its Systems
Hugging Face disclosed a breach it says was run end to end by an autonomous AI agent, and revealed that safety guardrails blocked frontier models from...
AI security is now a core part of cybersecurity. In AIstify’s AI Security section, we cover how models are attacked, defended, and operated safely - from prompt injection and data leakage to supply-chain risk and model misuse. We track vendor tooling, red-teaming, evaluations, and the policies shaping secure deployment across cloud and edge. Whether you are defending systems or building them, this hub keeps you current on threats, mitigations, and the standards emerging around trustworthy AI.
Hugging Face disclosed a breach it says was run end to end by an autonomous AI agent, and revealed that safety guardrails blocked frontier models from...
Twenty-six Meta workers sued the company, alleging its AI systems used metrics like token consumption to select layoff targets, disadvantaging people on medical or family leave.
Cloudflare launched Precursor, a bot-detection system that watches how visitors move, click and type across a whole session to tell humans from increasingly capable AI agents.
Anthropic published the cybersecurity rules behind its redeployed Claude Fable 5 model and proposed an industry framework for scoring how dangerous an AI jailbreak is.
OpenAI has proposed handing the US government a roughly 5% stake, worth about $43 billion, to seed a public wealth fund and ease mounting political pressure...
The US Commerce Department lifted export controls on Anthropic's Claude Fable 5 and Mythos 5, ending an 18-day standoff and restoring the models globally from July...
GPT-5.6 Sol brings stronger coding and cyber capabilities with OpenAI's most robust safeguards, but a government-approved rollout echoes the Anthropic model ban.
Anthropic told US senators that Alibaba-linked operators used about 25,000 fake accounts to extract Claude's capabilities in what it calls its largest known distillation attack.
A new Anthropic privacy policy taking effect July 8 lets the company ask some flagged Claude users to upload government IDs and submit biometric selfies.
Anthropic's Fable 5 and Mythos 5 shutdown, the first US export control on an LLM, has Carney and EU lawmakers pushing to reduce reliance on American...