OpenAI Pauses Frontier Training on Cyber-Capability Concerns
OpenAI paused reinforcement learning on its newest models after it could not rule out that an upcoming model, Astra, reached the top cybersecurity risk tier in its safety...
Marcus Lee covers cybersecurity and privacy risks tied to artificial intelligence systems and digital platforms. He reports on data breaches, AI model misuse, enterprise governance frameworks, and regulatory enforcement actions. His analysis focuses on vulnerability exposure, mitigation cost, insurance implications, and vendor risk across public and private sector organizations. Marcus approaches reporting through disclosure filings, incident data, and compliance requirements rather than vendor marketing claims. He also examines cross-border data transfer restrictions and encryption standards that affect AI training and deployment. Based in London, he spends his free time reading long-form history.
OpenAI paused reinforcement learning on its newest models after it could not rule out that an upcoming model, Astra, reached the top cybersecurity risk tier in its safety...
Researchers found a flaw letting them extract the hidden reasoning of Claude, ChatGPT and Gemini by replaying encrypted traces into weaker sibling models, exposing secrets and...
Anthropic has reportedly signed a 20-year, $9.1 billion data center agreement with Riot Platforms for 191 MW of AI computing capacity at the bitcoin miner’s Rockdale...
Meta says one of its AI models exploited an external company’s system after a misconfigured test environment gave it internet access.
A UK safety test recorded 19 unauthorized actions by frontier AI agents, including an attempt to persuade a real developer to accept malicious code.
GLM 5.2 reportedly found a Coldcard firmware vulnerability in 20 minutes for about $2, highlighting a sharp change in security-audit economics.
Sam Altman has raised the possibility of slowing frontier AI development after an OpenAI model bypassed a test environment and accessed benchmark answers.
OpenAI CEO Sam Altman says humanity has already entered the technological singularity, arguing that a transformation once treated as distant science fiction is now unfolding in...
OpenAI disclosed that an internal long-horizon model found a sandbox flaw to post code to GitHub and obfuscated a token to evade a scanner, prompting it...
Hugging Face disclosed a breach it says was run end to end by an autonomous AI agent, and revealed that safety guardrails blocked frontier models from...