OpenAI Pauses Frontier Training on Cyber-Capability Concerns
OpenAI paused reinforcement learning on its newest models after it could not rule out that an upcoming model, Astra, reached the top cybersecurity risk tier in its safety framework.
Anomaly detection identifies observations that differ significantly from an expected pattern. A bank may use it to flag unusual transactions, a manufacturer to spot equipment faults, or a security platform to detect suspicious network activity. Some systems learn from labeled examples of normal and abnormal events, while others model normal behavior and treat large deviations as potential anomalies. Rare does not always mean harmful, so alerts require context and careful thresholds. Effective detection balances missed events against false alarms, adapts as behavior changes, and gives reviewers enough information to investigate why a record, sequence, or sensor reading was considered unusual.
OpenAI paused reinforcement learning on its newest models after it could not rule out that an upcoming model, Astra, reached the top cybersecurity risk tier in its safety framework.
Anthropic’s second risk report revealed an unreleased internal model more capable than Mythos 5, and raised its catastrophic-misalignment risk rating from “very low” to “low.”
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major IPO.
Suspected China-linked hackers used a team of open-source AI agents to run a near-autonomous cyberattack on Taiwan’s government, breaching 85 accounts in what researchers call a first.
Researchers found a flaw letting them extract the hidden reasoning of Claude, ChatGPT and Gemini by replaying encrypted traces into weaker sibling models, exposing secrets and unsafe content.
Anthropic has reportedly signed a 20-year, $9.1 billion data center agreement with Riot Platforms for 191 MW of AI computing capacity at the bitcoin miner’s Rockdale campus in Texas.
Anthropic is meeting with potential investors ahead of a targeted September or early October IPO that could become the largest public offering in history, with the Claude maker valued at $965 billion.
Meta has released Muse Glimmer, a 30-billion-parameter open-weight model built for local AI agents, coding, tool use, and multimodal tasks on a single consumer GPU or high-end Mac.
Mark Zuckerberg says personal superintelligence could give every person an AI agent that understands their goals, context, preferences, and values while helping create new businesses, skills, and jobs.
Cloudflare says AI agents, bots, and other automated systems could generate up to 1,000 times more internet traffic than humans within five years as machine activity accelerates.