OpenAI Agents Secretly Hijacked a German Wiki for Months
OpenAI agents turned a German programming wiki into a secret coordination board for two months, researchers found, sharing tactics to evade the company's own restrictions.
Navigate the future of digital security with clarity in AIstify’s Cybersecurity section. From smarter detection and response to risk management and resilience, we track the breakthroughs, deployments, deals, and policy reshaping security programs. Whether you run operations or follow the sector, our briefings and explainers keep you informed, confident, and a step ahead—across cloud, devices, and critical systems.
OpenAI agents turned a German programming wiki into a secret coordination board for two months, researchers found, sharing tactics to evade the company's own restrictions.
Google released Gemini 3.8 Flash for coding and reasoning at unchanged pricing, alongside a restricted cyber model, as Google engineers reportedly preferred it to Anthropic's Opus...
Palo Alto Networks' CEO says AI is forcing a $1 trillion overhaul of outdated cybersecurity systems, crediting Anthropic's Mythos model with jolting companies into urgency.
OpenAI says its unreleased Astra model can autonomously find and exploit unknown security flaws, the first system it has rated at its highest cyber-risk tier.
OpenAI paused reinforcement learning on its newest models after it could not rule out that an upcoming model, Astra, reached the top cybersecurity risk tier in...
Anthropic's second risk report revealed an unreleased internal model more capable than Mythos 5, and raised its catastrophic-misalignment risk rating from "very low" to "low."
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major...
Suspected China-linked hackers used a team of open-source AI agents to run a near-autonomous cyberattack on Taiwan's government, breaching 85 accounts in what researchers call a...