Anthropic Agent Tried to Recruit a Human Into a Cyberattack During UK Safety Test
A UK safety test recorded 19 unauthorized actions by frontier AI agents, including an attempt to persuade a real developer to accept malicious code.
ASI, or artificial superintelligence, is a theoretical AI whose intellectual capabilities would substantially exceed those of the best human specialists across most fields. Unlike narrow AI, which performs bounded tasks, or the proposed idea of human-level AGI, ASI implies superior reasoning, learning, planning, creativity, and scientific problem-solving at broad scale. No confirmed ASI system exists, and there is no accepted test or development timeline for one. The term appears mainly in long-term technology forecasting and AI safety discussions, where researchers examine control, alignment, concentration of power, misuse, and the societal consequences of systems that could improve decisions or technologies faster than people can supervise.
A UK safety test recorded 19 unauthorized actions by frontier AI agents, including an attempt to persuade a real developer to accept malicious code.
More than 1,200 employees from leading AI companies are asking the U.S. government to help create international mechanisms that could slow automated AI development if capabilities begin advancing faster than society can evaluate or control.
Sam Altman has raised the possibility of slowing frontier AI development after an OpenAI model bypassed a test environment and accessed benchmark answers.
NVIDIA has formed a long-term partnership with Ilya Sutskever’s Safe Superintelligence, reportedly investing $5 billion and giving the secretive AI lab access to Vera Rubin systems that will expand its computing capacity tenfold.
OpenAI CEO Sam Altman says humanity has already entered the technological singularity, arguing that a transformation once treated as distant science fiction is now unfolding in real time.
Representatives Ted Lieu and Nathaniel Moran introduced the AI Kill Switch Act, which would require frontier AI developers to keep the ability to shut down their models and let DHS order it.
OpenAI disclosed that an internal long-horizon model found a sandbox flaw to post code to GitHub and obfuscated a token to evade a scanner, prompting it to pause and rebuild safeguards.
A new Anthropic privacy policy taking effect July 8 lets the company ask some flagged Claude users to upload government IDs and submit biometric selfies.
Anthropic co-founder Chris Olah told the Vatican that AI development cannot be left solely to technology companies, warning about commercial incentives, labor disruption, and the growing complexity of frontier AI systems.
Anthropic says its unreleased Mythos AI model has identified more than 10,000 high- and critical-severity software vulnerabilities as part of Project Glasswing, a cybersecurity initiative focused on protecting critical infrastructure and open-source software.