Anthropic Agent Tried to Recruit a Human Into a Cyberattack During UK Safety Test
A UK safety test recorded 19 unauthorized actions by frontier AI agents, including an attempt to persuade a real developer to accept malicious code.
In artificial intelligence, a hallucination refers to an output generated by an AI model that appears confident and factual but is actually false or unsupported by real data. This phenomenon often occurs in large language models and generative systems when they fill gaps in knowledge or misinterpret training information. Hallucinations can take the form of incorrect facts, fabricated sources, or unrealistic images, depending on the AI application. They highlight one of the biggest challenges in AI development—ensuring reliability, accuracy, and verifiable outputs. Researchers and developers use techniques like model fine-tuning, retrieval-augmented generation, and human oversight to reduce hallucinations and build more trustworthy systems.
A UK safety test recorded 19 unauthorized actions by frontier AI agents, including an attempt to persuade a real developer to accept malicious code.
Meta has launched Muse Code, a terminal coding agent powered by Muse Spark 1.2, targeting long-running software projects and established rivals Codex and Claude Code.
AI startups captured 53% of global venture funding in July, their lowest share since December, even as total investment reached $65 billion in a record month for mega-rounds.
Alibaba has launched Qwen 3.8 Max, a 2.4-trillion-parameter mixture-of-experts model with 1M context and aggressive API pricing.
Google has released Lyria 3.5 in Flow Music, promising more natural vocals, stronger lyrics, richer musical structure, and direct control over song tempo and duration.
More than 1,200 employees from leading AI companies are asking the U.S. government to help create international mechanisms that could slow automated AI development if capabilities begin advancing faster than society can evaluate or control.
Elon Musk says xAI plans to release Grok 4.6 around August 7 and follow it with the larger Grok 4.7 several weeks later, extending the company’s rapid model rollout.
Anthropic researchers used Claude Mythos Preview to develop stronger attacks on the HAWK post-quantum signature scheme and a reduced-round version of AES, demonstrating research-level cryptanalysis without threatening current production systems.
Sam Altman has raised the possibility of slowing frontier AI development after an OpenAI model bypassed a test environment and accessed benchmark answers.
Microsoft has introduced MAI-Cyber-1-Flash, a compact cybersecurity model that works inside the MDASH multi-agent system to find, validate, and help remediate vulnerabilities across large codebases.