European Banks Scrutinize Anthropic’s Mythos Model Over Cybersecurity Risks
German banks and regulators are assessing risks tied to Anthropic’s Mythos model as concerns grow over AI-driven cyber threats to financial systems.
When an AI system follows the letter of an instruction but misses its intent, the problem is one of AI alignment. Alignment aims to keep model behavior consistent with human goals, safety constraints, and acceptable social values, including in unfamiliar situations. The challenge is that people express preferences imperfectly, measurable objectives can reward shortcuts, and values may conflict across users or contexts. Researchers and product teams address this gap through preference training, behavioral policies, red-team evaluations, access controls, uncertainty handling, and human review. Alignment is therefore not a single training technique; it is an ongoing process that combines model design, governance, testing, and operational oversight.
German banks and regulators are assessing risks tied to Anthropic’s Mythos model as concerns grow over AI-driven cyber threats to financial systems.
OpenAI is scaling its Trusted Access for Cyber program and introducing GPT-5.4-Cyber to support vetted defenders as AI-driven security risks accelerate.
OpenAI has introduced a policy blueprint aimed at strengthening U.S. child safety protections in the age of AI. The framework focuses on laws, reporting standards, and built-in safeguards.
Google is adding new mental health features to Gemini, including crisis detection tools and direct hotline access. The company is also committing $30 million to expand global support services.
Anthropic has signed an agreement with the Australian government to collaborate on AI safety and research. The deal includes funding for scientific institutions and expanded use of Claude in healthcare and education.
Anthropic has launched the Anthropic Institute to study the societal, economic, and governance challenges posed by advanced AI systems. The initiative will combine research from engineers, economists, and social scientists.
Despite President Trump’s directive to cease federal use of Anthropic’s Claude AI, U.S. military forces reportedly employed the model for intelligence, target selection, and battlefield simulations in airstrikes on Iran.
Anthropic unveils a revised Responsible Scaling Policy with a Frontier Safety Roadmap, regular Risk Reports, and clearer separation between company commitments and industry recommendations.
A new platform called Rent a Human allows AI agents to outsource tasks to real people when automation falls short, highlighting an unusual hybrid model of human-in-the-loop labor.
Artificial intelligence has become the top investment theme for global family offices, while cryptocurrencies remain largely sidelined, according to JPMorgan’s latest global survey.