OpenAI Unveils AI Child Safety Policy Blueprint
OpenAI has introduced a policy blueprint aimed at strengthening U.S. child safety protections in the age of AI. The framework focuses on laws, reporting standards, and built-in safeguards.
An adversarial attack attempts to make an AI system fail by supplying carefully designed input. The change may be almost invisible to a person while causing a model to misclassify an image, follow a malicious instruction, or reveal protected behavior. Attackers can exploit knowledge of the model or probe it through repeated queries. Defenses include adversarial training, input validation, access controls, monitoring, and tests that simulate realistic threats. No single defense removes every risk because attacks evolve alongside models. Security teams therefore evaluate the complete system, including connected tools, data pipelines, user permissions, and the consequences of an incorrect or manipulated output.
OpenAI has introduced a policy blueprint aimed at strengthening U.S. child safety protections in the age of AI. The framework focuses on laws, reporting standards, and built-in safeguards.
Anthropic has launched Project Glasswing with major tech partners to use advanced AI for identifying and fixing software vulnerabilities. The move comes as AI models reach unprecedented offensive cyber capabilities.
Google is adding new mental health features to Gemini, including crisis detection tools and direct hotline access. The company is also committing $30 million to expand global support services.
Anthropic has signed an agreement with the Australian government to collaborate on AI safety and research. The deal includes funding for scientific institutions and expanded use of Claude in healthcare and education.
Anthropic unveils a revised Responsible Scaling Policy with a Frontier Safety Roadmap, regular Risk Reports, and clearer separation between company commitments and industry recommendations.
OpenAI CEO Sam Altman said the creator of the viral AI agent OpenClaw is joining the company, while the project will continue as an open source initiative supported by OpenAI.
A new platform called Rent a Human allows AI agents to outsource tasks to real people when automation falls short, highlighting an unusual hybrid model of human-in-the-loop labor.
Artificial intelligence has become the top investment theme for global family offices, while cryptocurrencies remain largely sidelined, according to JPMorgan’s latest global survey.
SpaceX has acquired Elon Musk’s AI company xAI, combining rockets, satellites, and artificial intelligence into a vertically integrated effort aimed at scaling AI compute beyond Earth.
Moltbook, built on OpenClaw agentic AI, lets bots interact and form communities. Experts warn of security risks and governance challenges with AI-driven social networks.