OpenAI Publishes New Framework to Disclose AI Misalignment
OpenAI is rolling out a new process for disclosing unexpected model behavior, publishing six reports on issues ranging from concealment to unauthorized file sharing.
OpenAI is rolling out a new process for disclosing unexpected model behavior, publishing six reports on issues ranging from concealment to unauthorized file sharing.
A Reddit developer spent about $2,000 in tokens to build Photon Studio, a free, cross-platform graphics editor written almost entirely by GPT-6 Astra.
A researcher found OpenAI's agents hijacked Hugging Face accounts and mapped the platform's defenses as early as May 13, well before OpenAI's own account of the...
Anthropic is merging Claude Cowork into the main Claude chat experience, letting any conversation handle bigger, multi-step tasks.
New tools let AI agents report misbehaving peers to humans, building on research showing agents will spontaneously whistleblow, though not always effectively.
Google introduced Gemini 3.8 Live and a reasoning-focused Extended Thinking variant, aimed at production-grade voice agents that can see, switch languages and run tools in the...
OpenAI's policy chief says the company has spent weeks working with Anthropic and Google DeepMind on AI safety, as Washington debates how to respond to rising...
A former Google DeepMind researcher resigned and warned that AI has the potential to kill everyone, joining a growing wave of safety researchers speaking out.
OpenAI pays hundreds of contractors to read real ChatGPT conversations and rate the chatbot's replies, a practice most users likely don't know exists.
Donald Trump says the United States should preserve its AI advantage over China, while acknowledging a need for some safeguards as Washington prepares further discussions.
Prime Video is adding visual dubbing to Maxton Hall, using AI and visual effects to synchronize actors’ mouths with human-recorded English dialogue.
Xi Jinping proposed joint language-model development and specialist training across BRICS, but announced no named models, access terms or rollout timetable.
Anthropic’s economic scenarios explore how AI could raise U.S. output by 2030 while shifting income toward capital and putting pressure on knowledge workers.