Vitalik Buterin Says AI’s Real Danger Is Who Controls It
Ethereum’s Vitalik Buterin argued the gravest AI risk is not rogue superintelligence but a few companies or governments seizing control of it, in a widely shared thread on AI’s future.
Data poisoning is an attack that changes or inserts examples in a model’s training pipeline to corrupt what the system learns. An attacker may aim to reduce overall accuracy, create a hidden backdoor triggered by a specific pattern, bias decisions against a target, or influence a generative model’s responses. Poisoning can enter through public data collection, compromised suppliers, user feedback, labels, retrieval indexes, or repeated fine-tuning. Defenses include provenance tracking, access control, anomaly detection, robust training, duplicate analysis, trusted validation sets, and review of unexpected behavior. The attack is difficult to diagnose because harmful examples may appear ordinary and their effect may emerge only after deployment.
Ethereum’s Vitalik Buterin argued the gravest AI risk is not rogue superintelligence but a few companies or governments seizing control of it, in a widely shared thread on AI’s future.
Cloudflare launched Precursor, a bot-detection system that watches how visitors move, click and type across a whole session to tell humans from increasingly capable AI agents.
Anthropic published the cybersecurity rules behind its redeployed Claude Fable 5 model and proposed an industry framework for scoring how dangerous an AI jailbreak is.
OpenAI has proposed handing the US government a roughly 5% stake, worth about $43 billion, to seed a public wealth fund and ease mounting political pressure in Washington.
The US Commerce Department lifted export controls on Anthropic’s Claude Fable 5 and Mythos 5, ending an 18-day standoff and restoring the models globally from July 1.
GPT-5.6 Sol brings stronger coding and cyber capabilities with OpenAI’s most robust safeguards, but a government-approved rollout echoes the Anthropic model ban.
Anthropic told US senators that Alibaba-linked operators used about 25,000 fake accounts to extract Claude’s capabilities in what it calls its largest known distillation attack.
A new Anthropic privacy policy taking effect July 8 lets the company ask some flagged Claude users to upload government IDs and submit biometric selfies.
Anthropic CEO Dario Amodei published an essay urging binding, FAA-style safety testing for frontier AI models, marking a shift from the company’s earlier focus on transparency.
Anthropic released a new security plugin for Claude Code that reviews AI-generated code for vulnerabilities while it is being written. The system uses a separate Claude instance to scan changes in the background and automatically suggest fixes before code reaches pull requests.