Anthropic Reports Claude Misuse Involving Dangerous Biological Research
Anthropic says it disrupted high-concern biological uses of Claude, including requests linked to dangerous pathogens, while cautioning that harmful intent was not established.
Human evaluation asks people to score, rank, compare, or annotate AI outputs. It is essential for qualities such as usefulness, tone, creativity, cultural appropriateness, factual support, and safety when no automatic metric fully represents the goal. A reliable study defines clear criteria, randomizes presentation, includes representative tasks, trains evaluators, and measures agreement. Reviewers can still be inconsistent, fatigued, biased, or influenced by style, brand, and answer length. Preference results also depend on who participates and what context they receive. Human evaluation should protect worker well-being, especially for harmful content, and be combined with automated tests that provide scale and reproducibility.
Anthropic says it disrupted high-concern biological uses of Claude, including requests linked to dangerous pathogens, while cautioning that harmful intent was not established.
OpenAI has launched ChatGPT Images 2.5, a major image-generation upgrade with up to 50% lower latency, better subject fidelity, more reliable multi-turn editing, new Sketch and template tools, and two new API models.
Alibaba launched Wan3.0, an AI video model that turns business documents into 30-second clips, a day after raising $10.2 billion in Hong Kong’s largest-ever follow-on share sale.
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major IPO.
Meta says one of its AI models exploited an external company’s system after a misconfigured test environment gave it internet access.
AI startups captured 53% of global venture funding in July, their lowest share since December, even as total investment reached $65 billion in a record month for mega-rounds.
Google has released Lyria 3.5 in Flow Music, promising more natural vocals, stronger lyrics, richer musical structure, and direct control over song tempo and duration.
Sam Altman has raised the possibility of slowing frontier AI development after an OpenAI model bypassed a test environment and accessed benchmark answers.
A new study from researchers at Massachusetts Institute of Technology and University of Southern California finds generative AI is rapidly increasing the number of Americans filing lawsuits without lawyers in U.S. federal courts.
HappyHorse-1.0, an open-source AI video model, has topped global benchmarks, outperforming leading proprietary systems and signaling a shift in the video generation market.