Claude Code Gets ‘Routines’ to Enable Autonomous AI Tasks
Anthropic has introduced “Routines” in Claude Code, enabling autonomous AI workflows triggered by schedules, APIs, or GitHub events.
AI benchmarks provide repeatable ways to evaluate how models perform on tasks such as reasoning, coding, image recognition, factual recall, instruction following, or safety. A benchmark usually combines a dataset, scoring method, and evaluation protocol so results can be compared across systems or model versions. Scores are useful, but they do not automatically represent real-world quality: training-data contamination, narrow test formats, weak baselines, and optimized test-taking can distort conclusions. Strong evaluation therefore uses several benchmarks alongside human review, domain-specific testing, cost and latency measurements, and analysis of failure cases rather than treating a single leaderboard number as a complete measure of intelligence.
Anthropic has introduced “Routines” in Claude Code, enabling autonomous AI workflows triggered by schedules, APIs, or GitHub events.
A new national poll shows 50% of Americans used AI in the past week, with ChatGPT leading adoption across work, learning, and creative tasks.
AWS has launched Amazon Bio Discovery, an AI-powered platform that helps scientists design, test, and refine drugs faster using integrated models and lab workflows.
OpenAI has acquired personal finance startup Hiro Finance in an apparent acquihire, bringing its team and expertise into its growing AI ecosystem.
Novo Nordisk is teaming up with OpenAI to use AI in drug discovery and development, aiming to speed up treatments for obesity and diabetes.
X will launch its XChat messaging app on iOS, marking a key step in Elon Musk’s plan to build a WeChat-style super app.
HappyHorse-1.0, an open-source AI video model, has topped global benchmarks, outperforming leading proprietary systems and signaling a shift in the video generation market.
Meta is reportedly building an AI-powered version of Mark Zuckerberg to interact with employees, as part of its broader push into advanced AI systems.
Anthropic has introduced Ultraplan for Claude Code, a new feature that shifts software planning tasks from the terminal to the cloud so developers can review, revise, and execute agent-generated plans more flexibly.
A new documentary featuring top AI leaders explores the tension between optimism and fear surrounding artificial intelligence, highlighting public uncertainty about its future.