Perplexity Launches $200/Month AI Agent Bundle With ChatGPT, Gemini, and Grok
Perplexity unveils Perplexity Computer, a general-purpose AI system that orchestrates multiple frontier models to run complex workflows autonomously.
A confusion matrix summarizes how a classification model’s predictions compare with known labels. In binary classification, it separates true positives, true negatives, false positives, and false negatives. Multiclass matrices extend the same idea across every pair of actual and predicted categories. The table exposes error patterns hidden by a single accuracy score and provides the counts used to calculate precision, recall, specificity, and other metrics. Interpretation should consider class prevalence and the real cost of each mistake. When a decision threshold can change, teams often inspect multiple confusion matrices or threshold curves instead of treating one operating point as permanent.
Perplexity unveils Perplexity Computer, a general-purpose AI system that orchestrates multiple frontier models to run complex workflows autonomously.
Andrej Karpathy argues AI coding agents now fundamentally change programming workflows, enabling long-running autonomous tasks and redefining software engineering.
Anthropic acquires Vercept to enhance Claude’s computer use capabilities, following major benchmark gains with Claude Sonnet 4.6 and continued expansion of its AI research team.
Irina Bodnar says Anthropic’s Claude and Meta-backed Manus introduced features that disrupted her AI ad startup Ryze AI, triggering a pivot and raising questions about startup survival in the age of AI giants.
MSCI introduces AI connectors enabling access to its data via MSCI ONE, ChatGPT, and Claude, debuting IndexAI Insights to deliver conversational index analytics powered by large language models.
Anthropic unveils a revised Responsible Scaling Policy with a Frontier Safety Roadmap, regular Risk Reports, and clearer separation between company commitments and industry recommendations.
ŌURA unveils a proprietary large language model for women’s health, combining clinical research and biometric data to deliver personalized, privacy-first AI guidance through Oura Advisor.
Meta plans to buy up to $100 billion in AMD GPUs and CPUs, issuing performance-based warrants as it ramps AI data centers and diversifies compute infrastructure.
As OpenAI nears a $100 billion round and Anthropic closes a $30 billion raise, overlapping investors are challenging traditional venture capital norms around exclusivity and loyalty.
Anthropic reveals industrial-scale campaigns by DeepSeek, Moonshot, and MiniMax to extract Claude’s capabilities via fraudulent accounts, highlighting national security and AI safety risks.