OpenAI Disbands Its Catastrophic-Risk Preparedness Team
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major IPO.
Explore AIstify's latest reporting, research, and expert analysis tagged with "responsible ai", collected in one continuously updated archive.
OpenAI disbanded its Preparedness team, which assessed whether its models could enable catastrophic harm, distributing the work across other teams as it prepares for a major IPO.
A black box model is an AI system whose internal decision process is difficult to inspect or explain, even when its predictions are useful.
Anthropic has committed $10 million CAD to Canadian research institutions to fund responsible AI development, announcing partnerships with eight universities and institutes including Mila, the Vector Institute, and Amii.
Partnership on AI is a nonprofit coalition that brings together companies, civil society groups, academics, and media organizations to study responsible AI practices.
A model card is structured documentation describing an AI model’s intended uses, evaluation results, limitations, and risks.
Bias is a systematic distortion in AI results that can produce unfair or inaccurate outcomes for particular groups, situations, or data patterns.
Explainable AI uses methods that help people understand, evaluate, and challenge how an AI system reaches a decision or prediction.
Human evaluation uses people to judge AI outputs on qualities that automated metrics cannot measure reliably.
Human-centered AI designs artificial intelligence around human goals, well-being, agency, accessibility, and accountable oversight.
AI alignment is the practice of directing AI systems toward intended goals and human values while keeping their behavior safe and controllable.
Equalized odds requires a classifier to have matching true-positive and false-positive rates across protected groups.
Differential privacy limits how much the presence of one person’s data can influence a released result or trained model.
Demographic parity is satisfied when a model produces a selected outcome at equal rates across defined demographic groups.
A white box model is an AI system with transparent, inspectable decision logic that supports explanation, auditing, debugging, and human review.
Human-in-the-loop (HITL) is an AI workflow where people review, guide, correct, or approve model decisions and sensitive automated actions.