Standard A starting point · leads to 2
AI Safety
Research and practices aimed at ensuring AI systems are safe, reliable, and beneficial, especially as capabilities increase.
Where it sits
Explore nearby
Shipping AI AI Alignment Ensuring AI systems behave in accordance with human values and intentions, a central challenge in AI safety. Evaluation Robustness A model's ability to maintain performance under distribution shifts, adversarial attacks, or noisy inputs. Shipping AI Interpretability Understanding the internal workings of AI models, including which features influence predictions and why. Shipping AI AI Governance Policies, frameworks, and practices for responsible development and deployment of AI systems. Shipping AI Fairness Ensuring AI systems treat all individuals and groups equitably, without discrimination based on protected attributes.