Risk and safety

AI safety

The field concerned with preventing harm from AI systems, from everyday failures such as biased or false outputs to serious risks from misuse or loss of control of highly capable systems. People use the term with very different scopes, so ask which is meant.


Related terms

AI incident

An event where an AI system causes or nearly causes harm, such as a chatbot giving customers wrong refund terms or an …

AI text detector

A tool that claims to tell whether text was written by AI. They are unreliable: they produce false positives, particul…

Adversarial example

An input subtly altered to fool a model, such as an image with changes invisible to people that makes a classifier see…

Alignment

Making an AI system's behaviour match the intentions and values of the people it serves, including in situations its d…

Automation bias

The human tendency to over-trust an automated system's output and stop checking it, especially when it is usually righ…

Bias

Systematic unfairness in an AI system's outputs, such as a CV screener that favours one group or an image generator th…

Beyond the definition

Knowing the word is the easy part

A short assessment scores you across five skill areas and builds a path through 18 modules, skipping whatever you already know. Free, no card.

Find your level