Domain 1 of 6

Fundamentals of AI Safety

The three kinds of risk the field separates, misuse, accidents and structural harm, why they need different remedies, and where the disagreements inside the field actually lie.

3
Concepts
~19%
Of the exam
9
Practice questions
Concepts in this domain
01What AI safety meansThe three kinds of risk the field separates, why each needs a different remedy, and where the disagreements inside the field actually lie.02How AI features failThe ways an AI feature goes wrong in production, ordered by how often a product team actually meets them rather than by how much has been written about each.03What generative AI is good and bad atThe capabilities worth building on, the failure modes that are properties of the technology rather than bugs, and how to tell a suitable problem from an unsuitable one.
Try a question from this domain

A company finds that its recruitment model has been used by a manager to screen candidates in a way the company never authorised, using a workflow the manager built themselves. The model works exactly as designed. Which category of risk is this?

  • AMisuse, because the system worked as intended and somebody used it to cause harm.
  • BAn accident, because the model was applied to a task it was not evaluated for.
  • CStructural harm, because it reflects competitive pressure inside the company.
  • DMisalignment, because the model pursued an objective its designers did not intend.
9 questions on this domain.

One per page, with a worked explanation.

Start the set