Skip to content

Safety

This course is about using AI without getting hurt by it. It covers what you may share and how far to trust an answer, and what can go wrong once an agent is allowed to act on your behalf.

0%
  1. getting started
  2. halfway
  3. almost there
  4. complete
Lesson plan (18 lessons, 18 live)
#LessonStatusModeMinutesCoversServesExerciseAfterIssue
Responsible use
1What may go into an AI toolsafety/responsible-useliveexplanation35 estimateresponsible-usedecides-what-to-sharediscloses-ai-userespects-licensesdo
2Redacting a document before you paste itsafety/redact-before-you-pastelivetutorial30 estimateresponsible-usedecides-what-to-sharedoWhat may go into an AI tool
3Saying that AI helped, and crediting what it copiedsafety/saying-ai-helpedlivetutorial25 estimateresponsible-usediscloses-ai-userespects-licensesdoWhat may go into an AI tool
Recognizing failure
4What an invented fact looks likesafety/spotting-hallucinationlivetutorial25 estimatefailure-modeschecks-claimsjudge
5Seeing bias across many answerssafety/bias-in-patternslivetutorial30 estimatefailure-modescalibrates-trustjudgeWhat an invented fact looks like
6When the tool is usually rightsafety/when-the-tool-is-usually-rightliveexplanation20 estimatefailure-modescalibrates-trustjudgeWhat an invented fact looks like
Verifying outputs
7Checking habits that run when you are in a hurrysafety/checking-habitslivetutorial25 estimateverificationchecks-claimscalibrates-trustdoWhat an invented fact looks like, When the tool is usually right#120
8Following a claim back to its sourcesafety/following-the-sourcelivetutorial20 estimateverificationchecks-claimsjudgeChecking habits that run when you are in a hurry#122
9Asking a model to disagreesafety/asking-a-model-to-disagreelivetutorial20 estimateverificationspots-sycophancydoChecking habits that run when you are in a hurry#124
10Marking what a person has checkedsafety/marking-what-a-person-checkedliveexplanation35 estimateverificationchecks-claimskeeps-a-check-habitjudgeFollowing a claim back to its source#126
11Checking what an agent changedsafety/checking-what-an-agent-changedlivetutorial20 estimateverificationcalibrates-trustdoChecking habits that run when you are in a hurry#128
Agent risk
12Why agent safety is differentsafety/agent-riskliveexplanation30 estimateagent-risknames-blast-radiuschooses-human-in-looprecognizes-injectiondo
13Sizing an agent's blast radiussafety/sizing-the-blast-radiuslivetutorial20 estimateagent-risknames-blast-radiuschooses-human-in-loopdoWhy agent safety is different, Checking what an agent changed#130
14Tracing a planted instruction to the tool that leakssafety/tracing-a-planted-instructionlivetutorial20 estimateagent-riskrecognizes-injectionjudgeWhy agent safety is different#132
Governance and oversight
15Writing an AI policy a team can applysafety/writing-an-ai-policylivetutorial25 estimategovernancesets-oversightdoWhy agent safety is different#135
16Reading agent logs and rerunning after a model changesafety/reading-agent-logslivetutorial25 estimategovernancesets-oversightdoWriting an AI policy a team can apply#138
17Assessing the risk of a use case before it startssafety/assessing-a-use-caselivetutorial25 estimategovernancesets-oversightnames-blast-radiusdoWriting an AI policy a team can apply#141
18Introduction to the EU AI Actsafety/eu-ai-actliveexplanation30 estimategovernancesets-oversightdiscloses-ai-usejudgeWhy agent safety is different