Build the layer that decides what a system will not do, well enough to say which defence stops which attack, to measure what each one costs in wrongly blocked requests, and to stop treating a prompt instruction as a security control
What a Model Refuses, and Why

Every deployed model has something deciding what it will not answer. This course takes that layer apart: what is training, what is a filter, what a jailbreak actually exploits, and why instructions in a prompt are not a boundary.
8 lessons, written and corrected before you arrived. Reading them here needs no account. The first reads the whole way through; the others open and then stop, because a page nobody owns cannot tell who is reading it. Starting the course gives you your own copy, where every idea has problems standing under it and you can ask about any sentence.