Phase 18: Ethics, Safety & Alignment

Frontier Safety Frameworks — RSP, PF, FSF

Three major-lab frameworks define the 2026 industry governance of frontier capability. Anthropic Responsible Scaling Policy v3.0 (February 2026) introduces tiered AI Safety Levels (ASL-1 through ASL-5+), modeled on biosafety levels, with ASL-3 activated May 2025 for CBRN-relevant models. OpenAI Preparedness Framework v2 (April 2025) defines five criteria for tracked capabilities and separates Capabilities Reports from Safeguards Reports. DeepMind Frontier Safety Framework v3.0 (September 2025) introduces Critical Capability Levels including a new Harmful Manipulation CCL. All three now include competitor-adjustment clauses allowing deferral if peer labs ship without comparable safeguards. Cross-lab alignment remains structural, not terminological: "Capability Thresholds," "High Capability thresholds," and "Critical Capability Levels" denote analogous constructs. Describe Anthropic's ASL tier structure and what activated ASL-3. Name the five OpenAI Preparedness Framework v2 criteria for tracked capabilities. Describe DeepMind's Critical Capability Level structure and the Harmful Manipulation CCL. Explain the competitor-adjustment clauses and why they matter for race dynamics. Define a safety case and describe the three-pillar structure (monitoring, illegibility, incapability). Lessons 7-17 establish that deception is possible, dual-use capability exists, and evaluation has limits. A lab with a frontier-capable model needs an internal governance structure that: Defines thresholds for when new safeguards are required. Defines required evaluations before scaling. Describes what a safety case looks like. Handles the race-dynamic problem (if competitors ship without safeguards, what do you do?).…

Frontier Safety Frameworks — RSP, PF, FSF: Three major-lab frameworks define the 2026 industry governance of frontier capability. Anthropic Responsible…

This free lesson is part of the AI Engineering from Scratch curriculum. Read the full explanation, run the lesson code, and verify the result in the interactive reader or from the repository source.

Browse the complete course catalog or open this lesson on GitHub.