As AI systems gain increased autonomy and influence, ensuring their safe and beneficial operation becomes paramount. AI safety is not a single property but a collection of principles, frameworks, and practices designed to reduce risk while preserving model capability. This article surveys the foundational principles that underpin modern AI safety practice.
From alignment to governance, understanding these principles is essential for researchers, developers, and policymakers working at the frontier of autonomous systems.