indexAI Safety & Security#ai-safety#security#risk

AI Safety and Security

AI safety and security is the discipline of making AI systems resistant to misuse, accidents, data exposure, and model-mediated attacks. The core move is to treat the model as an unreliable component inside a security boundary, not as the boundary.

Mental model

An AI application crosses trust boundaries whenever untrusted data can influence model output and that output can reach data, code, money, or people. Security therefore constrains authority and validates effects outside the model; a prompt is never the sole enforcement layer.

Roadmap: threat landscape to assurance

Data and action risk

Controls and assurance

Connects to: Autonomy and Control · Evaluation · AI Ethics and Governance

Core sources