Guardrails
Technical mechanisms and constraints built into AI systems to prevent harmful outputs or behaviors.
Definitions (2)
The set of ten proposed mandatory requirements applying across the AI lifecycle and supply chain (accountable official, lifecycle risk management, data governance and provenance, pre‑deployment and adversarial testing, transparent reporting and disclosures, meaningful human oversight, ongoing monitoring and incident reporting, supply‑chain transparency, record‑keeping/documentation, and conformity assessment) that developers and deployers must implement for high‑risk AI.
Operational safeguards including risk assessment frameworks, model validation and testing, supplier assurance, and security‑ and privacy‑by‑design controls intended to prevent or limit harms from AI systems in public service use.
Related Terms
AI Safety
Measures and engineering practices ensuring AI systems do not cause unreasonable harm to people, property, or the environment during intended use or foreseeable misuse....
Alignment
Ensuring an AI system’s objectives, outputs and behaviour are consistent with specified human goals, values, laws and expectations....
Red Teaming
A structured, adversarial testing process that simulates intentional misuse to find vulnerabilities, failure modes, and harms in AI systems before deployment....