What is…
Guardrails
Rules and checks around an AI system that block unsafe, off-topic or wrong outputs and actions.
Guardrails include input filters, output validators, permission limits (read-only access, spending caps), approval steps before risky actions, and sandboxes that keep agents from touching things they shouldn't.
For agents, the most important guardrail is simple: require a human's OK before anything irreversible — sending, deleting, paying, publishing.
🧠 Test yourself
Which of these describes Guardrails?
Related terms
🎮 Learn AI by playing
40 bite-size missions, boss battles and a certificate. Free.
Start the bootcamp →☀️ One AI term every morning
Plus the day's top stories, in your inbox by 8am.