Skip to content

What is…

Guardrails

Rules and checks around an AI system that block unsafe, off-topic or wrong outputs and actions.

Guardrails include input filters, output validators, permission limits (read-only access, spending caps), approval steps before risky actions, and sandboxes that keep agents from touching things they shouldn't.

For agents, the most important guardrail is simple: require a human's OK before anything irreversible — sending, deleting, paying, publishing.

🧠 Test yourself

Which of these describes Guardrails?

Related terms

🎮 Learn AI by playing

40 bite-size missions, boss battles and a certificate. Free.

Start the bootcamp →

☀️ One AI term every morning

Plus the day's top stories, in your inbox by 8am.