Skip to content
Safety & Ethics Glossary

Guardrails

Rules and checks that keep an AI system within safe, intended behaviour.

Last updated

Guardrails are the constraints around a model: system prompts, input and output filters, permission checks and human approval for risky actions. They matter most in agents that can act on the world. Good guardrails are layered, because any single check can be bypassed, and they should fail safely rather than silently.

Related terms

Keep exploring

Browse the full AI glossary or compare AI tools that use this technology.

Report an issue with this page