← Back to the AI Glossary 📖 Safety & Alignment

Guardrails

Rules and filters built around an AI system to prevent it from producing harmful, off-topic, or unwanted output.

Guardrails can be built into the model itself during training, or added as a separate check on top of its output before a user ever sees it.
One of 60 free AI glossary terms
Plain-language definitions for the AI jargon you'll actually run into — no email needed, ever, for this section.
Browse the Full Glossary →