# AI guardrails

> The checks around a model that constrain what it can say and do — input filters, output validation, permissioned tools — so a bad response fails safely.

A prompt instruction is a request, not a boundary: a determined user can talk a model past 'never reveal this'. Guardrails are enforcement instead — code that blocks the prompt-injection pattern before it reaches the model, validates that the answer came from the allowed sources, and gates tool calls like 'send email' or 'refund' behind real permissions.

The honest model of a language model is a fast, confident, occasionally wrong employee: you would not give that employee the database password and no review. Guardrails are the review, implemented where the model cannot talk its way around it.

## Related terms

- https://dfieldsolutions.com/en/glossary/prompt-injection.md
- https://dfieldsolutions.com/en/glossary/tool-calling.md
- https://dfieldsolutions.com/en/glossary/llm-evaluation.md

---

Source: https://dfieldsolutions.com/en/glossary/ai-guardrails
DField Solutions — Dunakeszi, Hungary — dezso@dfieldsolutions.com
Booking: see https://dfieldsolutions.com/en/contact
