All terms

Glossary

Guardrails

Technical boundaries around a language model that define what may go in, what may come out, and what it can trigger.

Guardrails sit in three places: in front of the model when checking input, behind it when checking output, and around the tools it may call.

The third is by far the most effective. What a model says can be corrected. What it triggers cannot.