Setting up Guardrails

Hard rules your AI Avatar will not break. When to use them vs System Prompt boundaries, how to write them, and what guardrails cannot do.

HomeHelp CenterAI Avatars › Setting up Guardrails

Guardrails are the hard rules your avatar will not break. Where Objectives are soft goals ("make sure we get there"), Guardrails are firm limits ("never go there"). Use them for compliance language, brand-safety rules, and anything where being wrong has consequences.

Examples of guardrails

Guardrails vs System Prompt boundaries

You can write "never give medical advice" into the System Prompt's Boundaries section. So why use a Guardrail?

Rule of thumb: tone and personality go in the System Prompt. Rules with consequences go in Guardrails.

Verbal vs visual modality

How to write a good guardrail

One rule per guardrail. "Never quote unverified pricing and never discuss competitors" should be two guardrails, not one. Say what to do instead. "Don't give medical advice → say 'I'd recommend asking your doctor' instead." Pure prohibitions confuse the avatar; redirects are easier to follow. Test on edge cases. Ask the avatar the exact question your guardrail is meant to prevent. If it answers it anyway, the rule isn't strict enough. Each persona can hold up to 10 guardrails. Fewer is better — every guardrail is one more thing the avatar checks against every reply. 3–5 well-chosen rules beat 10 vague ones.

What guardrails are not

If you have an existing list of "things the bot must never say" from your legal or compliance team, that's literally what guardrails are for. Paste each line as a separate guardrail.

Related articles in AI Avatars

Help Center · Learn · Privacy Policy · Terms of Service · Contact