Guardrails
The Guardrails tab is a safety and control interface designed to restrict and guide a character's AI behavior. It allows creators to define the ethical and topical boundaries of the conversation.
This API is available only on the Paid Plans.
Introduction
The Guardrails tab provides a robust control interface to define your character’s ethical and topical boundaries, preventing off-topic or inappropriate responses. By entering specific constraints in the Instructions field and toggling custom guardrails on, you ensure your AI character responds in a safe and professional manner, and is strictly aligned with your intended brand voice.

Main Features and Sections
1. Custom Guardrails Toggle
This toggle acts as the master switch to activate or deactivate your character's safety oversight. When enabled, it forces the AI character to prioritize your specific behavioral rules over its general conversational logic.
2. Guardrails Instruction
This text field allows you to define explicit safety rules and response boundaries using up to 7,500 words. It is where you input constraints like "Do not talk about politics" to ensure the character remains professional and on-topic.
Examples
This example shows how you can implement guardrails for your AI character in Convai Playground
Conclusion
The Guardrails tab is an essential safety tool that empowers creators to enforce strict behavioral boundaries and content restrictions on their AI characters. By combining custom instructions with real-time testing, it ensures that interactions remain secure, on-brand, and free from undesirable topics or hallucinations.
Last updated
Was this helpful?