Skip to main content
Monitor flagged conversations and evaluate how your agent handles harmful content. Safety dashboard

Metrics

Configuring filters

Safety filters are configured project-wide in Behavior and overridden per channel in Voice and Messaging. See Safety filters for the full reference – categories, severity levels, language support, and how filters fit with Guardrails.

Safety filters

Configure per-channel filter categories and severity levels.

Standard dashboard

Day-to-day performance monitoring: containment, call volume, and duration.

Conversation review

Inspect individual flagged conversations in full transcript view.
Last modified on July 10, 2026