Configure prompt injection attack protection
Activate or deactivate prompt injection attack detection settings to protect all generative AI interactions on your instance from malicious inputs and unintended model behaviors.
Before you begin
Role required: sn_generative_ai.nsa_admin
About this task
AI Guardian detects and logs prompt injection attempts across all generative AI applications and features on your instance. You can configure AI Guardian to block the AI-generated response when an attack is detected.
Prompt injection detection is enabled by default for all ServiceNow Otto skills, except Platform and custom skills, which can be configured manually. The default action is block and log, with a medium severity threshold. When a skill has its own setting, AI Guardian automatically applies the more protective of the two settings, the skill-level setting or the instance-level setting.
You can export logs for review. For more information, see Export ServiceNow Otto Guardian logs.
Procedure
Result
Prompt injection detection is configured on your instance for all generative AI workflows. AI Guardian detects prompt injection attempts based on the severity level you selected and responds according to the action you configured. When a skill has its own setting, the more protective setting applies automatically.