---
sourceDocument: Yokohama Enable AI
sourceDocumentLink: https://servicenow-prod.fluidtopics.net/r/yokohama/intelligent-experiences

 Release :

    - yokohama

ft:locale :

    - en-US

ft:publication_title :

    - Yokohama Enable AI

ft:clusterId :

    - platai

bundleId :

    - platai

workflow :

    - Platform


---

# Activate offensiveness protection for generative AI

# Activate offensiveness protection for generative AI {#ariaid-title1}

Release version: Yokohama  
Updated July 31, 2025  
![](https://www.servicenow.com/docs/portal-asset/ico-clock) 1 minute to read  
Activate
offensiveness detection to log or block offensive content generated by Now Assist skills and workflows.

## Before you begin

Role required: sn_generative_ai.nsa_admin

## About this task

Generative AI output is probabilistic, which means that the same input can produce different outputs. Some of the AI generated content may be offensive, which includes toxic, sexist, or other harmful language. AI Guardian enables you to detect offensive content in both inputs and outputs, and logs the event when it is detected. You can also configure it to block offensive material so that users see a standard error message instead of the generated response.  
Note:  
Offensiveness detection applies only to specific Now Assist skills and workflows. It is not available for all Now Assist applications. For more information about the list of skills that support offensiveness detection, see [Now Assist Guardian](https://servicenow-prod.fluidtopics.net/H0bVbM_NdghPjzLAF~OnJg#now-assist-guardian "AI Guardian is built on the ServiceNow Small Language Model (SLM) and monitors generative AI interactions to detect offensive content, prompt injection attacks, and sensitive topics.").

You can export logs for
review. For more
information, see [Export Now Assist Guardian logs](https://servicenow-prod.fluidtopics.net/n_5MubOPi1HhhBTXleT4EQ "Export logs from AI Guardian to get insights into how often different guardrails are being detected and used.").

## Procedure

1. Navigate to AllNow Assist AdminSettings.
2. In the side panel, select the Now Assist GuardianOffensiveness tab.
3. Go to the Available for you tab to see which workflows you can choose from.  
   Offensiveness guardrails
   that are already activated
   appear in the Active tab.
4. Select Activate for the workflow that you want to enable offensiveness detection.
5. Select your impact detection.
   * Select Log only to record the events when offensive content is detected. The content is still shown to the user.
   * Select Block and log to record the event and prevents the content from being shown to the user. The user sees a standard error message instead.

   {#activate-offensiveness-protection-for-generative-ai__choices_zdn_xrr_q3c}  

6. Select Save.

## Result

Offensiveness
detection guardrail is enabled on your instance for
the selected
workflow. Events are logged when offensive
content is detected or generated.

## What to do next

You can enable offensiveness
detection for
separately for each supported Now Assist application and workflow.
Repeat this task for each workflow on which you want offensiveness protection
enabled.

To
change the detection impact for an active workflow, select more options (![More options icon.]()) icon in the list of active workflows and
then select
Edit.

To deactivate offensiveness protection
for a
workflow,
select more
options (![More options icon.]()) icon in the list of active workflows and then select Deactivate
.

*[\>]: and then


