---
sourceDocument: Yokohama ServiceNow AI Platform Administration
sourceDocumentLink: https://www.servicenow.com/docs/r/yokohama/platform-administration

 Release :

    - yokohama

ft:locale :

    - en-US

ft:publication_title :

    - Yokohama ServiceNow AI Platform Administration

ft:clusterId :

    - platadm

bundleId :

    - platadm

workflow :

    - Platform


---

# Configure crawl settings

# Configure crawl settings for a HubSpot external content connector {#ariaid-title1}

* Release version: Yokohama
* 
* Updated October 9, 2025
* 
* ![](https://www.servicenow.com/docs/portal-asset/ico-clock) 3 minutes to read

Specify the tickets you want your HubSpot external content connector to crawl. Define inclusion or exclusion filters to dictate the types of content the crawl retrieves and feeds to AI Search for indexing.

## Before you begin

A connector administrator must have already created the HubSpot external content connector that you want to configure crawl settings for. To learn about this procedure, see [Create a HubSpot external content connector](https://www.servicenow.com/docs/ApLcysFNR0G_QIpB~wf31g "Create an external content connector to retrieve searchable content and security principals from your HubSpot source system.").

Role required: sn_ext_conn.xcc_admin

## About this task

This task is optional. By default, the HubSpot external content connector crawls all tickets from its specified source system and sends all content, and binary attachment files with all supported file extensions, to AI Search for indexing. Only perform this task if you want the connector to use any of the following non-default settings:

* Inclusion or exclusion filters for the tickets to crawl when running content crawls
* Inclusion or exclusion filters for the binary attachment file extensions to retrieve when running content crawls
{#configure-crawl-settings-hubspot-external-content-connector__ul_onp_x1f_tdc}

Content is only retrieved from the source system if it passes all of your configured crawl setting filters. If any crawl setting filter excludes a
content item, the external content connector doesn't retrieve it.  
Important:  
By default, an external content connector can index up to one million (1,000,000) documents from its source system. When a connector exceeds this limit, it continues to crawl the source system, but
only sends document deletions and updates to AI Search for indexing, ignoring new documents. The connector logs an error message for every 10,000 documents it crawls beyond the indexing limit.{#configure-crawl-settings-hubspot-external-content-connector__store-app-external-content-connectors-indexing-limit-numeric-ph}

When a connector's indexed document count exceeds 800,000, a warning message appears in the connector's UI to indicate that it's approaching the indexing
limit. If the connector reaches the indexing limit, an error message appears in its UI.

If one of your connectors reaches the indexing limit, you can update its crawl settings and file inclusion/exclusion filters to reduce the number of documents it retrieves. Alternately, if you need to index
more than 1,000,000 documents, you can create a Customer Service and Support case at <https://support.servicenow.com/now> to request a limit increase for the connector.

## Procedure

1. Navigate to AllExternal Content ConnectorsExternal Content Admin Home. {#configure-crawl-settings-hubspot-external-content-connector__navigate-ext-content-connectors-admin-home-step}
{#configure-crawl-settings-hubspot-external-content-connector__navigate-ext-content-connectors-admin-home-step}
2. In the Connectors list, select the record for the HubSpot external content connector whose settings you want to modify.
3. In the connector editor's Settings tab, select Crawl settings. {#configure-crawl-settings-hubspot-external-content-connector__select-crawl-settings-step}
{#configure-crawl-settings-hubspot-external-content-connector__select-crawl-settings-step}
4. Select one of the following Tickets options:
   * To crawl all tickets from the source system, select Crawl all tickets.
   * To crawl only tickets that have been modified in the last three months, select Modified last 3 months.

   * To crawl only tickets that have been modified in the last six months, select Modified last 6 months.

   * To crawl only tickets that have been modified in the last year, select Modified last year.

   {#configure-crawl-settings-hubspot-external-content-connector__choices_cdr_s4x_sdc}
5. Select one of the following Attachments options:
   * To retrieve all attachments with supported file extensions from the source system, select Crawl all attachments.
   * To retrieve only attachments with specified file extensions from the source system, select Include only these file extensions, then use the File extensions to include field
     to enter attachment file extensions you want the connector to include when crawling.

     As an example, you might enter <kbd class="ph userinput">.docx</kbd> to retrieve only attachments with the Microsoft Word file format.
   * To retrieve all attachments except those with specified file extensions from the source system, select Exclude only these file extensions, then use the File extensions to exclude field to enter attachment file extensions you want the connector to exclude when crawling.

     As an example, you might enter <kbd class="ph userinput">.csv</kbd> to exclude attachments with the Comma-Separated Values (CSV) file format.

   {#configure-crawl-settings-hubspot-external-content-connector__choices_i3b_wt5_wgc}  
   For details on the supported attachment file extensions, see [Binary file extensions supported in External Content Connectors](https://www.servicenow.com/docs/7Xl2aL3AUF0LNjpE7xKH1Q "Connector administrators can restrict the binary file types an external content connector retrieves by specifying file extensions in inclusion or exclusion filters.").
6. Select Save and validate.
{#configure-crawl-settings-hubspot-external-content-connector__steps_yqk_jjw_22c}

## Result

The HubSpot external content connector is updated with your modified crawl settings.

## What to do next

To retrieve content from your HubSpot source system using your modified crawl settings, create and run a one-time content crawl for your HubSpot external content connector. To learn about creating and running one-time content
crawls, see [Create a content crawl for an external content connector](https://www.servicenow.com/docs/cHKxITeSqLvmLnYo9QVnHg "Retrieve searchable content and metadata from your source system with a content crawl. Run the crawl as a one-time task or schedule it to run on a recurring basis.").

*[\>]: and then


