Create crawl job

Create a crawl job to crawl the given URL with specified depth and page limits.

Header

xi-api-keystringOptional

Anfrage

This endpoint expects an object.
urlstringErforderlich
URL to a page of documentation that the agent will have access to in order to interact with users.
max_pagesintegerOptional1-10000Standardwert ist 1000

Maximum number of pages to crawl (1-10,000), defaults to 1000.

patternstringOptional<=2048 characters
If set, only URLs that match this pattern are included.
sitemap_urlslist of stringsOptional

List of URLs to crawl from sitemap (optional, overrides automatic URL discovery).

parent_folder_idstringOptional
If set, the created document or folder will be placed inside the given folder.
enable_auto_syncbooleanOptionalStandardwert ist false

Whether to enable auto-sync for this URL document.

auto_removebooleanOptionalStandardwert ist false

Whether to automatically remove the document if the URL becomes unavailable. Only applicable when auto-sync is enabled.

auto_discoverbooleanOptionalStandardwert ist false

Automatically discover and add new pages linked from already-crawled pages during auto-sync. Requires enable_auto_sync=true.

minimum_frequency_daysintegerOptional1-180

Minimum frequency (in days) at which the underlying eligible documents are refreshed. The actual interval may be shorter, never longer. Defaults to 7, tightened to the parent folder’s frequency if that is stricter. Only applicable when auto-sync is enabled.

max_depthintegerOptionalStandardwert ist 3Deprecated

Deprecated - this field is a no-op and will be removed in a future version.

Antwort

Successful Response
idstring
typeenumStandardwert ist discovery
Erlaubte Werte:
root_folder_idstring
statusstring
created_atinteger
folder_pathlist of objectsOptional
The folder path segments leading to the root folder, from root to parent folder.

Fehler

422
Crawl Jobs Create Request Unprocessable Entity Error