Skip to content
Scalekit Docs
Talk to an Engineer Dashboard

Connect AI agents to Algolia Crawler

Scalekit connector
Open markdown

The Algolia Crawler connector lets your AI agent act in each user's Algolia Crawler account. Each user connects their own Algolia Crawler username and password once, and Scalekit sends it with every call, so your agent never handles credentials. It comes with 20 tools.

Tools
20
What they doRead · write · destructive
9 · 9 · 29 read9 write2 destructive
Users sign in with

Setup

  1. Install the SDK

    Terminal window
    npm install @scalekit-sdk/node dotenv
  2. Set your credentials

    Add your Scalekit credentials to your .env file. Find values in app.scalekit.com > Developers > API Credentials.

    .env
    SCALEKIT_ENVIRONMENT_URL=<your-environment-url>
    SCALEKIT_CLIENT_ID=<your-client-id>
    SCALEKIT_CLIENT_SECRET=<your-client-secret>
  3. Create the Algolia Crawler connection

    In AgentKit > Connections, create an Algolia Crawler connection. The name you give it is the connection_name your code passes. See Configure connections.

Tools

Pass the exact name to execute_tool
Try in PlaygroundRequest a tool
  • algoliacrawler_get_config_versionRetrieve one saved version of a crawler's configuration.Read-only

    Get Config Version

    Retrieve one saved version of a crawler's configuration. Returns the version number, creation time, author id, and the full configuration object for that version (start URLs, actions, index prefix, schedule, rate limit, and other crawler settings). Use this to inspect or restore an earlier configuration. Use list_config_versions to find version numbers and get_crawler for the current configuration. Requires a crawler id from list_crawlers and a version number from list_config_versions.

    Inputs

    idstringrequired
    Unique ID (UUID) of the crawler. Use list_crawlers to find it. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
    versionintegerrequired
    Version number of the crawler configuration to retrieve. Version 1 is the initial configuration used when the crawler was created. Use list_config_versions to find valid numbers. Example: 3.
  • algoliacrawler_get_crawl_run_fileRetrieve the downloadable log file for a single crawl run of a crawler.Read-only

    Get Crawl Run File

    Retrieve the downloadable log file for a single crawl run of a crawler. Returns a JSON object with a file string holding the log file content for that run. Use this to debug or audit one crawl run. Use list_crawl_runs first to find the log id, and get_url_stats for aggregate counts instead of per-run logs. Requires a crawler id from list_crawlers and a log id from list_crawl_runs.

    Inputs

    idstringrequired
    Unique ID (UUID) of the crawler. Use list_crawlers to find it. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
    logIdstringrequired
    Unique ID (UUID) of the crawler log (crawl run) to download. Use list_crawl_runs to find it. Example: a2ebb507-ef64-4b6b-9d84-ef66baaa7a80.
  • algoliacrawler_get_crawlerGet the details of one Algolia crawler by id, optionally including its full configuration.Read-only

    Get Crawler

    Get the details of one Algolia crawler by id, optionally including its full configuration. Returns the crawler name, created and updated timestamps, whether it is running, reindexing, or blocked (with the blocking error and blocking task id when blocked), and the last reindex start and end times; with the configuration option set, also returns the configuration. Use this for the state or configuration of a known crawler. Use list_crawlers to find ids and get_task_status to follow an asynchronous task. Requires a crawler id from list_crawlers or create_crawler.

    Inputs

    idstringrequired
    Crawler ID (UUID) of the crawler to retrieve. Get it from the list crawlers tool or from the response of the create crawler tool. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
    withConfigboolean
    Set to true to include the crawler's full configuration in the response. Leave empty or false to return only status information such as name, timestamps, and running state. Example: true.
  • algoliacrawler_get_task_statusRetrieve the status of a crawler task, showing whether it is still pending or has completed.Read-only

    Get Task Status

    Retrieve the status of a crawler task, showing whether it is still pending or has completed. Returns an object with a pending boolean. Use this to poll a task started by another crawler tool, such as a run or reindex, until pending is false. Use cancel_task to unblock a crawler whose task failed. Requires a crawler id from list_crawlers and a task id returned by the tool that started the task.

    Inputs

    idstringrequired
    Unique ID (UUID) of the crawler. Use list_crawlers to find it. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
    taskIDstringrequired
    Unique ID (UUID) of the task to check. It is returned when a crawler action such as run, reindex, or crawl URLs is started. Example: 98458796-b7bb-4703-8b1b-785c1080b110.
  • algoliacrawler_get_url_statsRetrieve crawl statistics for a crawler, broken down by URL status.Read-only

    Get URL Stats

    Retrieve crawl statistics for a crawler, broken down by URL status. Returns the total count of crawled URLs and a data array where each entry has a status (DONE, SKIPPED, or FAILED), a category (fetch, extraction, indexing, or success), a reason, a readable explanation, and the number of URLs with that status. Use this to see why URLs were skipped or failed. Use list_crawl_runs for per-run history and get_crawler for configuration. Requires a crawler id from list_crawlers.

    Inputs

    idstringrequired
    Unique ID (UUID) of the crawler. Use list_crawlers to find it. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
  • algoliacrawler_list_config_versionsList the saved configuration versions of a crawler, including who authored each change.Read-only

    List Config Versions

    List the saved configuration versions of a crawler, including who authored each change. Returns a paginated response with the current page, items per page, total count, and an items array of version number, creation time, and author id. Every configuration update adds a new version. Use this to find a version number before calling get_config_version. The list does not include the configuration itself. Requires a crawler id from list_crawlers.

    Inputs

    idstringrequired
    Unique ID (UUID) of the crawler. Use list_crawlers to find it. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
    itemsPerPageinteger
    Number of configuration versions to return per page, from 1 to 100. Defaults to 20. Use together with page to walk through large result sets. Example: 20.default 20
    pageinteger
    Page of results to retrieve, from 1 to 100. Defaults to 1. Compare the total in the response with itemsPerPage to decide whether more pages exist. Example: 1.default 1
  • algoliacrawler_list_crawl_runsList the recorded crawl runs (crawler logs) for a crawler, optionally filtered by date range or URL status.Read-only

    List Crawl Runs

    List the recorded crawl runs (crawler logs) for a crawler, optionally filtered by date range or URL status. Returns a logs array and a meta object with the total number of matching records. Each log has its id, config id, reindex id, start and completion times, file sizes, expiry, status, access count, and counts of done, skipped, and failed URLs. Use this to find a log id before calling get_crawl_run_file or delete_crawl_runs. Use get_url_stats for aggregate URL status counts across the crawler. Requires a crawler id from list_crawlers.

    Inputs

    idstringrequired
    Unique ID (UUID) of the crawler. Use list_crawlers to find it. Example: e0f6db8a-24f5-4092-83a4-1b2c6cb6d809.
    fromstring
    Only return runs on or after this date, as a Unix timestamp in seconds, passed as a string. Leave empty for no lower bound. Example: 1762264044.
    limitinteger
    Maximum number of runs to return, from 1 to 1000. Defaults to 10. Use together with offset to page through results. Example: 10.default 10
    offsetinteger
    Number of runs to skip before returning results, used for pagination together with limit. Leave empty to start from the first run. Example: 10.
    orderstring
    Sort direction of the results, either ASC or DESC. Leave empty to use the server default order. Example: DESC.one of ASCDESC
    statusstring
    Filter runs by crawled URL status. Must be one of DONE, SKIPPED, or FAILED. Leave empty to return runs regardless of status. Example: FAILED.one of DONESKIPPEDFAILED
    untilstring
    Only return runs on or before this date, as a Unix timestamp in seconds, passed as a string. Leave empty for no upper bound. Example: 1762264044.
  • algoliacrawler_list_crawlersList the Algolia crawlers on the account, optionally filtered by crawler name or Algolia application ID.Read-only

    List Crawlers

    List the Algolia crawlers on the account, optionally filtered by crawler name or Algolia application ID. Returns a paginated response with the current page, items per page, total count, and an items array of crawler ids and names. Use this to find a crawler id before calling a tool that needs one. The list only carries id and name, so use the get-crawler tool for configuration details.

    Inputs

    itemsPerPageinteger
    Number of crawlers to return per page, from 1 to 100. Defaults to 20. Use together with page to walk through large result sets.default 20
    namestring
    Crawler name used to filter the response, up to 64 characters. Leave empty to list crawlers regardless of name. Example: test-crawler.
    pageinteger
    Page of results to retrieve, from 1 to 100. Defaults to 1. Compare the total in the response with itemsPerPage to decide whether more pages exist.default 1