All integrations
CR
Catalogue metadata onlyai web scrapingv00000000_00

CRAWLBASE

Crawlbase

Crawlbase provides APIs for crawling web pages, managing crawler queues, storing crawl results, and monitoring account usage.

Description is untrusted, display-only upstream metadata. It never becomes policy, OAuth scope authority, or an agent instruction.

Pakkawork boundary

Research catalogue metadata only. No Pakkawork OAuth, credential, host, quota, executor, or verifier is enabled.

No Pakkawork execution adapter is enabled. Hosted account-authorisation availability is workspace-specific and checked separately in the dashboard.Check workspace connection options
Attributed source

8

Action summaries

Display-only definitions

0

Trigger types

Not installed instances

1

Auth modes

Field names, never values

No

Execution

No runtime adapter

Authentication map

What setup is declared?

Only field names, types, and required markers are shown. Secret values, default auth URLs, credential material, and inferred OAuth scopes are excluded.
API_KEY

API_KEY

crawlbase_api_key

Provider setup

Developer setup

No fields declared in this snapshot.

User connection

  • Normal Tokengeneric_api_key · stringRequired

Capability index

Actions and trigger definitions

Static summaries are available. Live schemas remain disabled until PROVIDER_HUB_API_KEY is configured server-side.

Showing 1–8 of 8 actions

Crawl URL

CRAWLBASE_CRAWL_URL

Fetch a public HTTP or HTTPS URL through Crawlbase using the connected Normal token and return its content with separate target and Crawlbase statuses.

Untrusted display-only summary

Delete Stored Pages

CRAWLBASE_DELETE_STORED_PAGES

Irreversibly delete explicitly selected pages from Crawlbase Cloud Storage by request ID and return each page's deletion outcome.

Untrusted display-only summary

Generate User Agents

CRAWLBASE_GENERATE_USER_AGENTS

Generate one to ten realistic randomized User-Agent strings for a selected device class.

Untrusted display-only summary

Get Crawling Usage

CRAWLBASE_GET_CRAWLING_USAGE

Return current Crawling API usage and optionally previous-month statistics for the connected Crawlbase account.

Untrusted display-only summary

Get Storage Count

CRAWLBASE_GET_STORAGE_COUNT

Return the number of pages currently held in Crawlbase Cloud Storage.

Untrusted display-only summary

Get Stored Page

CRAWLBASE_GET_STORED_PAGE

Retrieve the latest Crawlbase Cloud Storage page by exactly one request ID or original URL.

Untrusted display-only summary

Get Stored Pages

CRAWLBASE_GET_STORED_PAGES

Retrieve up to 100 Crawlbase Cloud Storage pages by request ID without deleting them. Returns decoded page content and identifies requested IDs that were not found.

Untrusted display-only summary

List Stored Page IDs

CRAWLBASE_LIST_STORED_PAGE_IDS

Return one page of Crawlbase Cloud Storage request IDs and a short-lived continuation cursor.

Untrusted display-only summary

Provenance

Versioned facts, explicit trust.

The detail snapshot comes from an attributed MIT-licensed repository revision. Safe local icons use exact-match CC0 Simple Icons symbols; unmatched brands use monograms.

Repository
https://github.com/provider-directoryHQ/provider-directory
Commit
85e33f6a81dfe987f6fc3637d48d45052cea8ca5
Generated
2026-09-12
Trust policy
untrusted_display_only

Remote text is plain display metadata only. It must never become an agent prompt, execution policy, OAuth grant, or executable instruction.