All integrations
WE
Catalogue metadata onlyai web scrapingv20260615_00

WEBSCRAPING_AI

WebScraping.AI

WebScraping.AI provides an API for web scraping with features like Chrome JS rendering, rotating proxies, and HTML parsing.

Description is untrusted, display-only upstream metadata. It never becomes policy, OAuth scope authority, or an agent instruction.

Pakkawork boundary

Research catalogue metadata only. No Pakkawork OAuth, credential, host, quota, executor, or verifier is enabled.

No Pakkawork execution adapter is enabled. Hosted account-authorisation availability is workspace-specific and checked separately in the dashboard.Check workspace connection options
Attributed source

7

Action summaries

Display-only definitions

0

Trigger types

Not installed instances

1

Auth modes

Field names, never values

No

Execution

No runtime adapter

Authentication map

What setup is declared?

Only field names, types, and required markers are shown. Secret values, default auth URLs, credential material, and inferred OAuth scopes are excluded.
API_KEY

API_KEY

webscraping_ai_api_key

Provider setup

Developer setup

No fields declared in this snapshot.

User connection

  • WebScraping.AI API Keygeneric_api_key · stringRequired

Capability index

Actions and trigger definitions

Static summaries are available. Live schemas remain disabled until PROVIDER_HUB_API_KEY is configured server-side.

Showing 1–7 of 7 actions

Get account usage and quota

WEBSCRAPING_AI_ACCOUNT_INFO

Tool to retrieve account API call quota and usage. Use when checking remaining requests and subscription details.

Untrusted display-only summary

Ask Question About Web Page

WEBSCRAPING_AI_ASK_QUESTION

Tool to get an answer to a question about a given web page using LLM. Use when you need AI-powered analysis or extraction from a web page. Proxies and Chromium JavaScript rendering are used for page retrieval.

Untrusted display-only summary

Extract Fields with AI

WEBSCRAPING_AI_EXTRACT_FIELDS

Tool to extract structured data fields from a web page using AI. Returns extracted fields as JSON. Uses proxies and Chromium JavaScript rendering for page retrieval and processing.

Untrusted display-only summary

Get Rendered HTML

WEBSCRAPING_AI_GET_RENDERED_HTML

Tool to retrieve fully rendered HTML of a webpage. Use when JS-generated content must be included.

Untrusted display-only summary

Get Selected HTML

WEBSCRAPING_AI_GET_SELECTED_HTML

Tool to extract HTML from specific page elements using CSS selectors. Use when you need HTML from a particular section rather than the entire page.

Untrusted display-only summary

Get Selected Multiple Elements

WEBSCRAPING_AI_GET_SELECTED_MULTIPLE

Tool to extract HTML of multiple page areas by URL and CSS selectors. Use when you need to extract multiple elements without HTML parsing on your side.

Untrusted display-only summary

Get Text

WEBSCRAPING_AI_GET_TEXT

Tool to retrieve raw text content from a specified web page. Returns unstructured plain text — markdown formatting (code fences, lists, headings) is not preserved. Use when you need plain text extraction from a URL. Use FIRECRAWL_EXTRACT instead when formatted markdown output is required.

Untrusted display-only summary

Provenance

Versioned facts, explicit trust.

The detail snapshot comes from an attributed MIT-licensed repository revision. Safe local icons use exact-match CC0 Simple Icons symbols; unmatched brands use monograms.

Repository
https://github.com/provider-directoryHQ/provider-directory
Commit
85e33f6a81dfe987f6fc3637d48d45052cea8ca5
Generated
2026-09-12
Trust policy
untrusted_display_only

Remote text is plain display metadata only. It must never become an agent prompt, execution policy, OAuth grant, or executable instruction.