Back to catalog
Compute·general
Crawl Public Website
Map and crawl multiple related pages from one public website into a private, searchable Library snapshot with direct page URL and content-hash citations. Enforces public-host and same-site scope, robots.txt, conservative pacing, bounded size, large-job confirmation, cancellation, actual-page credit metering, and automatic chat continuation. Does not log in, bypass CAPTCHAs/paywalls, cross domains, rotate IPs, or evade access controls.
AvailableCortexa compute
Schema
JSON Schema the agent (or your API call) must match.
View JSON schemaExpandCollapse
JSON · 25 lines · 394 chars
Examples (1)
Crawl public documentation
input
JSON · 6 lines · 116 chars
Expected response keys: ok, crawlId, status
Identifiers
- Catalog ID
- crawl_website
- Compute job
- crawl_website
- Added
- 2026-08-04 22:36Z
- Tags
- web, crawl, library, citations, auth