ScienceBenchmarksDocsCatalogPricing
Sign inStart research
Back to catalog
Compute·general

Crawl Public Website

Map and crawl multiple related pages from one public website into a private, searchable Library snapshot with direct page URL and content-hash citations. Enforces public-host and same-site scope, robots.txt, conservative pacing, bounded size, large-job confirmation, cancellation, actual-page credit metering, and automatic chat continuation. Does not log in, bypass CAPTCHAs/paywalls, cross domains, rotate IPs, or evade access controls.

AvailableCortexa compute

Schema

JSON Schema the agent (or your API call) must match.

View JSON schemaExpandCollapse
JSON · 25 lines · 394 chars

Examples (1)

Crawl public documentation

input
JSON · 6 lines · 116 chars
Expected response keys: ok, crawlId, status

Identifiers

Catalog ID
crawl_website
Compute job
crawl_website
Added
2026-08-04 22:36Z
Tags
web, crawl, library, citations, auth
Cortexa.

The agent for research teams. 1.8K+ research tools across scientific and professional fields, with sources attached to the claims they support.

Product

  • Cortexa for science
  • Research benchmark
  • Documentation
  • Integrations
  • Tool catalog
  • Security & privacy
  • Pricing

Get started

  • Start research
  • Sign in
  • Developer API
  • MCP server

Support

  • Help center
  • Contact us
  • Terms of Service
  • Privacy Policy

© 2026 Cortexa. All rights reserved.

TermsPrivacy·For research context only · Not medical, legal, or financial advice.