AptlyStar

fastCRW

Scrape, search, crawl, and map web data

Instructions d'utilisation

Integrate fastCRW into the workflow. Scrape pages, search the web, crawl entire sites, and map URL structures. fastCRW is a Firecrawl-compatible web scraper in a single binary — self-host or cloud.

Outils

crw_scrape

Extract structured content from web pages with comprehensive metadata support. Converts content to markdown or HTML while capturing SEO metadata, Open Graph tags, and page information.

Entrée

ParamètreTypeObligatoireDescription
urlstringOuiThe URL to scrape content from (e.g., "https://example.com/page"\)
scrapeOptionsjsonNonOptions for content scraping
baseUrlstringNonBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringOuifastCRW API key
pricingper_requestNonNo description
rateLimitstringNonNo description

Sortie

ParamètreTypeDescription
markdownstringPage content in markdown format
htmlstringRaw HTML content of the page
metadataobjectPage metadata including SEO and Open Graph information
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code (e.g., "en")
↳ sourceURLstringOriginal source URL that was scraped
↳ statusCodenumberHTTP status code of the response
↳ keywordsstringPage meta keywords
↳ robotsstringRobots meta directive (e.g., "follow, index")
↳ ogTitlestringOpen Graph title
↳ ogDescriptionstringOpen Graph description
↳ ogUrlstringOpen Graph URL
↳ ogImagestringOpen Graph image URL
↳ ogLocaleAlternatearrayAlternate locale versions for Open Graph
↳ ogSiteNamestringOpen Graph site name
↳ errorstringError message if scrape failed

Search for information on the web using fastCRW

Entrée

ParamètreTypeObligatoireDescription
querystringOuiThe search query to use
baseUrlstringNonBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringOuifastCRW API key
pricingper_requestNonNo description
rateLimitstringNonNo description

Sortie

ParamètreTypeDescription
dataarraySearch results data with scraped content and metadata
↳ titlestringSearch result title from search engine
↳ descriptionstringSearch result description/snippet from search engine
↳ urlstringURL of the search result
↳ markdownstringPage content in markdown (when sources include scraped content)
↳ metadataobjectMetadata about the search result page
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ errorstringError message if scrape failed

crw_crawl

Crawl entire websites and extract structured content from all accessible pages

Entrée

ParamètreTypeObligatoireDescription
urlstringOuiThe website URL to crawl (e.g., "https://example.com" or "https://docs.example.com/guide"\)
maxPagesnumberNonMaximum number of pages to crawl (e.g., 50, 100, 500). Default: 100
maxDepthnumberNonMaximum depth to crawl from the starting URL (e.g., 1, 2, 3). Controls how many levels deep to follow links
formatsjsonNonOutput formats for scraped content (e.g., ["markdown"], ["markdown", "html"], ["markdown", "links"])
excludePathsjsonNonURL paths to exclude from crawling (e.g., ["/blog/", "/admin/", "/*.pdf"])
includePathsjsonNonURL paths to include in crawling (e.g., ["/docs/", "/api/"]). Only these paths will be crawled
onlyMainContentbooleanNonExtract only main content from pages
baseUrlstringNonBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringOuifastCRW API key
pricingper_requestNonNo description
rateLimitstringNonNo description

Sortie

ParamètreTypeDescription
pagesarrayArray of crawled pages with their content and metadata
↳ markdownstringPage content in markdown format
↳ htmlstringProcessed HTML content of the page
↳ rawHtmlstringUnprocessed raw HTML content
↳ linksarrayArray of links found on the page
↳ metadataobjectPage metadata from crawl operation
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ ogLocaleAlternatearrayAlternate locale versions
totalnumberTotal number of pages found during crawl

crw_map

Get a complete list of URLs from any website quickly and reliably. Useful for discovering all pages on a site without crawling them.

Entrée

ParamètreTypeObligatoireDescription
urlstringOuiThe base URL to map and discover links from (e.g., "https://example.com"\)
limitnumberNonMaximum number of links to return (e.g., 100, 1000, 5000)
baseUrlstringNonBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringOuifastCRW API key
pricingper_requestNonNo description
rateLimitstringNonNo description

Sortie

ParamètreTypeDescription
successbooleanWhether the mapping operation was successful
linksarrayArray of discovered URLs from the website

On this page