AptlyStar

fastCRW

Scrape, search, crawl, and map web data

Usage Instructions

Integrate fastCRW into the workflow. Scrape pages, search the web, crawl entire sites, and map URL structures. fastCRW is a Firecrawl-compatible web scraper in a single binary — self-host or cloud.

Tools

crw_scrape

Extract structured content from web pages with comprehensive metadata support. Converts content to markdown or HTML while capturing SEO metadata, Open Graph tags, and page information.

Input

ParameterTypeRequiredDescription
urlstringYesThe URL to scrape content from (e.g., "https://example.com/page"\)
scrapeOptionsjsonNoOptions for content scraping
baseUrlstringNoBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringYesfastCRW API key
pricingper_requestNoNo description
rateLimitstringNoNo description

Output

ParameterTypeDescription
markdownstringPage content in markdown format
htmlstringRaw HTML content of the page
metadataobjectPage metadata including SEO and Open Graph information
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code (e.g., "en")
↳ sourceURLstringOriginal source URL that was scraped
↳ statusCodenumberHTTP status code of the response
↳ keywordsstringPage meta keywords
↳ robotsstringRobots meta directive (e.g., "follow, index")
↳ ogTitlestringOpen Graph title
↳ ogDescriptionstringOpen Graph description
↳ ogUrlstringOpen Graph URL
↳ ogImagestringOpen Graph image URL
↳ ogLocaleAlternatearrayAlternate locale versions for Open Graph
↳ ogSiteNamestringOpen Graph site name
↳ errorstringError message if scrape failed

Search for information on the web using fastCRW

Input

ParameterTypeRequiredDescription
querystringYesThe search query to use
baseUrlstringNoBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringYesfastCRW API key
pricingper_requestNoNo description
rateLimitstringNoNo description

Output

ParameterTypeDescription
dataarraySearch results data with scraped content and metadata
↳ titlestringSearch result title from search engine
↳ descriptionstringSearch result description/snippet from search engine
↳ urlstringURL of the search result
↳ markdownstringPage content in markdown (when sources include scraped content)
↳ metadataobjectMetadata about the search result page
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ errorstringError message if scrape failed

crw_crawl

Crawl entire websites and extract structured content from all accessible pages

Input

ParameterTypeRequiredDescription
urlstringYesThe website URL to crawl (e.g., "https://example.com" or "https://docs.example.com/guide"\)
maxPagesnumberNoMaximum number of pages to crawl (e.g., 50, 100, 500). Default: 100
maxDepthnumberNoMaximum depth to crawl from the starting URL (e.g., 1, 2, 3). Controls how many levels deep to follow links
formatsjsonNoOutput formats for scraped content (e.g., ["markdown"], ["markdown", "html"], ["markdown", "links"])
excludePathsjsonNoURL paths to exclude from crawling (e.g., ["/blog/", "/admin/", "/*.pdf"])
includePathsjsonNoURL paths to include in crawling (e.g., ["/docs/", "/api/"]). Only these paths will be crawled
onlyMainContentbooleanNoExtract only main content from pages
baseUrlstringNoBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringYesfastCRW API key
pricingper_requestNoNo description
rateLimitstringNoNo description

Output

ParameterTypeDescription
pagesarrayArray of crawled pages with their content and metadata
↳ markdownstringPage content in markdown format
↳ htmlstringProcessed HTML content of the page
↳ rawHtmlstringUnprocessed raw HTML content
↳ linksarrayArray of links found on the page
↳ metadataobjectPage metadata from crawl operation
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ ogLocaleAlternatearrayAlternate locale versions
totalnumberTotal number of pages found during crawl

crw_map

Get a complete list of URLs from any website quickly and reliably. Useful for discovering all pages on a site without crawling them.

Input

ParameterTypeRequiredDescription
urlstringYesThe base URL to map and discover links from (e.g., "https://example.com"\)
limitnumberNoMaximum number of links to return (e.g., 100, 1000, 5000)
baseUrlstringNoBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringYesfastCRW API key
pricingper_requestNoNo description
rateLimitstringNoNo description

Output

ParameterTypeDescription
successbooleanWhether the mapping operation was successful
linksarrayArray of discovered URLs from the website

On this page