AptlyStar

fastCRW

Scrape, search, crawl, and map web data

Nutzungsanweisungen

Integrate fastCRW into the workflow. Scrape pages, search the web, crawl entire sites, and map URL structures. fastCRW is a Firecrawl-compatible web scraper in a single binary — self-host or cloud.

Tools

crw_scrape

Extract structured content from web pages with comprehensive metadata support. Converts content to markdown or HTML while capturing SEO metadata, Open Graph tags, and page information.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe URL to scrape content from (e.g., "https://example.com/page"\)
scrapeOptionsjsonNeinOptions for content scraping
baseUrlstringNeinBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringJafastCRW API key
pricingper_requestNeinNo description
rateLimitstringNeinNo description

Ausgabe

ParameterTypBeschreibung
markdownstringPage content in markdown format
htmlstringRaw HTML content of the page
metadataobjectPage metadata including SEO and Open Graph information
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code (e.g., "en")
↳ sourceURLstringOriginal source URL that was scraped
↳ statusCodenumberHTTP status code of the response
↳ keywordsstringPage meta keywords
↳ robotsstringRobots meta directive (e.g., "follow, index")
↳ ogTitlestringOpen Graph title
↳ ogDescriptionstringOpen Graph description
↳ ogUrlstringOpen Graph URL
↳ ogImagestringOpen Graph image URL
↳ ogLocaleAlternatearrayAlternate locale versions for Open Graph
↳ ogSiteNamestringOpen Graph site name
↳ errorstringError message if scrape failed

Search for information on the web using fastCRW

Eingabe

ParameterTypErforderlichBeschreibung
querystringJaThe search query to use
baseUrlstringNeinBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringJafastCRW API key
pricingper_requestNeinNo description
rateLimitstringNeinNo description

Ausgabe

ParameterTypBeschreibung
dataarraySearch results data with scraped content and metadata
↳ titlestringSearch result title from search engine
↳ descriptionstringSearch result description/snippet from search engine
↳ urlstringURL of the search result
↳ markdownstringPage content in markdown (when sources include scraped content)
↳ metadataobjectMetadata about the search result page
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ errorstringError message if scrape failed

crw_crawl

Crawl entire websites and extract structured content from all accessible pages

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe website URL to crawl (e.g., "https://example.com" or "https://docs.example.com/guide"\)
maxPagesnumberNeinMaximum number of pages to crawl (e.g., 50, 100, 500). Default: 100
maxDepthnumberNeinMaximum depth to crawl from the starting URL (e.g., 1, 2, 3). Controls how many levels deep to follow links
formatsjsonNeinOutput formats for scraped content (e.g., ["markdown"], ["markdown", "html"], ["markdown", "links"])
excludePathsjsonNeinURL paths to exclude from crawling (e.g., ["/blog/", "/admin/", "/*.pdf"])
includePathsjsonNeinURL paths to include in crawling (e.g., ["/docs/", "/api/"]). Only these paths will be crawled
onlyMainContentbooleanNeinExtract only main content from pages
baseUrlstringNeinBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringJafastCRW API key
pricingper_requestNeinNo description
rateLimitstringNeinNo description

Ausgabe

ParameterTypBeschreibung
pagesarrayArray of crawled pages with their content and metadata
↳ markdownstringPage content in markdown format
↳ htmlstringProcessed HTML content of the page
↳ rawHtmlstringUnprocessed raw HTML content
↳ linksarrayArray of links found on the page
↳ metadataobjectPage metadata from crawl operation
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ ogLocaleAlternatearrayAlternate locale versions
totalnumberTotal number of pages found during crawl

crw_map

Get a complete list of URLs from any website quickly and reliably. Useful for discovering all pages on a site without crawling them.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe base URL to map and discover links from (e.g., "https://example.com"\)
limitnumberNeinMaximum number of links to return (e.g., 100, 1000, 5000)
baseUrlstringNeinBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringJafastCRW API key
pricingper_requestNeinNo description
rateLimitstringNeinNo description

Ausgabe

ParameterTypBeschreibung
successbooleanWhether the mapping operation was successful
linksarrayArray of discovered URLs from the website

On this page