AptlyStar

fastCRW

Scrape, search, crawl, and map web data

使用方法

Integrate fastCRW into the workflow. Scrape pages, search the web, crawl entire sites, and map URL structures. fastCRW is a Firecrawl-compatible web scraper in a single binary — self-host or cloud.

ツール

crw_scrape

Extract structured content from web pages with comprehensive metadata support. Converts content to markdown or HTML while capturing SEO metadata, Open Graph tags, and page information.

入力

パラメータ型必須説明
urlstringはいThe URL to scrape content from (e.g., "https://example.com/page"\)
scrapeOptionsjsonいいえOptions for content scraping
baseUrlstringいいえBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringはいfastCRW API key
pricingper_requestいいえNo description
rateLimitstringいいえNo description

出力

パラメータ型説明
markdownstringPage content in markdown format
htmlstringRaw HTML content of the page
metadataobjectPage metadata including SEO and Open Graph information
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code (e.g., "en")
↳ sourceURLstringOriginal source URL that was scraped
↳ statusCodenumberHTTP status code of the response
↳ keywordsstringPage meta keywords
↳ robotsstringRobots meta directive (e.g., "follow, index")
↳ ogTitlestringOpen Graph title
↳ ogDescriptionstringOpen Graph description
↳ ogUrlstringOpen Graph URL
↳ ogImagestringOpen Graph image URL
↳ ogLocaleAlternatearrayAlternate locale versions for Open Graph
↳ ogSiteNamestringOpen Graph site name
↳ errorstringError message if scrape failed

Search for information on the web using fastCRW

入力

パラメータ型必須説明
querystringはいThe search query to use
baseUrlstringいいえBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringはいfastCRW API key
pricingper_requestいいえNo description
rateLimitstringいいえNo description

出力

パラメータ型説明
dataarraySearch results data with scraped content and metadata
↳ titlestringSearch result title from search engine
↳ descriptionstringSearch result description/snippet from search engine
↳ urlstringURL of the search result
↳ markdownstringPage content in markdown (when sources include scraped content)
↳ metadataobjectMetadata about the search result page
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ errorstringError message if scrape failed

crw_crawl

Crawl entire websites and extract structured content from all accessible pages

入力

パラメータ型必須説明
urlstringはいThe website URL to crawl (e.g., "https://example.com" or "https://docs.example.com/guide"\)
maxPagesnumberいいえMaximum number of pages to crawl (e.g., 50, 100, 500). Default: 100
maxDepthnumberいいえMaximum depth to crawl from the starting URL (e.g., 1, 2, 3). Controls how many levels deep to follow links
formatsjsonいいえOutput formats for scraped content (e.g., ["markdown"], ["markdown", "html"], ["markdown", "links"])
excludePathsjsonいいえURL paths to exclude from crawling (e.g., ["/blog/", "/admin/", "/*.pdf"])
includePathsjsonいいえURL paths to include in crawling (e.g., ["/docs/", "/api/"]). Only these paths will be crawled
onlyMainContentbooleanいいえExtract only main content from pages
baseUrlstringいいえBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringはいfastCRW API key
pricingper_requestいいえNo description
rateLimitstringいいえNo description

出力

パラメータ型説明
pagesarrayArray of crawled pages with their content and metadata
↳ markdownstringPage content in markdown format
↳ htmlstringProcessed HTML content of the page
↳ rawHtmlstringUnprocessed raw HTML content
↳ linksarrayArray of links found on the page
↳ metadataobjectPage metadata from crawl operation
↳ titlestringPage title
↳ descriptionstringPage meta description
↳ languagestringPage language code
↳ sourceURLstringOriginal source URL
↳ statusCodenumberHTTP status code
↳ ogLocaleAlternatearrayAlternate locale versions
totalnumberTotal number of pages found during crawl

crw_map

Get a complete list of URLs from any website quickly and reliably. Useful for discovering all pages on a site without crawling them.

入力

パラメータ型必須説明
urlstringはいThe base URL to map and discover links from (e.g., "https://example.com"\)
limitnumberいいえMaximum number of links to return (e.g., 100, 1000, 5000)
baseUrlstringいいえBase URL for self-hosted fastCRW (defaults to https://fastcrw.com/api\)
apiKeystringはいfastCRW API key
pricingper_requestいいえNo description
rateLimitstringいいえNo description

出力

パラメータ型説明
successbooleanWhether the mapping operation was successful
linksarrayArray of discovered URLs from the website

On this page