AptlyStar

Context.dev

Scrape, crawl, search, extract, and enrich web and brand data

Nutzungsanweisungen

Integrate Context.dev into the workflow. Scrape pages to markdown or HTML, capture screenshots, list images, crawl entire sites, map sitemaps, search the web, extract structured data and products, pull design systems, classify industries, and retrieve brand assets by domain, name, email, ticker, or transaction ΓÇö all from one API.

Tools

context_dev_scrape_markdown

Scrape any URL and return clean, LLM-ready markdown content.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe full URL to scrape (must include http:// or https://)
useMainContentOnlybooleanNeinReturn only main content, excluding headers, footers, and navigation
includeLinksbooleanNeinPreserve hyperlinks in the markdown output (default: true)
includeImagesbooleanNeinInclude image references in the markdown output (default: false)
includeFramesbooleanNeinRender iframe contents inline (default: false)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 86400000)
waitForMsnumberNeinBrowser wait time after page load in milliseconds (0-30000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
markdownstringPage content as clean markdown
urlstringThe scraped URL

context_dev_scrape_html

Scrape any URL and return the raw HTML content of the page.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe full URL to scrape (must include http:// or https://)
useMainContentOnlybooleanNeinReturn only main content, excluding headers, footers, and navigation
includeFramesbooleanNeinRender iframe contents inline into the returned HTML (default: false)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 86400000)
waitForMsnumberNeinBrowser wait time after page load in milliseconds (0-30000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
htmlstringRaw HTML content of the page
urlstringThe scraped URL
typestringDetected content type (html, xml, json, text, csv, markdown, svg, pdf)

context_dev_scrape_images

Discover every image asset on a page, with optional dimension and type enrichment.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe full URL to scrape images from (must include http:// or https://)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 86400000)
waitForMsnumberNeinBrowser wait time after page load in milliseconds (0-30000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
enrichResolutionbooleanNeinMeasure image dimensions (enables 5-credit enrichment)
enrichHostedUrlbooleanNeinHost images on a CDN and return their URL and MIME type (enables enrichment)
enrichClassificationbooleanNeinClassify each image by visual asset type (enables enrichment)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
successbooleanWhether the scrape succeeded
imagesarrayDiscovered image assets with source, element, type, and optional enrichment
↳ srcstringImage source URL or data
↳ elementstringSource element (img, svg, link, source, video, css, object, meta, background)
↳ typestringImage representation (url, html, base64)
↳ altstringAlt text
↳ enrichmentjsonOptional enrichment (width, height, mimetype, url, type) when requested
urlstringThe scraped URL

context_dev_screenshot

Capture a screenshot of any web page and store it as a downloadable image file.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe full URL to capture (must include http:// or https://)
fullScreenshotbooleanNeinCapture the full scrollable page instead of just the viewport (default: false)
handleCookiePopupbooleanNeinAttempt to dismiss cookie banners before capturing (default: false)
viewportWidthnumberNeinViewport width in pixels (240-7680, default: 1920)
viewportHeightnumberNeinViewport height in pixels (240-4320, default: 1080)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 86400000)
waitForMsnumberNeinPost-load delay before capturing in milliseconds (0-30000, default: 3000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
filefileStored screenshot image file
screenshotUrlstringPublic URL of the captured screenshot
screenshotTypestringScreenshot type (viewport or fullPage)
domainstringDomain that was captured
widthnumberScreenshot width in pixels
heightnumberScreenshot height in pixels

context_dev_crawl

Crawl an entire website and return each discovered page as clean markdown.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe starting URL to crawl (must include http:// or https://)
maxPagesnumberNeinMaximum number of pages to crawl (1-500, default: 100)
maxDepthnumberNeinMaximum link depth from the starting URL (0 = start page only)
urlRegexstringNeinRegex pattern to filter which URLs are crawled
includeLinksbooleanNeinPreserve hyperlinks in the markdown output (default: true)
includeImagesbooleanNeinInclude image references in the markdown output (default: false)
useMainContentOnlybooleanNeinStrip headers, footers, and sidebars from each page (default: false)
followSubdomainsbooleanNeinFollow links to subdomains of the starting domain (default: false)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 86400000)
waitForMsnumberNeinBrowser wait time after page load in milliseconds (0-30000)
stopAfterMsnumberNeinSoft crawl time budget in milliseconds (10000-110000, default: 80000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
resultsarrayCrawled pages with markdown content and per-page metadata
↳ markdownstringPage content as markdown
↳ metadatajsonPage metadata (url, title, crawlDepth, statusCode)
metadataobjectCrawl summary (numUrls, maxCrawlDepth, numSucceeded, numFailed, numSkipped)

context_dev_map

Build a sitemap of a domain and return every discovered page URL.

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to build a sitemap for (e.g., "example.com")
maxLinksnumberNeinMaximum number of URLs to return (1-100000, default: 10000)
urlRegexstringNeinRE2-compatible regex to filter URLs (max 256 chars)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
domainstringThe domain that was mapped
urlsarrayAll page URLs discovered from the sitemap
metaobjectSitemap discovery stats (sitemapsDiscovered, sitemapsFetched, errors)

Search the web with natural language and optionally scrape results to markdown.

Eingabe

ParameterTypErforderlichBeschreibung
querystringJaThe natural language search query (1-500 characters)
includeDomainsarrayNeinOnly return results from these domains
excludeDomainsarrayNeinExclude results from these domains
freshnessstringNeinRecency filter (last_24_hours, last_week, last_month, last_year)
queryFanoutbooleanNeinExpand the query into parallel variants for broader coverage
markdownEnabledbooleanNeinScrape each result page to markdown (default: false)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
resultsarraySearch results with url, title, description, relevance, and optional markdown
↳ urlstringResult page URL
↳ titlestringResult page title
↳ descriptionstringResult snippet/description
↳ relevancestringRelevance rating (high, medium, low)
↳ markdownjsonScraped markdown for the result (when markdown scraping is enabled)
querystringThe query that was searched

context_dev_extract

Crawl a website and extract structured data matching a provided JSON schema.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe starting website URL (must include http:// or https://)
schemajsonJaJSON Schema describing the structure of the data to extract
instructionsstringNeinOptional extraction guidance for link prioritization (max 2000 chars)
factCheckbooleanNeinRequire extracted values to be grounded in page facts (default: false)
followSubdomainsbooleanNeinFollow links on subdomains of the starting domain (default: false)
maxPagesnumberNeinMaximum number of pages to analyze (1-50, default: 5)
maxDepthnumberNeinMaximum link depth from the starting URL
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 604800000)
stopAfterMsnumberNeinSoft crawl time budget in milliseconds (10000-110000, default: 80000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringExtraction status
urlstringThe starting URL that was crawled
urlsAnalyzedarrayURLs that were analyzed during extraction
datajsonStructured data matching the requested schema
metadataobjectCrawl summary (numUrls, maxCrawlDepth, numSucceeded, numFailed, numSkipped)

context_dev_extract_product

Detect and extract structured product details from a single product page URL.

Eingabe

ParameterTypErforderlichBeschreibung
urlstringJaThe product page URL (must include http:// or https://)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 604800000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
isProductPagebooleanWhether the URL is a product page
platformstringDetected platform (amazon, tiktok_shop, etsy, generic)
productobjectExtracted product details
↳ namestringProduct name
↳ descriptionstringProduct description
↳ pricenumberProduct price
↳ currencystringPrice currency
↳ billing_frequencystringBilling frequency (monthly, yearly, one_time, usage_based)
↳ pricing_modelstringPricing model (per_seat, flat, tiered, freemium, custom)
↳ urlstringProduct URL
↳ categorystringProduct category
↳ featuresjsonProduct features
↳ target_audiencejsonTarget audience
↳ tagsjsonProduct tags
↳ image_urlstringPrimary product image URL
↳ imagesjsonProduct image URLs
↳ skustringProduct SKU

context_dev_extract_products

Extract the product catalog from a brand's website by domain (beta).

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to extract products from (e.g., "example.com")
maxProductsnumberNeinMaximum number of products to extract (1-12)
maxAgeMsnumberNeinCache duration in milliseconds (0-2592000000, default: 604800000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
productsarrayExtracted products with pricing, features, and metadata
↳ namestringProduct name
↳ descriptionstringProduct description
↳ pricenumberProduct price
↳ currencystringPrice currency
↳ billing_frequencystringBilling frequency (monthly, yearly, one_time, usage_based)
↳ pricing_modelstringPricing model (per_seat, flat, tiered, freemium, custom)
↳ urlstringProduct URL
↳ categorystringProduct category
↳ featuresjsonProduct features
↳ target_audiencejsonTarget audience
↳ tagsjsonProduct tags
↳ image_urlstringPrimary product image URL
↳ imagesjsonProduct image URLs
↳ skustringProduct SKU

context_dev_scrape_fonts

Extract the font families, usage stats, and font files used by a domain.

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to extract fonts from (e.g., "example.com")
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringExtraction status
domainstringThe domain that was analyzed
fontsarrayFonts with usage statistics and fallbacks
↳ fontstringFont family name
↳ usesjsonWhere the font is used
↳ fallbacksjsonFallback font families
↳ num_elementsnumberNumber of elements using the font
↳ num_wordsnumberNumber of words rendered in the font
↳ percent_wordsnumberPercent of words using the font
↳ percent_elementsnumberPercent of elements using the font
fontLinksjsonFont family download links keyed by font name (type, files, category)

context_dev_scrape_styleguide

Extract a domain's design system: colors, typography, spacing, shadows, and UI components.

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to extract the styleguide from (e.g., "example.com")
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringExtraction status
domainstringThe domain that was analyzed
styleguidejsonDesign system: mode, colors, typography, elementSpacing, shadows, fontLinks, components

context_dev_classify_naics

Classify a brand into NAICS industry codes from its domain or company name.

Eingabe

ParameterTypErforderlichBeschreibung
inputstringJaBrand domain or company name to classify (e.g., "stripe.com" or "Stripe")
minResultsnumberNeinMinimum number of codes to return (1-10, default: 1)
maxResultsnumberNeinMaximum number of codes to return (1-10, default: 5)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringClassification status
domainstringResolved domain
typestringInput type that was resolved
codesarrayMatched NAICS codes with name and confidence
↳ codestringIndustry code
↳ namestringIndustry name
↳ confidencestringMatch confidence (high, medium, low)

context_dev_classify_sic

Classify a brand into SIC industry codes from its domain or company name.

Eingabe

ParameterTypErforderlichBeschreibung
inputstringJaBrand domain or company name to classify (e.g., "stripe.com" or "Stripe")
typestringNeinSIC taxonomy version: "original_sic" (default) or "latest_sec"
minResultsnumberNeinMinimum number of codes to return (1-10, default: 1)
maxResultsnumberNeinMaximum number of codes to return (1-10, default: 5)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringClassification status
domainstringResolved domain
typestringInput type that was resolved
classificationstringSIC taxonomy version used (original_sic or latest_sec)
codesarrayMatched SIC codes with name, confidence, and group metadata
↳ codestringIndustry code
↳ namestringIndustry name
↳ confidencestringMatch confidence (high, medium, low)
↳ majorGroupstringMajor group code (original_sic only)
↳ majorGroupNamestringMajor group name (original_sic only)
↳ officestringSEC office (latest_sec only)

context_dev_get_brand

Retrieve brand data for a domain: logos, colors, backdrops, socials, address, and industry.

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to retrieve brand data for (e.g., "airbnb.com")
forceLanguagestringNeinOverride the detected language with a supported language code
maxSpeedbooleanNeinSkip time-consuming operations for a faster response (default: false)
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringRetrieval status
brandobjectBrand data object
↳ domainstringBrand domain
↳ titlestringBrand title
↳ descriptionstringBrand description
↳ sloganstringBrand slogan
↳ colorsjsonBrand colors (hex and name)
↳ logosjsonBrand logos with mode, colors, resolution, and type
↳ backdropsjsonBrand backdrop images
↳ socialsjsonSocial media profiles (type and url)
↳ addressjsonBrand address
↳ stockjsonStock info (ticker and exchange)
↳ is_nsfwbooleanWhether the brand contains adult content
↳ emailstringBrand contact email
↳ phonestringBrand contact phone
↳ industriesjsonIndustry taxonomy (eic industry/subindustry pairs)
↳ linksjsonKey brand links (careers, privacy, terms, blog, pricing)
↳ primary_languagestringPrimary language of the brand site

context_dev_get_brand_by_name

Retrieve brand data by company name: logos, colors, socials, address, and industry.

Eingabe

ParameterTypErforderlichBeschreibung
namestringJaCompany name to retrieve brand data for (3-30 chars, e.g., "Apple Inc")
countryGlstringNeinISO 2-letter country code to prioritize (e.g., "us")
forceLanguagestringNeinOverride the detected language with a supported language code
maxSpeedbooleanNeinSkip time-consuming operations for a faster response (default: false)
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringRetrieval status
brandobjectBrand data object
↳ domainstringBrand domain
↳ titlestringBrand title
↳ descriptionstringBrand description
↳ sloganstringBrand slogan
↳ colorsjsonBrand colors (hex and name)
↳ logosjsonBrand logos with mode, colors, resolution, and type
↳ backdropsjsonBrand backdrop images
↳ socialsjsonSocial media profiles (type and url)
↳ addressjsonBrand address
↳ stockjsonStock info (ticker and exchange)
↳ is_nsfwbooleanWhether the brand contains adult content
↳ emailstringBrand contact email
↳ phonestringBrand contact phone
↳ industriesjsonIndustry taxonomy (eic industry/subindustry pairs)
↳ linksjsonKey brand links (careers, privacy, terms, blog, pricing)
↳ primary_languagestringPrimary language of the brand site

context_dev_get_brand_by_email

Retrieve brand data from a work email address. Free/disposable emails are rejected (422).

Eingabe

ParameterTypErforderlichBeschreibung
emailstringJaWork email address; the domain is extracted (free providers are rejected)
forceLanguagestringNeinOverride the detected language with a supported language code
maxSpeedbooleanNeinSkip time-consuming operations for a faster response (default: false)
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringRetrieval status
brandobjectBrand data object
↳ domainstringBrand domain
↳ titlestringBrand title
↳ descriptionstringBrand description
↳ sloganstringBrand slogan
↳ colorsjsonBrand colors (hex and name)
↳ logosjsonBrand logos with mode, colors, resolution, and type
↳ backdropsjsonBrand backdrop images
↳ socialsjsonSocial media profiles (type and url)
↳ addressjsonBrand address
↳ stockjsonStock info (ticker and exchange)
↳ is_nsfwbooleanWhether the brand contains adult content
↳ emailstringBrand contact email
↳ phonestringBrand contact phone
↳ industriesjsonIndustry taxonomy (eic industry/subindustry pairs)
↳ linksjsonKey brand links (careers, privacy, terms, blog, pricing)
↳ primary_languagestringPrimary language of the brand site

context_dev_get_brand_by_ticker

Retrieve brand data for a public company by its stock ticker symbol.

Eingabe

ParameterTypErforderlichBeschreibung
tickerstringJaStock ticker symbol (e.g., "AAPL", "GOOGL", "BRK.A")
tickerExchangestringNeinExchange code for the ticker (e.g., "NASDAQ", "NYSE", "LSE"). Default: NASDAQ
forceLanguagestringNeinOverride the detected language with a supported language code
maxSpeedbooleanNeinSkip time-consuming operations for a faster response (default: false)
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringRetrieval status
brandobjectBrand data object
↳ domainstringBrand domain
↳ titlestringBrand title
↳ descriptionstringBrand description
↳ sloganstringBrand slogan
↳ colorsjsonBrand colors (hex and name)
↳ logosjsonBrand logos with mode, colors, resolution, and type
↳ backdropsjsonBrand backdrop images
↳ socialsjsonSocial media profiles (type and url)
↳ addressjsonBrand address
↳ stockjsonStock info (ticker and exchange)
↳ is_nsfwbooleanWhether the brand contains adult content
↳ emailstringBrand contact email
↳ phonestringBrand contact phone
↳ industriesjsonIndustry taxonomy (eic industry/subindustry pairs)
↳ linksjsonKey brand links (careers, privacy, terms, blog, pricing)
↳ primary_languagestringPrimary language of the brand site

context_dev_get_brand_simplified

Retrieve essential brand data for a domain: title, colors, logos, and backdrops.

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to retrieve simplified brand data for (e.g., "airbnb.com")
maxAgeMsnumberNeinCache max age in milliseconds (86400000-31536000000, default: 7776000000)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringRetrieval status
brandobjectSimplified brand data (domain, title, colors, logos, backdrops)
↳ domainstringBrand domain
↳ titlestringBrand title
↳ colorsjsonBrand colors (hex and name)
↳ logosjsonBrand logos with mode, colors, resolution, and type
↳ backdropsjsonBrand backdrop images

context_dev_identify_transaction

Identify the brand behind a raw bank/card transaction descriptor and return its brand data.

Eingabe

ParameterTypErforderlichBeschreibung
transactionInfostringJaThe raw transaction descriptor or identifier to resolve to a brand
countryGlstringNeinISO 2-letter country code from the transaction (e.g., "us", "gb")
citystringNeinCity name to prioritize in the search
mccstringNeinMerchant Category Code for the business category
phonenumberNeinPhone number from the transaction for verification
highConfidenceOnlybooleanNeinEnforce additional verification steps for higher confidence (default: false)
forceLanguagestringNeinOverride the detected language with a supported language code
maxSpeedbooleanNeinSkip time-consuming operations for a faster response (default: false)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringIdentification status
brandobjectBrand data for the identified merchant
↳ domainstringBrand domain
↳ titlestringBrand title
↳ descriptionstringBrand description
↳ sloganstringBrand slogan
↳ colorsjsonBrand colors (hex and name)
↳ logosjsonBrand logos with mode, colors, resolution, and type
↳ backdropsjsonBrand backdrop images
↳ socialsjsonSocial media profiles (type and url)
↳ addressjsonBrand address
↳ stockjsonStock info (ticker and exchange)
↳ is_nsfwbooleanWhether the brand contains adult content
↳ emailstringBrand contact email
↳ phonestringBrand contact phone
↳ industriesjsonIndustry taxonomy (eic industry/subindustry pairs)
↳ linksjsonKey brand links (careers, privacy, terms, blog, pricing)
↳ primary_languagestringPrimary language of the brand site

context_dev_prefetch_domain

Queue a domain for brand-data prefetching to reduce latency on later requests (subscribers; 0 credits).

Eingabe

ParameterTypErforderlichBeschreibung
domainstringJaThe domain to prefetch brand data for (e.g., "example.com")
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringPrefetch status
messagestringHuman-readable prefetch result message
domainstringThe domain queued for prefetching

context_dev_prefetch_by_email

Queue an email's domain for brand-data prefetching to reduce later latency (subscribers; 0 credits). Free/disposable emails are rejected.

Eingabe

ParameterTypErforderlichBeschreibung
emailstringJaWork email address whose domain should be prefetched (free providers rejected)
timeoutMSnumberNeinRequest timeout in milliseconds (1000-300000)
apiKeystringJaContext.dev API key

Ausgabe

ParameterTypBeschreibung
statusstringPrefetch status
messagestringHuman-readable prefetch result message
domainstringThe domain queued for prefetching

On this page