article
Optional
Clean article body text and reading metrics
url
Optional
Original requested URL
finalUrl
Optional
Final resolved URL after redirects
httpStatus
Optional
HTTP status code
crawlerEngine
Optional
Engine used for scraping (cheerio or playwright)
title
Optional
Page title
description
Optional
Meta description
metadata
Optional
SEO, OpenGraph, and Twitter tags
markdown
Optional
Clean LLM-ready markdown content
article
Optional
Clean article body text and reading metrics
structuredData
Optional
Parsed Schema.org JSON-LD structured data objects
contacts
Optional
Extracted email addresses, telephone numbers, and social links
tables
Optional
Extracted HTML tables in structured JSON format
media
Optional
Extracted images, videos, and documents
links
Optional
Internal and external links
rawHtml
Optional
Raw unparsed HTML
screenshotUrl
Optional
URL to captured full-page screenshot
pdfUrl
Optional
URL to generated full-page PDF
responseTimeMs
Optional
Response time in milliseconds
scrapedAt
Optional
ISO timestamp of when the page was scraped