Article body
articleBody
Optional
Full article text, joined from div.text > div.text-long paragraph blocks (concatenated across all blocks for longform/syndicated pieces).
Channel NewsAsia (CNA) Scraper
Pricing
from $3.50 / 1,000 results
Extract full article text or the newest headlines from Channel NewsAsia (channelnewsasia.com) -- no account or API key needed.
Input
_input
Optional
The exact input value (URL, or 'latest') that produced this record.
Source strategy
_source
Optional
Which extraction strategy emitted this record (e.g. S1-jsonld+html, S2-rss).
Scraped at
_scrapedAt
Optional
UTC ISO 8601 timestamp when this record was scraped.
Error
_error
Optional
Present only on failed extractions -- the error kind.
Error detail
_errorDetail
Optional
Free-form diagnostic detail, present only alongside _error.
Headline
headline
Optional
Article headline (from ld+json NewsArticle node).
Description
description
Optional
Article dek/description. On latest-mode (RSS) records this is the RSS item's instead (often empty).
Article body
articleBody
Optional
Full article text, joined from div.text > div.text-long paragraph blocks (concatenated across all blocks for longform/syndicated pieces).
Date published
datePublished
Optional
ISO 8601 publish date, from ld+json.
Date modified
dateModified
Optional
ISO 8601 last-modified date, from ld+json.
Image
image
Optional
Lead image URL, extracted from the ld+json image object's 'url' field.
Title
title
Optional
Headline from the RSS item (latest mode).
Link
link
Optional
Article URL (latest mode).
GUID
guid
Optional
RSS item GUID (latest mode).
Publish date
pubDate
Optional
RSS pubDate string (latest mode).
Category
category
Optional
RSS item category (latest mode).
Thumbnail URL
thumbnailUrl
Optional
RSS media:thumbnail url attribute (latest mode).