Recipe, Event & Job Data Extractor (schema.org JSON-LD) avatar

Recipe, Event & Job Data Extractor (schema.org JSON-LD)

Pricing

$1.00 / 1,000 items

Go to Apify Store
Recipe, Event & Job Data Extractor (schema.org JSON-LD)

Recipe, Event & Job Data Extractor (schema.org JSON-LD)

Extract clean recipes (ingredients, steps, times, nutrition), events (dates, venue, tickets), job postings (salary, location, remote), products, articles, businesses, FAQs and more from any website that publishes schema.org data for Google. One row per item, normalized fields. $1 per 1,000 items.

Pricing

$1.00 / 1,000 items

Rating

0.0

(0)

Developer

SwiftKit

SwiftKit

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Share

Millions of websites describe their content in schema.org structured data so Google can show rich results: recipe cards, event listings, Google for Jobs, product prices. This tool reads that data from any page and turns it into clean, normalized rows, one per item:

  • Recipes: name, author, ingredients list, step-by-step instructions, prep / cook / total minutes, yield, category, cuisine, keywords, calories and nutrition, rating
  • Events: name, start and end, status (scheduled, cancelled, moved online), venue, address, coordinates, online URL, organizer, performers, ticket price, currency, availability, ticket link
  • Job postings: title, company, date posted, valid through, employment type, remote flag, locations, salary range with currency and unit, description
  • Products: name, brand, SKU, GTIN, price range, currency, availability, rating
  • Articles: headline, authors, publisher, published and modified dates, section, keywords
  • Businesses: name, address, coordinates, phone, opening hours, price range, cuisine, rating
  • FAQs (question and answer pairs), how-to guides, courses, apps & software

Paste the pages, or start from an index page (a recipe category, an events calendar, a careers page) and let it follow links on the same site.

Who it's for

  • Food apps and meal planners: ingredients, steps and times from recipe sites.
  • Event aggregators and city guides: dates, venues and tickets from venue and organizer sites.
  • Job boards and recruiters: postings from company career pages that publish Google for Jobs data.
  • SEO teams: see exactly what structured data a page exposes (turn on raw JSON-LD).

Input

OptionDefaultWhat it does
Pages–Any URLs
Only these typesallRecipes, events, jobs, products, articles…
Follow links on the same siteoffCrawl from index pages
Only follow links containing–e.g. /recipe/
Max pages / Max items50 / 1,000Limits
Include raw JSON-LDoffThe original object too
Respect robots.txtonSkip disallowed pages

Output

A real recipe (loveandlemons.com):

{
"status": "ok",
"itemType": "Recipe",
"name": "BEST Hummus",
"author": "Jeanine Donofrio",
"prepMinutes": 5,
"totalMinutes": 5,
"ingredients": ["…9 items…"],
"instructions": ["…2 steps…"],
"rating": 4.96
}

A real job posting (a Lever career page):

{
"itemType": "JobPosting",
"title": "Client Partner - Emerging & Scaled, Independent Agency (UK)",
"company": "Spotify",
"datePosted": "2026-09-17",
"employmentType": ["Permanent"],
"remote": false,
"locations": [{ "city": "London" }]
}

Generic "Article" wrappers and site-wide organization boilerplate are skipped automatically when a page has a specific item (a recipe, a job…), so you don't pay for noise.

statusMeaning
okItem extracted.
no_structured_dataThe page has no supported structured data. Not charged.
blocked_by_robots_txt / blocked_by_bot_protectionThe site doesn't want bots. Not charged.
unreachableThe page didn't load (some big publishers answer bots with 402/403). Not charged.

Use it from your AI agent (MCP)

Claude, Cursor and other MCP clients can call this tool directly through Apify's MCP server. Add it as an MCP server / connector:

{
"mcpServers": {
"apify": {
"url": "https://mcp.apify.com?tools=swiftkit/structured-data",
"headers": { "Authorization": "Bearer <YOUR_APIFY_TOKEN>" }
}
}
}

Then just ask, for example: "Get the ingredients and total time from this recipe: https://…". Each call is billed like a normal run.

Pricing

You pay per extracted item. Every other status is free. See the Pricing tab.

Limits, honestly

  • Only data the site publishes as JSON-LD in its HTML. Pages that add it with JavaScript aren't covered.
  • Some large publishers block automated visitors entirely (for example with HTTP 402 or 403). The tool respects that and moves on.
  • Recipe instructions and descriptions are the site's own text: use them in line with the site's terms and copyright (ingredients and facts are generally fine; republishing whole recipes may not be).
  • People are never extracted as items on their own; author and organizer names appear only as part of the item they belong to.

More tools from SwiftKit

Questions?

Open an issue on the Issues tab.