Recipe, Event & Job Data Extractor (schema.org JSON-LD)
Pricing
$1.00 / 1,000 items
Recipe, Event & Job Data Extractor (schema.org JSON-LD)
Extract clean recipes (ingredients, steps, times, nutrition), events (dates, venue, tickets), job postings (salary, location, remote), products, articles, businesses, FAQs and more from any website that publishes schema.org data for Google. One row per item, normalized fields. $1 per 1,000 items.
Pricing
$1.00 / 1,000 items
Rating
0.0
(0)
Developer
SwiftKit
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
3 days ago
Last modified
Categories
Share
Millions of websites describe their content in schema.org structured data so Google can show rich results: recipe cards, event listings, Google for Jobs, product prices. This tool reads that data from any page and turns it into clean, normalized rows, one per item:
- Recipes: name, author, ingredients list, step-by-step instructions, prep / cook / total minutes, yield, category, cuisine, keywords, calories and nutrition, rating
- Events: name, start and end, status (scheduled, cancelled, moved online), venue, address, coordinates, online URL, organizer, performers, ticket price, currency, availability, ticket link
- Job postings: title, company, date posted, valid through, employment type, remote flag, locations, salary range with currency and unit, description
- Products: name, brand, SKU, GTIN, price range, currency, availability, rating
- Articles: headline, authors, publisher, published and modified dates, section, keywords
- Businesses: name, address, coordinates, phone, opening hours, price range, cuisine, rating
- FAQs (question and answer pairs), how-to guides, courses, apps & software
Paste the pages, or start from an index page (a recipe category, an events calendar, a careers page) and let it follow links on the same site.
Who it's for
- Food apps and meal planners: ingredients, steps and times from recipe sites.
- Event aggregators and city guides: dates, venues and tickets from venue and organizer sites.
- Job boards and recruiters: postings from company career pages that publish Google for Jobs data.
- SEO teams: see exactly what structured data a page exposes (turn on raw JSON-LD).
Input
| Option | Default | What it does |
|---|---|---|
| Pages | – | Any URLs |
| Only these types | all | Recipes, events, jobs, products, articles… |
| Follow links on the same site | off | Crawl from index pages |
| Only follow links containing | – | e.g. /recipe/ |
| Max pages / Max items | 50 / 1,000 | Limits |
| Include raw JSON-LD | off | The original object too |
| Respect robots.txt | on | Skip disallowed pages |
Output
A real recipe (loveandlemons.com):
{"status": "ok","itemType": "Recipe","name": "BEST Hummus","author": "Jeanine Donofrio","prepMinutes": 5,"totalMinutes": 5,"ingredients": ["…9 items…"],"instructions": ["…2 steps…"],"rating": 4.96}
A real job posting (a Lever career page):
{"itemType": "JobPosting","title": "Client Partner - Emerging & Scaled, Independent Agency (UK)","company": "Spotify","datePosted": "2026-09-17","employmentType": ["Permanent"],"remote": false,"locations": [{ "city": "London" }]}
Generic "Article" wrappers and site-wide organization boilerplate are skipped automatically when a page has a specific item (a recipe, a job…), so you don't pay for noise.
| status | Meaning |
|---|---|
ok | Item extracted. |
no_structured_data | The page has no supported structured data. Not charged. |
blocked_by_robots_txt / blocked_by_bot_protection | The site doesn't want bots. Not charged. |
unreachable | The page didn't load (some big publishers answer bots with 402/403). Not charged. |
Use it from your AI agent (MCP)
Claude, Cursor and other MCP clients can call this tool directly through Apify's MCP server. Add it as an MCP server / connector:
{"mcpServers": {"apify": {"url": "https://mcp.apify.com?tools=swiftkit/structured-data","headers": { "Authorization": "Bearer <YOUR_APIFY_TOKEN>" }}}}
Then just ask, for example: "Get the ingredients and total time from this recipe: https://…". Each call is billed like a normal run.
Pricing
You pay per extracted item. Every other status is free. See the Pricing tab.
Limits, honestly
- Only data the site publishes as JSON-LD in its HTML. Pages that add it with JavaScript aren't covered.
- Some large publishers block automated visitors entirely (for example with HTTP 402 or 403). The tool respects that and moves on.
- Recipe instructions and descriptions are the site's own text: use them in line with the site's terms and copyright (ingredients and facts are generally fine; republishing whole recipes may not be).
- People are never extracted as items on their own; author and organizer names appear only as part of the item they belong to.
More tools from SwiftKit
- Product Page Scraper: prices and stock from any shop, with price tracking
- Job Search API: 10,000+ company career sites in one search
- Website to Markdown for AI: clean page content for LLMs
Questions?
Open an issue on the Issues tab.