Perplexity Scraper — Answers, Citations & Sources avatar

Perplexity Scraper — Answers, Citations & Sources

Pricing

from $6.40 / 1,000 premium calls

Go to Apify Store
Perplexity Scraper — Answers, Citations & Sources

Perplexity Scraper — Answers, Citations & Sources

Ask Perplexity any prompt at scale and get the answer, its markdown, every citation, related questions and the source domains behind it as structured JSON. For answer-engine optimisation and AI search monitoring.

Pricing

from $6.40 / 1,000 premium calls

Rating

0.0

(0)

Developer

ScrapeBadger

ScrapeBadger

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Categories

Share

What does Perplexity Scraper do?

Perplexity Scraper extracts structured data from Perplexity and delivers it as clean JSON — no proxies, browsers or anti-bot handling on your side. It is a practical Perplexity API alternative for teams who need Perplexity data on a schedule rather than a one-off export.

It covers 1 different Perplexity surface behind a single Mode dropdown, so one Actor replaces a rack of single-purpose scrapers.

Why scrape Perplexity?

  • Every surface in one place — pick a mode, fill two fields, run.
  • Anti-bot handled upstream — residential proxy rotation, TLS and browser fingerprinting, and CAPTCHA solving happen inside the ScrapeBadger API.
  • Pay per call, not per row — a page of results costs the same as one lookup, so bulk work stays cheap.
  • Runs on the Apify platform — scheduling, monitoring, webhooks, the API, and integrations with Make, Zapier, Google Sheets, Slack and Airtable all work out of the box.
  • Partial results are kept — a transient upstream failure ends the run cleanly instead of discarding what you already paid for.

What data can Perplexity Scraper extract?

Every record is pushed to the dataset as its own row. The most useful fields are below; the full record carries considerably more.

FieldTypeDescription
promptstringThe prompt that was sent
answerstringThe full answer text returned by the model
citation_countnumberHow many sources the answer cited
modelstringModel that produced the answer
latency_msnumberHow long the upstream call took

Perplexity scraping modes

Pick one Mode; the input form marks the fields it needs.

ModeWhat it returnsCharged as
AskOne Perplexity answer with citations, related questions and sources.premium-call

How to scrape Perplexity

  1. Click Try for free and sign in to Apify.
  2. Open Settings → Environment variables and add SCRAPEBADGER_API_KEY with your key from scrapebadger.com, ticking Secret.
  3. Choose a Mode from the dropdown.
  4. Fill in the fields that mode needs — the description on each field says which modes use it.
  5. Set Max items to cap the run.
  6. Press Start and watch the dataset fill up.
  7. Export as JSON, CSV, Excel or XML, or pull it from the API tab.

How much will it cost to scrape Perplexity?

This Actor is pay per event: one event per API call, whatever that call returns. Fetching a page of 100 records costs the same as fetching one record, so larger pages are cheaper per row.

EventPrice per eventWhat triggers it
premium-call$0.0080A premium call — an LLM answer or brand-visibility measurement.

Apify also charges its standard $0.001 actor start fee per run. ScrapeBadger credits are consumed on your own account on top of this.

Input

See the Input tab for every option with inline documentation. A minimal run looks like this:

{
"mode": "Ask",
"prompt": "What are the best web scraping APIs in 2026?",
"country": "US",
"max_items": 100
}

Max items caps the run; the Actor pages until it reaches that number or Perplexity runs out of results.

Output

You can download the dataset in various formats such as JSON, HTML, CSV, or Excel, or read it straight from the Apify API. Each record is one row:

{
"prompt": "<prompt>",
"answer": "<answer>",
"citation_count": 1234,
"model": "<model>",
"latency_ms": 1234
}

What can you do with Perplexity data?

A few things teams actually build with this Actor.

Perplexity returns its own follow-up questions with every answer. That list is effectively a 'People Also Ask' feed for answer engines, and it makes an excellent content calendar.

Audit citation share in your category

Perplexity is the most citation-forward of the assistants. Counting how often your domain appears across a prompt set gives you a hard number for how visible you are in AI search.

Feed an agent with sourced answers

Because every claim carries a citation, Perplexity output is easier to verify downstream than a raw model completion — useful as a research step inside a larger agent pipeline.

Monitor competitors' AI visibility

Run competitor-shaped prompts on a schedule and watch who gets named. A competitor appearing in answers where they did not last month usually means a content or PR push you should know about.

Tips for faster, cheaper runs

  • Raise the page-size field (count, limit, per_page — whichever the mode exposes) before raising Max items. Fewer, bigger calls cost less.
  • Use the cheap reference and autocomplete modes to resolve IDs before spending on the expensive detail modes.
  • Schedule incremental runs with a low Max items rather than one huge sweep; you get fresher data and a smaller bill.
  • Chain Actors with webhooks to push new rows straight into your warehouse.

Requirements

This Actor calls the ScrapeBadger API on your behalf, so it needs your key:

  1. Get one at scrapebadger.com — there is a free tier.
  2. In Settings → Environment variables, add SCRAPEBADGER_API_KEY with your sb_live_… key and tick Secret.

Credits are consumed on your ScrapeBadger account in addition to the Apify event price.

Frequently asked questions

Scraping publicly available data from Perplexity is generally legal, and this Actor only ever reads pages a logged-out visitor could see. What you then do with the data is what matters — read the disclaimer below before collecting anything that could be personal data.

Do I need my own proxies?

No. Proxy rotation, browser fingerprinting and anti-bot handling all happen upstream in the ScrapeBadger API, so there is nothing to configure here.

Do I need a ScrapeBadger account?

Yes. The Actor calls the ScrapeBadger API on your behalf, so it needs your API key in the SCRAPEBADGER_API_KEY environment variable. There is a free tier at scrapebadger.com — see Requirements below.

How many results can I get in one run?

Set Max items to whatever you need. The Actor keeps paging until it hits that number or Perplexity runs out of results, and stops cleanly either way.

How much does one run cost?

Events start at $0.0080 and are charged once per API call, not per row — so a page of 100 results costs the same as a single lookup. Apify adds its standard $0.001 actor start fee per run.

What happens if the run fails halfway through?

Whatever was already scraped stays in the dataset. The run ends with a status message explaining where it stopped instead of throwing your results away.

Can I run this on a schedule or from my own code?

Yes. Use the Schedules tab for recurring runs, or the API tab to start runs and read the dataset from your own application.

Does this use the official API?

No — and that is the point. It reads the answer the public web interface gives a normal visitor, which is what your customers actually see. API answers differ from web answers because the web product runs its own retrieval and grounding.

Will the same prompt give the same answer twice?

No. These models are non-deterministic and their retrieval changes constantly. That is why the useful pattern is to run the same prompt on a schedule and track the trend, rather than reading a single answer as fact.

Is scraping Perplexity reliable?

Anti-bot interstitials are a fact of life on Perplexity. A blocked call is retried four times with exponential backoff against fresh exits, and if a run dies partway it keeps everything already scraped rather than throwing the dataset away. Upstream availability is monitored continuously.

Your feedback and support

Found a bug, missing a field, or need a mode that is not here? Open a ticket on the Issues tab, or email support@scrapebadger.com. The API tab has everything you need to run this Actor programmatically.

Disclaimer

Our Actors are ethical and do not extract any private user data, such as email addresses, gender, or location. They only extract what the user has chosen to share publicly. We therefore believe that our Actors, when used for ethical purposes by Apify users, are safe. However, you should be aware that your results could contain personal data. Personal data is protected by the GDPR in the European Union and by other regulations around the world. You should not scrape personal data unless you have a legitimate reason to do so. If you're unsure whether your reason is legitimate, consult your lawyers.