Global Parliamentary Data Scraper
Pricing
from $0.25 / 1,000 parliamentary records
Global Parliamentary Data Scraper
Collect official parliamentary bills, debates, reports and votes across Europe and the US. Get normalized records with dates and source links, ready for JSON, CSV or Excel export. Congress.gov requires an API key.
Collect normalized political and parliamentary data from official sources in European countries (France, Denmark, Ireland, the Netherlands, Poland, Austria, Estonia, Italy, Lithuania, Portugal, Slovakia, Spain, Czechia, Finland, Sweden, Greece, Croatia, the United Kingdom, and Germany), plus the European Parliament and the United States Congress. The Actor is designed for researchers, journalists, civic-technology teams, policy analysts, monitoring workflows, and AI agents that need consistent legislative records without maintaining separate collectors.
Start with 20 records
Choose France — Assemblée nationale, keep All compatible categories, leave date filters empty, and use metadata mode for a first run. Export the collected records as JSON, CSV or Excel. No Congress API key is needed for this example.
What can this Actor collect?
- French National Assembly sessions, news, legislative dossiers, and reports.
- French Senate reports, legislative texts, dossiers, sessions, and news.
- U.S. Congress bills, resolutions, and amendments.
- UK Parliament bills, debates, publications, and committee material.
- German Bundestag news, debates, documents, and sessions.
- European Parliament resolutions, parliamentary documents, debates, and questions.
- Danish Folketing bills, resolutions, questions, documents, meetings, committees, and votes through the official ODA API.
- Irish Oireachtas bills, debates, parliamentary questions, and votes through its official Open Data API.
- Dutch Tweede Kamer bills, documents, meetings, questions, votes, and committees through its official OData API.
- Polish Sejm bills, sittings, votes, committees, written questions, and interpellations through its official API.
- Bulgarian National Assembly bills from its official parliamentary source.
- Austrian Nationalrat plenary sessions through the official Parliament Open Data endpoint.
- Estonian Riigikogu bills and votes through its official Open Data API.
- Italian parliamentary bills and votes through the Camera and Senato Linked Open Data endpoints.
- Portuguese parliamentary initiatives and votes through the Assembleia da República Open Data files.
- Spanish bills through the Congreso de los Diputados Open Data files.
- Czech bills and written questions through the Chamber of Deputies official document list.
- Finnish bills, committee reports, and written questions through the Eduskunta Open Data API.
- Swedish bills, committee reports, and plenary protocols through the Riksdag Open Data API, with an official website fallback.
- Greek laws through the Hellenic Parliament Open Data API.
- Croatian bills through the official Parliament RSS feed; Lithuanian and Slovak bills through their official parliamentary lists.
The Actor uses official APIs, feeds, and institutional pages. It does not collect social-media posts or election-campaign content.
Key features
- One normalized output format across all supported institutions.
- Inclusive
fromDateandtoDatefilters. - Stable identifiers and URL-based deduplication.
- Balanced results when All compatible categories is selected.
- Transparent
dateStatusanddateSourcefields. - Fast metadata mode for large National Assembly documents.
- Optional full-text downloading.
- Isolated source errors, so one unavailable category does not cancel valid results.
- Results available in the Apify dataset for JSON, CSV, Excel, API, integrations, schedules, and monitoring.
How to use it
- Select a political or parliamentary source, or Toutes les sources (un seul Run) to aggregate every configured source into one Dataset.
- Choose one content category or All compatible categories.
- Enter an optional start and end date in
YYYY-MM-DDformat. - Choose the maximum number of records.
- For U.S. Congress data, provide your Congress.gov API key in the secret input field. Start without date filters for your first trial.
- Click Start, then open the Dataset and choose Export for JSON, CSV or Excel. Read Collection status and source warnings in Output if results are missing.
When the aggregate mode is selected, the Actor launches the configured source runs automatically and combines their datasets. It requires execution on the Apify platform (not local mode).
The aggregate mode launches every source displayed in the input list. Belgium, Slovenia, Cyprus, Malta, and Hungary are intentionally not displayed for now because their official sites either block automated access or do not provide a reproducible public feed.
Example input
{"source": "assemblee","category": "all","maxRequestsPerCrawl": 20,"includeFullText": false}
Real output sample
Inspect a 20-record French National Assembly sample collected on September 4, 2026: JSON sample · CSV sample. The sample includes 4 record types. Downloading these existing files does not start a new paid run. This is a limited snapshot, not a complete archive.
Output fields
Each record can contain country, institution, language, source, title, date, type, url, text, collectedAt, id, dateStatus, and dateSource.
Dates are never invented. If an official source does not provide a reliable date, the Actor returns date: null, dateStatus: missing, and dateSource: null.
Pricing
$0.25 per 1,000 parliamentary records ($0.00025 per Dataset item), plus an Actor-start event of $0.00005 per allocated GB, with a minimum of one event. At the default 512 MB, 20 records cost $0.005 in record events plus $0.00005 for the start event. Check the Pricing tab and run charge limit before starting.
Errors are stored separately in COLLECTION_STATUS, never as a billable parliamentary record. The start event can still apply to a failed run, and any genuine records saved before an interruption remain billable.
All-sources mode has a different cost: it launches up to 23 child runs. The requested record limit applies to each source, not to the combined result. Child starts, child Dataset records and records copied into the parent Dataset are separate chargeable events. The parent run's charge limit does not constitute a combined cap for all child runs. Start with one institution to estimate cost.
FAQ
Why do some records have no date?
Some official pages or APIs do not expose a trustworthy publication date. The Actor marks this transparently instead of guessing.
Why is a Congress.gov API key required?
Congress.gov requires an API key for its official API. The key is stored as a secret input by Apify.
Can I request complete document text?
Enable Download full text where supported by the selected collector. The amount of text available varies by source and document; some results contain metadata or summaries. Larger pages, especially session transcripts, take longer.
Can I automate recurring monitoring?
Yes. Use Apify schedules, API calls, webhooks, or integrations to run the Actor repeatedly and export new datasets.
Responsible use
This Actor collects publicly available information from official institutions. Source availability, page structure, and field completeness can change. Users remain responsible for complying with applicable laws and the terms of the source websites.
Reliability and coverage
Official sources may return fewer records than requested or be temporarily unavailable. A run with no matching records is marked failed with a diagnostic rather than returning an error as a document. Genuine records already collected are retained. In all-sources mode, inspect COLLECTION_STATUS for missing-source warnings; a partial result is not complete coverage. Source categories differ, and unsupported category choices fall back to the institution’s compatible categories.
The output preserves the original source language; it does not translate, verify political claims or determine the legal effect of a document. Date filtering is not a guarantee of exhaustive historical coverage.
Support
Open an issue on this Actor with the run ID, selected institution and expected category. Do not share API keys or private data.