Ensembl Genomics Scraper (Genes, Variants, Sequences)
Pricing
from $18.00 / 1,000 result items
Ensembl Genomics Scraper (Genes, Variants, Sequences)
Query the Ensembl genome reference for 200+ species. Look up genes by symbol or stable ID, list features in a genomic region, fetch DNA sequence, or resolve human variants (rsIDs). Returns biotype, coordinates, transcript IDs, descriptions, and assembly metadata.
Pricing
from $18.00 / 1,000 result items
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
8 days ago
Last modified
Categories
Share

𧬠Ensembl Genomics Scraper
π Export genes, variants, and DNA sequences in seconds. Look up by gene symbol, stable ID, chromosomal region, or human rsID across 20+ species. Returns biotype, coordinates, transcript IDs, sequence, allele frequencies, and assembly metadata.
The Ensembl Genomics Scraper queries the public Ensembl genome reference, the de facto open browser for vertebrate, model-organism, and select non-vertebrate genomes. It returns up to 30 structured fields per record, including stable ID, display name, object type, biotype, species, chromosome, start, end, strand, assembly, description, canonical transcript, source, logic name, molecule type, sequence length and sequence, variant name, variant class, minor allele and frequency, ancestral allele, allele string, most-severe consequence, mappings, evidence, synonyms, mode, query, and the scrape timestamp.
The catalog spans 20+ reference species including human, mouse, rat, zebrafish, fruit fly, roundworm, baker's yeast, thale cress, chicken, pig, cow, dog, cat, horse, sheep, rhesus macaque, chimpanzee, western clawed frog, medaka, and mosquito. This Actor returns gene lookups, region overlaps, sequence fetches, and human variant resolutions in one run.
| π― Target Audience | π‘ Primary Use Cases |
|---|---|
| Bioinformaticians, pharma research, genetics labs, academic researchers, computational biology students, biotech startups, precision-medicine teams | Gene annotation pipelines, variant impact analysis, comparative genomics, target identification, rsID resolution for GWAS, sequence retrieval for primer design |
π What the Ensembl Genomics Scraper does
Five query workflows in a single Actor:
- 𧬠Lookup by gene symbol. Resolve
BRCA2,TP53,EGFR, etc. to Ensembl stable IDs, coordinates, biotype, and canonical transcript. - π Lookup by stable ID. Pass
ENSG00000139618orENST00000380152for any Ensembl-supported species. - πΊοΈ Overlap region. Return all gene features inside
chromosome:start-end(e.g.7:140424943-140624564). - π§ͺ Sequence by ID. Fetch the raw DNA, cDNA, or protein sequence for any Ensembl stable ID.
- 𧬠Variation by rsID. Resolve human dbSNP rsIDs (e.g.
rs56116432,rs1042522) to allele frequencies, consequences, and ancestral alleles.
Each record bundles the relevant Ensembl-native fields, the species, the mode used, the original query string, and a collection timestamp.
π‘ Why it matters: the Ensembl genome browser is the most widely cited open genome reference in life sciences. Hand-coding a REST client means handling rate limits, schema-per-endpoint quirks, and pagination. This Actor delivers consistent records you can pipe straight into BI tools, notebooks, or pipelines.
π Data fields
Each record includes: alleleString, ancestralAllele, assemblyName, biotype, canonicalTranscript, chromosome, description, displayName, end, evidence, logicName, mappings, minorAllele, minorAlleleFreq, mode, molecule, mostSevereConsequence, objectType, query, scrapedAt, sequence, sequenceLength, source, species, stableId, start, strand, synonyms, varClass, variantName. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.
π How to use
- π Sign up. Create a free account with $5 credit (takes 2 minutes).
- π Open the Actor. Go to the Ensembl Genomics Scraper page on the Apify Store.
- π― Set input. Pick a mode, a species, and a query payload (symbols, stable IDs, region, or rsIDs). Set
maxItems. - π Run it. Click Start and let the Actor collect your data.
- π₯ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.
β±οΈ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.
π Recommended Actors
- π§ͺ KEGG Pathways Scraper - Biochemical pathways and orthologies
- π ArXiv Scraper - Pre-print research papers
- π¬ Figshare Scraper - Open scientific datasets and supplementary files
- 𧬠ClinicalTrials.gov Scraper - U.S. clinical trial registry
- π GBIF Biodiversity Scraper - Global biodiversity occurrence records
π‘ Pro Tip: browse the complete ParseForge collection for more open-science scrapers.
β οΈ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by Ensembl, EMBL-EBI, the Wellcome Sanger Institute, or NCBI/dbSNP. All trademarks mentioned are the property of their respective owners. Only publicly available open genome reference data is collected.
π Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.