HGNC Gene Symbols Scraper avatar

HGNC Gene Symbols Scraper

Pricing

from $15.00 / 1,000 result items

Go to Apify Store
HGNC Gene Symbols Scraper

HGNC Gene Symbols Scraper

Query the HUGO Gene Nomenclature Committee database for approved human gene symbols, names, aliases, chromosomal location, gene family, RefSeq, Ensembl, OMIM, UniProt, and external links. Export to JSON, CSV, or Excel for bioinformatics, genomics research, and pharmaceutical pipelines.

Pricing

from $15.00 / 1,000 result items

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

8 days ago

Last modified

Share

ParseForge Banner

๐Ÿงฌ HGNC Gene Symbols Scraper

๐Ÿš€ Export approved human gene symbols in seconds. Pull 43,000+ HGNC-approved gene records with cross-references to Ensembl, Entrez, UniProt, OMIM, and PubMed. No API key, no registration, no manual nomenclature lookups.

The HGNC Gene Symbols Scraper exports records from the HUGO Gene Nomenclature Committee, the official authority for assigning unique human gene symbols and names. Each record carries 27 fields including approved symbol, full name, chromosomal location, aliases, previous symbols, gene group, status, and cross-references to Ensembl, Entrez, UCSC, RefSeq, UniProt, OMIM, PubMed, MGD, RGD, CCDS, and Vega. HGNC nomenclature underpins virtually every modern human-genetics database and clinical-genomics pipeline.

Coverage spans 43,000+ approved gene symbols plus thousands of pseudogenes, withdrawn symbols, and reserved names. This Actor turns lookup-by-symbol, lookup-by-ID, and search-by-keyword into one-step exports as CSV, Excel, JSON, or XML.

๐ŸŽฏ Target Audience๐Ÿ’ก Primary Use Cases
Bioinformatics teams, clinical-genomics labs, pharma R&D, computational biologists, science writers, EHR vendorsVariant interpretation, gene-panel design, cross-DB joins, symbol normalization, literature mining, omics pipeline annotation

๐Ÿ“‹ What the HGNC Scraper does

Five lookup modes in a single run:

  • ๐Ÿ”ค Symbol lookup. Resolve approved symbols like BRCA1, TP53, EGFR, MYC, AKT1.
  • ๐Ÿ†” HGNC ID lookup. Resolve canonical HGNC IDs like 1100 or HGNC:1100.
  • ๐Ÿ”— Entrez Gene ID lookup. Cross-reference NCBI Entrez IDs back to HGNC records.
  • ๐Ÿงช UniProt accession lookup. Map protein accessions like P38398 to gene records.
  • ๐Ÿ” Free-text search. Query across symbols, names, aliases, and previous names.

Each record includes chromosomal location, locus type and group, alias and previous symbols, gene-family group, status, approval date, last-modified timestamp, and the complete cross-reference panel.

๐Ÿ’ก Why it matters: symbol nomenclature drifts. A gene approved as MLL in 2010 is now KMT2A. Pipelines and clinical reports that miss the update silently lose joins. This Actor returns the canonical, current HGNC record on every lookup so your annotations stay correct.

๐Ÿ“Š Data fields

Each record includes: aliasName, aliasSymbol, ccdsId, dateApprovedReserved, dateModified, ensemblGeneId, entrezId, geneGroup, hgncId, location, locusGroup, locusType, mgdId, name, omimId, prevName, prevSymbol, pubmedId, raw, refseqAccession, rgdId, scrapedAt, status, symbol, ucscId, uniprotIds, vegaId. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

๐Ÿš€ How to use

  1. ๐Ÿ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. ๐ŸŒ Open the Actor. Go to the HGNC Gene Symbols Scraper page on the Apify Store.
  3. ๐ŸŽฏ Set input. Choose a mode, paste your symbols or IDs into the values list, and set maxItems.
  4. ๐Ÿš€ Run it. Click Start and let the Actor resolve every lookup.
  5. ๐Ÿ“ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.

โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for more biomedical and reference-data scrapers.

โš ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by HGNC, HUGO, or EMBL-EBI. All trademarks mentioned are the property of their respective owners. Only publicly available gene nomenclature data is collected.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.