{"schemaVersion":1,"generatedAt":"2026-07-20T09:39:07.064Z","asOf":"2026-07-20","license":"Compiled aggregate counts (facts) re-served by SeqDesk under each source's terms; CC-BY for UniProt & AlphaFold. See https://seqdesk.org/data for full source list & licensing. Provided as-is, no warranty.","categories":{"sequence-archives":"Sequence archives","structures-proteins":"Structures & proteins","datasets-dois":"Datasets, DOIs & repositories","omics-specialized":"Omics & specialized archives","standards-vocab":"Standards & vocabularies","fair-literature":"FAIR adoption & literature","earth-environment":"Earth & environment","physics-materials":"Physics, space & materials","chemistry-compounds":"Chemistry & compounds","biodiversity":"Biodiversity & specimens","clinical-biomed":"Clinical & biomedical","open-data":"Open data & repositories","metadata-completeness":"Metadata & completeness","sequencing-technology":"Sequencing technology"},"licensing":{"summary":"Figures here are compiled aggregate counts — single totals updated weekly from public archives and cached by SeqDesk. They are facts, not reproductions of the underlying records. Each carries its source and retrieval date and is provided “as is” with no warranty. SeqDesk is independent and not endorsed by any listed organization.","providers":[{"name":"NCBI / U.S. National Library of Medicine","scope":"SRA, GenBank, RefSeq, ClinVar, dbSNP, Taxonomy, NCBI Datasets, NCBI Virus, GEO","license":"US-gov public domain · no use/distribution restrictions","url":"https://www.ncbi.nlm.nih.gov/home/about/policies/"},{"name":"EMBL-EBI","scope":"ENA, BioSamples, BioStudies/ArrayExpress, Europe PMC, MGnify, OLS/ENVO, ENA checklists","license":"EMBL-EBI Terms of Use · CC0-aligned","url":"https://www.ebi.ac.uk/about/terms-of-use/"},{"name":"UniProt Consortium","scope":"UniProtKB, Swiss-Prot, TrEMBL, InterPro, Pfam","license":"CC BY 4.0 (attribution required)","url":"https://creativecommons.org/licenses/by/4.0/"},{"name":"RCSB PDB / wwPDB","scope":"released structures","license":"CC0 1.0","url":"https://creativecommons.org/publicdomain/zero/1.0/"},{"name":"AlphaFold DB (Google DeepMind / EMBL-EBI)","scope":"predicted structures","license":"CC BY 4.0 (attribution required)","url":"https://creativecommons.org/licenses/by/4.0/"},{"name":"OpenAlex (OurResearch)","scope":"works, dataset works, FAIR-paper citations","license":"CC0","url":"https://creativecommons.org/publicdomain/zero/1.0/"},{"name":"DataCite & Crossref","scope":"dataset DOIs, total DOIs","license":"CC0 (metadata)","url":"https://datacite.org/"},{"name":"Zenodo · OSF · Dryad · re3data","scope":"records, projects, datasets, repositories","license":"open terms · counts are facts","url":"https://zenodo.org/"},{"name":"bioRxiv/medRxiv · OBO Foundry · GSC MIxS","scope":"preprints, ontologies, MIxS terms","license":"open / CC","url":"https://www.biorxiv.org/"},{"name":"GBIF · OBIS · iNaturalist","scope":"biodiversity occurrences, datasets, observations","license":"CC0 / CC BY per record · counts are facts","url":"https://www.gbif.org/terms"},{"name":"NCBI PubChem · ClinicalTrials.gov (NLM)","scope":"compounds, substances, bioassays, registered trials & results","license":"US-gov public domain","url":"https://www.ncbi.nlm.nih.gov/home/about/policies/"},{"name":"CDS Strasbourg · ESA Gaia · NASA/IPAC","scope":"SIMBAD, VizieR, Gaia DR3, Exoplanet Archive, EOSDIS CMR","license":"CC BY 4.0 / Gaia licence / US-gov open","url":"https://cds.unistra.fr/"},{"name":"CERN (Open Data · INSPIRE-HEP)","scope":"physics records & open datasets","license":"CC0 / open","url":"https://opendata.cern.ch/"},{"name":"Materials Project · OQMD · NOMAD","scope":"computational materials (OPTIMADE)","license":"CC BY 4.0","url":"https://materialsproject.org/about/terms"},{"name":"PANGAEA · ESGF (WCRP CMIP6)","scope":"Earth & climate datasets","license":"CC BY (per dataset)","url":"https://www.pangaea.de/"},{"name":"EMBL-EBI ChEMBL · ChEBI · GWAS Catalog · ENCODE · NeuroMorpho.Org","scope":"chemistry, ontologies, associations, experiments, neuron morphologies","license":"CC BY / open","url":"https://www.ebi.ac.uk/about/terms-of-use/"},{"name":"Harvard Dataverse · figshare · data.europa.eu · World Bank","scope":"cross-domain datasets & development indicators","license":"open terms · counts are facts","url":"https://dataverse.harvard.edu/"},{"name":"NHGRI (National Human Genome Research Institute)","scope":"DNA sequencing cost data (cost per genome, cost per Mb)","license":"U.S. Government work / public domain (cite NHGRI)","url":"https://www.genome.gov/about-genomics/fact-sheets/DNA-Sequencing-Costs-Data"}]},"count":1,"metrics":[{"id":"cost-per-genome","label":"Cost to sequence a human genome","category":"sequencing-technology","source":"NHGRI — DNA Sequencing Costs (Genome Sequencing Program)","unit":"USD per genome","tier":"secondary","flagship":false,"cadence":"release","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"headline":{"value":"$525","unit":"per genome · 2022"},"forecast":false,"series":[["2001-09-30",95263072],["2002-03-31",70175437],["2003-03-31",53751684],["2004-01-31",28780376],["2005-01-31",17534970],["2006-01-31",12585659],["2007-01-31",9408739],["2008-01-31",3063820],["2009-01-31",232735],["2010-01-31",46774],["2011-01-31",20963],["2012-01-31",7666],["2013-01-31",5671],["2014-01-31",4008],["2015-01-31",3970],["2016-05-31",1176],["2017-02-28",1015],["2018-02-28",1232],["2019-02-28",993],["2020-02-28",645],["2021-02-28",851],["2022-05-31",525]],"note":"The most-cited chart in genomics: NHGRI's cost to generate one high-quality human whole genome (~30x), from ~$95M in Sept 2001 to ~$525 by May 2022 — a >180,000-fold drop. The famous cliff from 2008 is where second-generation sequencing made the cost fall far faster than computing's Moore's Law. NHGRI updates this dataset on its own (irregular) schedule; the curve is intentionally NOT forecast — cost has roughly plateaued since the $1,000-genome era and straight-line extrapolation would mislead.","sources":[{"label":"NHGRI — DNA Sequencing Costs: Data from the NHGRI Genome Sequencing Program (GSP)","url":"https://www.genome.gov/about-genomics/fact-sheets/DNA-Sequencing-Costs-Data"},{"label":"NHGRI cost data table (May 2022 release, .xls)","url":"https://www.genome.gov/sites/default/files/media/files/2023-05/Sequencing_Cost_Data_Table_May2022.xls"}],"events":[{"date":"2005-09-01","label":"454 pyrosequencing","detail":"Margulies et al. (Nature 437:376) published the 454 picolitre-plate pyrosequencer, the first next-generation platform — roughly 100x the throughput of capillary Sanger instruments."},{"date":"2007-01-01","label":"Illumina/Solexa SBS","detail":"Solexa's sequencing-by-synthesis Genome Analyzer (Illumina acquired Solexa in 2007) brought gigabase-scale short reads, the workhorse chemistry behind most of this cost decline."},{"date":"2008-01-01","label":"Cost breaks Moore's Law","detail":"From 2008 the cost per genome fell far faster than computing's Moore's Law as second-generation sequencing scaled — the steep drop from ~$9.4M (Jan 2007) to ~$233k (Jan 2009)."},{"date":"2014-01-14","label":"$1,000 genome","detail":"Illumina launched the HiSeq X Ten on 14 Jan 2014, the first platform marketed as breaking the $1,000-per-genome barrier for population-scale whole-genome sequencing."},{"date":"2017-01-09","label":"NovaSeq 6000","detail":"Illumina introduced the NovaSeq series in January 2017, pushing per-genome cost toward a few hundred dollars at production scale."},{"date":"2022-09-29","label":"NovaSeq X / sub-$200","detail":"Illumina announced the NovaSeq X Series on 29 Sep 2022, claiming a ~$200 genome at maximum scale — beyond the end of the NHGRI series shown here."}]}]}