{"schemaVersion":1,"generatedAt":"2026-07-20T09:39:07.064Z","asOf":"2026-07-20","license":"Compiled aggregate counts (facts) re-served by SeqDesk under each source's terms; CC-BY for UniProt & AlphaFold. See https://seqdesk.org/data for full source list & licensing. Provided as-is, no warranty.","categories":{"sequence-archives":"Sequence archives","structures-proteins":"Structures & proteins","datasets-dois":"Datasets, DOIs & repositories","omics-specialized":"Omics & specialized archives","standards-vocab":"Standards & vocabularies","fair-literature":"FAIR adoption & literature","earth-environment":"Earth & environment","physics-materials":"Physics, space & materials","chemistry-compounds":"Chemistry & compounds","biodiversity":"Biodiversity & specimens","clinical-biomed":"Clinical & biomedical","open-data":"Open data & repositories","metadata-completeness":"Metadata & completeness","sequencing-technology":"Sequencing technology"},"licensing":{"summary":"Figures here are compiled aggregate counts — single totals updated weekly from public archives and cached by SeqDesk. They are facts, not reproductions of the underlying records. Each carries its source and retrieval date and is provided “as is” with no warranty. SeqDesk is independent and not endorsed by any listed organization.","providers":[{"name":"NCBI / U.S. National Library of Medicine","scope":"SRA, GenBank, RefSeq, ClinVar, dbSNP, Taxonomy, NCBI Datasets, NCBI Virus, GEO","license":"US-gov public domain · no use/distribution restrictions","url":"https://www.ncbi.nlm.nih.gov/home/about/policies/"},{"name":"EMBL-EBI","scope":"ENA, BioSamples, BioStudies/ArrayExpress, Europe PMC, MGnify, OLS/ENVO, ENA checklists","license":"EMBL-EBI Terms of Use · CC0-aligned","url":"https://www.ebi.ac.uk/about/terms-of-use/"},{"name":"UniProt Consortium","scope":"UniProtKB, Swiss-Prot, TrEMBL, InterPro, Pfam","license":"CC BY 4.0 (attribution required)","url":"https://creativecommons.org/licenses/by/4.0/"},{"name":"RCSB PDB / wwPDB","scope":"released structures","license":"CC0 1.0","url":"https://creativecommons.org/publicdomain/zero/1.0/"},{"name":"AlphaFold DB (Google DeepMind / EMBL-EBI)","scope":"predicted structures","license":"CC BY 4.0 (attribution required)","url":"https://creativecommons.org/licenses/by/4.0/"},{"name":"OpenAlex (OurResearch)","scope":"works, dataset works, FAIR-paper citations","license":"CC0","url":"https://creativecommons.org/publicdomain/zero/1.0/"},{"name":"DataCite & Crossref","scope":"dataset DOIs, total DOIs","license":"CC0 (metadata)","url":"https://datacite.org/"},{"name":"Zenodo · OSF · Dryad · re3data","scope":"records, projects, datasets, repositories","license":"open terms · counts are facts","url":"https://zenodo.org/"},{"name":"bioRxiv/medRxiv · OBO Foundry · GSC MIxS","scope":"preprints, ontologies, MIxS terms","license":"open / CC","url":"https://www.biorxiv.org/"},{"name":"GBIF · OBIS · iNaturalist","scope":"biodiversity occurrences, datasets, observations","license":"CC0 / CC BY per record · counts are facts","url":"https://www.gbif.org/terms"},{"name":"NCBI PubChem · ClinicalTrials.gov (NLM)","scope":"compounds, substances, bioassays, registered trials & results","license":"US-gov public domain","url":"https://www.ncbi.nlm.nih.gov/home/about/policies/"},{"name":"CDS Strasbourg · ESA Gaia · NASA/IPAC","scope":"SIMBAD, VizieR, Gaia DR3, Exoplanet Archive, EOSDIS CMR","license":"CC BY 4.0 / Gaia licence / US-gov open","url":"https://cds.unistra.fr/"},{"name":"CERN (Open Data · INSPIRE-HEP)","scope":"physics records & open datasets","license":"CC0 / open","url":"https://opendata.cern.ch/"},{"name":"Materials Project · OQMD · NOMAD","scope":"computational materials (OPTIMADE)","license":"CC BY 4.0","url":"https://materialsproject.org/about/terms"},{"name":"PANGAEA · ESGF (WCRP CMIP6)","scope":"Earth & climate datasets","license":"CC BY (per dataset)","url":"https://www.pangaea.de/"},{"name":"EMBL-EBI ChEMBL · ChEBI · GWAS Catalog · ENCODE · NeuroMorpho.Org","scope":"chemistry, ontologies, associations, experiments, neuron morphologies","license":"CC BY / open","url":"https://www.ebi.ac.uk/about/terms-of-use/"},{"name":"Harvard Dataverse · figshare · data.europa.eu · World Bank","scope":"cross-domain datasets & development indicators","license":"open terms · counts are facts","url":"https://dataverse.harvard.edu/"},{"name":"NHGRI (National Human Genome Research Institute)","scope":"DNA sequencing cost data (cost per genome, cost per Mb)","license":"U.S. Government work / public domain (cite NHGRI)","url":"https://www.genome.gov/about-genomics/fact-sheets/DNA-Sequencing-Costs-Data"}]},"count":96,"metrics":[{"id":"ena-read-run","label":"ENA raw read datasets","category":"sequence-archives","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2022-01-01",22000000],["2023-01-01",28000000],["2024-01-01",32000000],["2025-01-01",37000000],["2026-06-24",42613406],["2026-07-06",42834140],["2026-07-13",42924583],["2026-07-20",43021658]],"note":"Each record is one submitted sequencing run. Cumulative; only goes up.","events":[{"date":"2007-01-01","label":"SRA launched","detail":"NCBI established the Sequence Read Archive in 2007 as the first public repository for raw next-generation sequencing reads, the foundation later mirrored by ENA/EBI."},{"date":"2007-01-01","label":"Illumina NGS arrives","detail":"Solexa's Genome Analyzer (acquired by Illumina in 2007) brought gigabase-scale short-read sequencing to market, triggering the explosive growth in submitted read data."},{"date":"2008-01-01","label":"1000 Genomes Project","detail":"Launched in January 2008, this international effort to catalogue human variation drove a large early wave of high-coverage read submissions to public archives."},{"date":"2015-05-01","label":"Nanopore MinION","detail":"Oxford Nanopore's portable MinION became commercially available in May 2015, adding real-time long-read sequencing to the data deposited in ENA/SRA."},{"date":"2020-01-10","label":"SARS-CoV-2 genome","detail":"The first SARS-CoV-2 genome (Wuhan-Hu-1) was shared publicly on 10-11 January 2020, kicking off the largest pathogen-sequencing surge ever recorded in the archives."}]},{"id":"ena-sequence","label":"ENA assembled sequence records","category":"sequence-archives","source":"EMBL-EBI / ENA","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=sequence&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2010-01-01",111807831],["2011-01-01",128262962],["2012-01-01",144561929],["2013-01-01",158773513],["2014-01-01",190402903],["2015-01-01",203127832],["2016-01-01",215563691],["2017-01-01",226705977],["2018-01-01",234237431],["2019-01-01",245341344],["2020-01-01",250527682],["2021-01-01",264554470],["2022-01-01",285459049],["2023-01-01",291371095],["2024-01-01",298617527],["2025-01-01",303369429],["2026-01-01",308730625],["2026-06-24",312995979],["2026-07-06",313089857],["2026-07-13",313146987],["2026-07-20",313146987]]},{"id":"sra-records","label":"NCBI SRA records","category":"sequence-archives","source":"NCBI / Sequence Read Archive","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/einfo.fcgi?db=sra&version=2.0&retmode=json&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"einforesult.dbinfo.0.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",45092952],["2026-07-06",45221246],["2026-07-13",45416039],["2026-07-20",45542529]]},{"id":"genbank-nuccore","label":"GenBank nucleotide records","category":"sequence-archives","source":"NCBI / GenBank (nuccore)","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=nuccore&term=all%5Bsb%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2002-01-01",14981422],["2003-01-01",22359913],["2004-01-01",31291052],["2005-01-01",40858172],["2006-01-01",53227254],["2007-01-01",66171909],["2008-01-01",86147450],["2009-01-01",105144229],["2010-01-01",120540797],["2011-01-01",137515957],["2012-01-01",157162863],["2013-01-01",179092467],["2014-01-01",213912765],["2015-01-01",263467835],["2016-01-01",292580247],["2017-01-01",321891083],["2018-01-01",350474188],["2019-01-01",376446019],["2020-01-01",399641060],["2021-01-01",427755026],["2022-01-01",483121662],["2023-01-01",579251081],["2024-01-01",619046223],["2025-01-01",650469847],["2026-01-01",687532105],["2026-06-24",731510357],["2026-07-06",731882147],["2026-07-13",732140069],["2026-07-20",735368763]]},{"id":"ncbi-assemblies","label":"NCBI genome assemblies","category":"sequence-archives","source":"NCBI Datasets","unit":"assemblies","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.ncbi.nlm.nih.gov/datasets/v2/genome/taxon/1/dataset_report?page_size=1","method":"GET","parse":{"type":"json","path":"total_count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2009-12-31",14421],["2010-12-31",19769],["2011-12-31",26429],["2012-12-31",36218],["2013-12-31",53990],["2014-12-31",77202],["2015-12-31",114446],["2016-12-31",158180],["2017-12-31",216754],["2018-12-31",298005],["2019-12-31",638809],["2020-12-31",1013201],["2021-04-18",1000000],["2021-12-31",1289876],["2022-12-31",1658238],["2023-12-31",2204898],["2024-12-31",2814679],["2025-12-31",3351364],["2026-06-24",4183542],["2026-07-06",4220941],["2026-07-13",4248468],["2026-07-20",4266543]],"sources":[{"label":"NCBI Assembly growth (per-year assembly counts)","url":"https://www.ncbi.nlm.nih.gov/assembly"}]},{"id":"refseq-nuccore","label":"RefSeq curated nucleotide records","category":"sequence-archives","source":"NCBI / RefSeq","unit":"records","tier":"secondary","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=nuccore&term=srcdb_refseq%5Bprop%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2005-01-01",162780],["2006-01-01",321638],["2007-01-01",429360],["2008-01-01",720127],["2009-01-01",1020607],["2010-01-01",1268879],["2011-01-01",1680515],["2012-01-01",2097674],["2013-01-01",2616977],["2014-01-01",4943175],["2015-01-01",11319336],["2016-01-01",17496304],["2017-01-01",24224850],["2018-01-01",30969736],["2019-01-01",38899623],["2020-01-01",47456133],["2021-01-01",57966926],["2022-01-01",70052014],["2023-01-01",82264971],["2024-01-01",99217809],["2025-01-01",117135672],["2026-01-01",134045664],["2026-06-24",146863552],["2026-07-06",147043342],["2026-07-13",147252697],["2026-07-20",147690531]]},{"id":"mgnify-analyses","label":"MGnify metagenomic analyses","category":"sequence-archives","source":"EMBL-EBI / MGnify","unit":"analyses","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/metagenomics/api/v1/analyses?page=1","method":"GET","parse":{"type":"json","path":"meta.pagination.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",635107],["2026-07-06",635107],["2026-07-13",635107],["2026-07-20",635107]]},{"id":"genbank-wgs-bases","label":"GenBank WGS sequence","category":"sequence-archives","source":"NCBI / GenBank release notes","unit":"bases","tier":"flagship","flagship":true,"cadence":"release","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"series":[["2002-01-01",6702372564],["2004-01-01",35009256228],["2006-01-01",81611376856],["2008-01-01",141374971004],["2010-01-01",177385297156],["2012-01-01",356002922838],["2014-01-01",848977922022],["2016-01-01",1817189565845],["2018-01-01",3656719423096],["2020-01-01",11830842428018],["2022-01-01",19086596616569],["2024-01-01",32983029087303]],"note":"Whole-genome sequence in bases — the GenBank WGS division only (not the larger set-based WGS/TSA/TLS header total in the release notes). Each YYYY-01-01 point is that calendar year's December release, not literally Jan 1: e.g. 2024-01-01 = 32,983,029,087,303 is GenBank Release 264.0 (19 Dec 2024) and 2022-01-01 = 19,086,596,616,569 is Release 253.0 (15 Dec 2022). Spacing is a uniform 2 years, so the trend is intact; only the labels are year-end. Not exposed as a live API — update manually from https://www.ncbi.nlm.nih.gov/genbank/statistics/.","events":[{"date":"1982-01-01","label":"GenBank founded","detail":"The Los Alamos sequence library won a five-year NIGMS grant in 1982 and was christened GenBank, establishing the public nucleotide database whose WGS holdings this curve tracks."},{"date":"2005-09-01","label":"454 NGS arrives","detail":"Margulies et al. published the picolitre-reactor 454 pyrosequencing system in Nature (437:376-380), launching next-generation sequencing with ~100x the throughput of capillary instruments."},{"date":"2006-01-01","label":"Illumina/Solexa GA","detail":"Solexa launched the Genome Analyzer in 2006, bringing sequencing-by-synthesis short reads (1 Gb per run) that would soon dominate sequence submissions."},{"date":"2008-01-22","label":"1000 Genomes Proj.","detail":"An international consortium announced the 1000 Genomes Project on 22 Jan 2008, a population-scale effort that poured large volumes of human sequence into public databases."},{"date":"2014-01-14","label":"$1000 genome","detail":"Illumina introduced the HiSeq X Ten on 14 Jan 2014, the first platform to break the $1,000-per-genome barrier and slash the cost of large-scale sequencing."},{"date":"2020-01-11","label":"SARS-CoV-2 genome","detail":"The first SARS-CoV-2 genome (Wuhan-Hu-1) was made public via virological.org on 11 Jan 2020, igniting an unprecedented global viral-genome sequencing surge."}]},{"id":"human-genomes","label":"Human genomes sequenced","category":"sequence-archives","source":"SeqDesk estimate · Berkeley Genomics, UK Biobank, gnomAD, Stephens 2015","unit":"genomes (WGS)","tier":"flagship","flagship":true,"cadence":"curated","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"series":[["2012-01-01",1092],["2015-01-01",250000],["2020-01-01",500000],["2025-01-01",2000000]],"note":"SeqDesk-curated LOWER BOUND of cumulative human whole-genome (WGS, ~30x) sequencing worldwide: about 2 million by 2025 (Berkeley Genomics, deliberately conservative); the true figure is likely 2-5 million by 2026 and ultimately unknowable because most clinical/biobank genomes never reach public archives. WGS only — excludes millions of exomes (WES) and ~50M consumer SNP arrays (not sequencing). Component cohorts (UK Biobank ~491k, All of Us ~415k, gnomAD, Genomics England 100k) must NOT be summed (severe double-counting). Context: Stephens et al. 2015 projected 100M-2B human genomes by 2025 and Birney 2017 projected >60M clinical genomes — both overshot the realized ~2M by 50-1000x, a caution against straight-line extrapolation of any curve on this page.","sources":[{"label":"Berkeley Genomics — “How many human genomes have been sequenced?” (~2M lower bound)","url":"https://berkeleygenomics.org/articles/How_many_human_genomes_have_been_sequenced.html"},{"label":"Stephens et al. 2015, PLOS Biology — Big Data: Astronomical or Genomical? (100M-2B by 2025 projection)","url":"https://doi.org/10.1371/journal.pbio.1002195"},{"label":"UK Biobank WGS, Nature 2025 — 490,640 genomes at 32.5x","url":"https://www.nature.com/articles/s41586-024-08344-6"},{"label":"gnomAD v4 — 76,215 WGS + 730,947 WES","url":"https://gnomad.broadinstitute.org/news/2023-11-gnomad-v4-0/"},{"label":"1000 Genomes Project Phase 1 — 1,092 genomes","url":"https://www.nature.com/articles/nature11632"}],"forecast":false,"events":[{"date":"2001-02-12","label":"HGP draft genome","detail":"The Human Genome Project published the first working draft of the human genome, announced February 12, 2001 in Nature and Science."},{"date":"2003-04-14","label":"HGP complete","detail":"The Human Genome Project declared the essentially complete reference human genome finished on April 14, 2003, two years ahead of schedule."},{"date":"2012-11-01","label":"1000 Genomes","detail":"The 1000 Genomes Project published an integrated map from 1,092 human genomes, vastly expanding the catalog of human genetic variation across populations."},{"date":"2018-12-05","label":"100k Genomes UK","detail":"Genomics England reached its goal of sequencing 100,000 whole genomes from NHS patients, marking the first national-scale clinical genomics program."},{"date":"2022-03-31","label":"T2T-CHM13","detail":"The Telomere-to-Telomere Consortium published the first truly complete, gapless human genome (T2T-CHM13), adding ~200 million bases over the prior reference."},{"date":"2023-11-30","label":"UK Biobank 500k","detail":"UK Biobank released whole-genome sequences for all 500,000 participants, then the world's largest single set of human sequencing data."}]},{"id":"pdb-structures","label":"RCSB PDB released structures","category":"structures-proteins","source":"RCSB Protein Data Bank","unit":"structures","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://data.rcsb.org/rest/v1/holdings/current/entry_ids","method":"GET","parse":{"type":"array-length"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["1976-01-01",13],["1977-01-01",36],["1978-01-01",42],["1979-01-01",53],["1980-01-01",69],["1981-01-01",85],["1982-01-01",117],["1983-01-01",153],["1984-01-01",175],["1985-01-01",195],["1986-01-01",213],["1987-01-01",238],["1988-01-01",291],["1989-01-01",365],["1990-01-01",507],["1991-01-01",694],["1992-01-01",886],["1993-01-01",1582],["1994-01-01",2871],["1995-01-01",3812],["1996-01-01",4984],["1997-01-01",6548],["1998-01-01",8603],["1999-01-01",10959],["2000-01-01",13583],["2001-01-01",16393],["2002-01-01",19387],["2003-01-01",23531],["2004-01-01",28680],["2005-01-01",34014],["2006-01-01",40419],["2007-01-01",47553],["2008-01-01",54456],["2009-01-01",61744],["2010-01-01",69486],["2011-01-01",77411],["2012-01-01",86161],["2013-01-01",95499],["2014-01-01",105050],["2015-01-01",114292],["2016-01-01",125095],["2017-01-01",136156],["2018-01-01",147325],["2019-01-01",158804],["2020-01-01",172809],["2021-01-01",185393],["2022-01-01",199677],["2023-01-01",214175],["2024-01-01",229635],["2025-01-01",247251],["2026-01-01",256006],["2026-06-24",256006],["2026-07-06",256292],["2026-07-13",256550],["2026-07-20",256840]],"note":"RCSB search API needs a URL-encoded JSON query (return_counts:true -> total_count). Wire a custom fetcher.","sources":[{"label":"RCSB PDB — Overall Growth of Released Structures Per Year","url":"https://www.rcsb.org/stats/growth/growth-released-structures"}],"events":[{"date":"1971-10-01","label":"PDB founded","detail":"The Protein Data Bank launched at Brookhaven National Laboratory as a public archive for macromolecular structures, beginning with just seven X-ray crystallography entries."},{"date":"1994-01-01","label":"First CASP","detail":"The biennial Critical Assessment of Structure Prediction (CASP) experiment began, driving sustained demand for and deposition of experimentally solved structures used as blind-test targets."},{"date":"2000-06-01","label":"Protein Structure Init.","detail":"The NIH/NIGMS Protein Structure Initiative began a ~$764M structural genomics push that fed thousands of new structures into the PDB over the following decade."},{"date":"2013-01-01","label":"Cryo-EM revolution","detail":"Direct electron detectors triggered cryo-EM's resolution revolution, enabling near-atomic structures and a surge of cryo-EM depositions into the PDB."},{"date":"2020-11-30","label":"AlphaFold2 CASP14","detail":"DeepMind's AlphaFold2 won CASP14 with near-experimental accuracy, widely hailed as solving the 50-year protein-folding problem."},{"date":"2022-07-28","label":"AlphaFold DB 200M","detail":"DeepMind and EMBL-EBI expanded the AlphaFold DB to over 200 million predicted structures, covering nearly every cataloged protein and dwarfing the experimental PDB."}]},{"id":"alphafold","label":"AlphaFold DB predicted structures","category":"structures-proteins","source":"DeepMind / EMBL-EBI","unit":"structures","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ebisearch/ws/rest/alphafold?query=domain_source:alphafold&size=0&format=json","method":"GET","parse":{"type":"json","path":"hitCount","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",214683829]],"events":[{"date":"2020-11-30","label":"AlphaFold2 wins CASP14","detail":"DeepMind's AlphaFold2 won the CASP14 assessment with a median GDT score of 92.4, widely hailed as a solution to the 50-year protein-folding problem."},{"date":"2021-07-15","label":"AF2 code + Nature paper","detail":"DeepMind published the AlphaFold2 method in Nature and open-sourced the full code and model weights on GitHub, enabling widespread adoption."},{"date":"2021-07-22","label":"AlphaFold DB launch","detail":"DeepMind and EMBL-EBI launched the AlphaFold Protein Structure Database with ~360,000 predictions, including the entire human proteome."},{"date":"2022-07-28","label":"DB hits ~214M structures","detail":"The database expanded roughly 200-fold to over 214 million predicted structures, covering nearly every catalogued protein across ~1 million species."},{"date":"2024-10-09","label":"Nobel Prize Chemistry","detail":"Demis Hassabis and John Jumper shared the 2024 Nobel Prize in Chemistry for AlphaFold's protein-structure prediction, cementing its scientific impact."}]},{"id":"uniprot-total","label":"UniProtKB total entries","category":"structures-proteins","source":"UniProt / EMBL-EBI","unit":"records","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://rest.uniprot.org/uniprotkb/search?query=*&size=0&format=list","method":"GET","parse":{"type":"header","header":"x-total-results","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2011-01-11",13069501],["2012-01-25",19968488],["2013-01-09",29805788],["2014-01-22",52159208],["2015-01-07",89998523],["2016-01-20",60268458],["2017-01-18",74265355],["2018-01-31",108184003],["2019-01-16",140253338],["2020-02-26",178316438],["2021-02-10",208365010],["2022-02-23",230895644],["2023-03-01",246440937],["2024-01-24",250322721],["2025-02-05",253206171],["2026-06-10",149810139],["2026-06-24",149810139]],"sources":[{"label":"UniProt release statistics (per-release entry counts)","url":"https://ftp.uniprot.org/pub/databases/uniprot/current_release/knowledgebase/UniProtKB_release-statistics.html"}],"events":[{"date":"1986-07-01","label":"Swiss-Prot launched","detail":"Amos Bairoch released the first Swiss-Prot, a manually curated protein sequence database of ~4,000 entries, seeding the lineage that becomes UniProt."},{"date":"1996-11-01","label":"TrEMBL added","detail":"TrEMBL release 1 added 86,033 computer-annotated entries translated from EMBL/GenBank/DDBJ coding sequences, letting the database absorb new sequences far faster than manual curation."},{"date":"2002-01-01","label":"UniProt consortium","detail":"EBI, SIB, and PIR united Swiss-Prot, TrEMBL, and PIR-PSD into the UniProt consortium, consolidating protein data into one authoritative resource."},{"date":"2014-04-01","label":"Proteome flood","detail":"High-throughput genome sequencing drove UniProtKB toward ~90 million highly redundant sequences (e.g. 4,080 S. aureus proteomes), a sharp spike later trimmed by the reference-proteomes program."},{"date":"2021-07-22","label":"AlphaFold DB launch","detail":"DeepMind and EMBL-EBI launched the AlphaFold Protein Structure Database with over 365,000 predicted structures for the human proteome and 20 model organisms drawn from UniProt."},{"date":"2022-07-01","label":"AlphaFold 214M","detail":"AlphaFold DB expanded to cover over 214 million sequences spanning nearly all of UniProt, tying structural predictions to the bulk of the sequence database."}]},{"id":"swissprot","label":"UniProtKB/Swiss-Prot reviewed","category":"structures-proteins","source":"UniProt / EMBL-EBI","unit":"records","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://rest.uniprot.org/uniprotkb/search?query=reviewed:true&size=0&format=list","method":"GET","parse":{"type":"header","header":"x-total-results","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",575503]]},{"id":"interpro","label":"InterPro entries","category":"structures-proteins","source":"InterPro / EMBL-EBI","unit":"records","tier":"slow-moving","flagship":false,"cadence":"quarterly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/interpro/api/entry/interpro?page_size=1","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",54190]]},{"id":"pfam","label":"Pfam families","category":"structures-proteins","source":"Pfam / EMBL-EBI","unit":"families","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/interpro/api/entry/pfam?page_size=1","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",30134]]},{"id":"datacite-datasets","label":"DataCite dataset DOIs","category":"datasets-dois","source":"DataCite","unit":"DOIs","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.datacite.org/dois?resource-type-id=dataset&page%5Bsize%5D=0","method":"GET","parse":{"type":"json","path":"meta.total","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2012-01-01",286477],["2013-01-01",417258],["2014-01-01",516873],["2015-01-01",1159210],["2016-01-01",2089542],["2017-01-01",2916663],["2018-01-01",3691786],["2019-01-01",5455200],["2020-01-01",6553166],["2021-01-01",8134417],["2022-01-01",10813210],["2023-01-01",12986154],["2024-01-01",15706400],["2025-01-01",29627024],["2026-01-01",56486756],["2026-06-24",70797921],["2026-07-06",70925036],["2026-07-13",70971627],["2026-07-20",71791555]]},{"id":"datacite-total","label":"DataCite total DOIs","category":"datasets-dois","source":"DataCite","unit":"DOIs","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.datacite.org/dois?page%5Bsize%5D=0","method":"GET","parse":{"type":"json","path":"meta.total","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2012-01-01",787182],["2013-01-01",1242763],["2014-01-01",1704084],["2015-01-01",3108854],["2016-01-01",4792999],["2017-01-01",6937147],["2018-01-01",9787150],["2019-01-01",13307796],["2020-01-01",16982043],["2021-01-01",20794986],["2022-01-01",26334475],["2023-01-01",33993360],["2024-01-01",52207964],["2025-01-01",71696049],["2026-01-01",108542278],["2026-06-24",129785612],["2026-07-06",130419140],["2026-07-13",130611985],["2026-07-20",131571128]]},{"id":"zenodo","label":"Zenodo records","category":"datasets-dois","source":"CERN / Zenodo","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://zenodo.org/api/records?size=1","method":"GET","parse":{"type":"json","path":"hits.total","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2015-01-01",4236],["2016-01-01",17755],["2017-01-01",110105],["2018-01-01",311506],["2019-01-01",969630],["2020-01-01",1402412],["2021-01-01",1657147],["2022-01-01",2215494],["2023-01-01",2758043],["2024-01-01",3276956],["2025-01-01",4314863],["2026-01-01",5683774],["2026-06-24",6704706],["2026-07-06",6864540],["2026-07-13",6907886],["2026-07-20",6950103]]},{"id":"osf-nodes","label":"OSF public projects","category":"datasets-dois","source":"COS / OSF","unit":"projects","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.osf.io/v2/nodes/?page%5Bsize%5D=1","method":"GET","parse":{"type":"json","path":"links.meta.total","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2014-01-01",1769],["2015-01-01",4070],["2016-01-01",8253],["2017-01-01",19413],["2018-01-01",41615],["2019-01-01",79111],["2020-01-01",121952],["2021-01-01",181393],["2022-01-01",240759],["2023-01-01",294035],["2024-01-01",346961],["2025-01-01",501418],["2026-01-01",573472],["2026-06-24",618482],["2026-07-06",622790],["2026-07-13",625137],["2026-07-20",627245]]},{"id":"osf-registrations","label":"OSF registrations","category":"datasets-dois","source":"COS / OSF","unit":"registrations","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.osf.io/v2/registrations/?page%5Bsize%5D=1","method":"GET","parse":{"type":"json","path":"links.meta.total","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2014-01-01",1064],["2015-01-01",2048],["2016-01-01",5131],["2017-01-01",10680],["2018-01-01",18792],["2019-01-01",32277],["2020-01-01",48955],["2021-01-01",74996],["2022-01-01",102852],["2023-01-01",132854],["2024-01-01",165963],["2025-01-01",203434],["2026-01-01",250324],["2026-06-24",276165],["2026-07-06",279574],["2026-07-13",280944],["2026-07-20",282233]]},{"id":"dryad","label":"Dryad curated datasets","category":"datasets-dois","source":"Dryad","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://datadryad.org/api/v2/datasets?per_page=1","method":"GET","parse":{"type":"json","path":"total","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",71096],["2026-07-06",71255],["2026-07-13",71383],["2026-07-20",71487]]},{"id":"re3data","label":"re3data registered repositories","category":"datasets-dois","source":"re3data.org","unit":"repositories","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://www.re3data.org/api/v1/repositories","method":"GET","parse":{"type":"xml-tag-count","tag":"repository"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",3505],["2026-07-06",3507],["2026-07-13",3510],["2026-07-20",3515]],"note":"XML list — count <repository> elements (~1MB). Wire a custom fetcher; poll monthly."},{"id":"biosamples","label":"EBI BioSamples total","category":"omics-specialized","source":"EMBL-EBI / BioSamples","unit":"samples","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/biosamples/samples?size=1&page=0","method":"GET","parse":{"type":"json","path":"page.totalElements","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2014-01-01",2000000],["2018-01-01",5000000],["2021-01-01",18000000],["2026-06-24",53656082],["2026-07-06",53936004],["2026-07-13",54077762],["2026-07-20",54180273]],"note":"Sample-metadata backbone. Historical points from BioSamples NAR updates; latest is the live API count."},{"id":"ena-samples","label":"ENA samples (BioSamples mirror)","category":"omics-specialized","source":"EMBL-EBI / ENA","unit":"samples","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=sample&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2010-01-01",84098],["2011-01-01",238252],["2012-01-01",476313],["2013-01-01",942985],["2014-01-01",1377336],["2015-01-01",1921110],["2016-01-01",2852158],["2017-01-01",4029178],["2018-01-01",5651976],["2019-01-01",7568506],["2020-01-01",10162219],["2021-01-01",13355311],["2022-01-01",20162054],["2023-01-01",27240621],["2024-01-01",33823122],["2025-01-01",38863188],["2026-01-01",43816447],["2026-06-24",53402316],["2026-07-06",53655189],["2026-07-13",53790249],["2026-07-20",53934213]]},{"id":"geo-series","label":"NCBI GEO series (GSE)","category":"omics-specialized","source":"NCBI / GEO","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=gds&term=gse%5Bentry+type%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2005-01-01",1475],["2006-01-01",2869],["2007-01-01",4682],["2008-01-01",7393],["2009-01-01",10671],["2010-01-01",14983],["2011-01-01",20551],["2012-01-01",27341],["2013-01-01",34937],["2014-01-01",44095],["2015-01-01",53616],["2016-01-01",64070],["2017-01-01",77664],["2018-01-01",93026],["2019-01-01",106799],["2020-01-01",122734],["2021-01-01",141659],["2022-01-01",167145],["2023-01-01",191004],["2024-01-01",215730],["2025-01-01",243241],["2026-01-01",270500],["2026-06-24",287409],["2026-07-06",288653],["2026-07-13",289126],["2026-07-20",289649]]},{"id":"geo-samples","label":"NCBI GEO samples (GSM)","category":"omics-specialized","source":"NCBI / GEO","unit":"samples","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=gds&term=gsm%5Bentry+type%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2005-01-01",28246],["2006-01-01",63630],["2007-01-01",109408],["2008-01-01",187376],["2009-01-01",272277],["2010-01-01",383166],["2011-01-01",509292],["2012-01-01",672896],["2013-01-01",851671],["2014-01-01",1054392],["2015-01-01",1306436],["2016-01-01",1571528],["2017-01-01",1919542],["2018-01-01",2313107],["2019-01-01",2814554],["2020-01-01",3357000],["2021-01-01",4110457],["2022-01-01",4810445],["2023-01-01",5452713],["2024-01-01",6940444],["2025-01-01",7563792],["2026-01-01",8216848],["2026-06-24",8561384],["2026-07-06",8592462],["2026-07-13",8607925],["2026-07-20",8618271]]},{"id":"arrayexpress","label":"ArrayExpress studies","category":"omics-specialized","source":"EMBL-EBI / BioStudies","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/biostudies/api/v1/arrayexpress/search?pageSize=1","method":"GET","parse":{"type":"json","path":"totalHits","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",80510],["2026-07-06",80584],["2026-07-13",80616],["2026-07-20",80631]]},{"id":"biostudies","label":"BioStudies total studies","category":"omics-specialized","source":"EMBL-EBI / BioStudies","unit":"datasets","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/biostudies/api/v1/search?pageSize=1","method":"GET","parse":{"type":"json","path":"totalHits","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",3343764],["2026-07-06",3356171],["2026-07-13",3370410],["2026-07-20",3401386]]},{"id":"pride","label":"PRIDE Archive projects","category":"omics-specialized","source":"EMBL-EBI / PRIDE","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/pride/ws/archive/v2/projects/count","method":"GET","parse":{"type":"text","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",39927],["2026-07-06",40166],["2026-07-13",40264],["2026-07-20",40361]]},{"id":"metabolights","label":"MetaboLights studies","category":"omics-specialized","source":"EMBL-EBI / MetaboLights","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/metabolights/ws/studies","method":"GET","parse":{"type":"json","path":"studies","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",3097],["2026-07-06",3141],["2026-07-13",3163],["2026-07-20",3181]]},{"id":"clinvar","label":"ClinVar records","category":"omics-specialized","source":"NCBI / ClinVar","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=clinvar&term=all%5Bfilter%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2014-01-01",17144],["2015-01-01",41418],["2016-01-01",93235],["2017-01-01",218645],["2018-01-01",323982],["2019-01-01",457637],["2020-01-01",636017],["2021-01-01",819828],["2022-01-01",1151479],["2023-01-01",1628253],["2024-01-01",2408416],["2025-01-01",3115127],["2026-01-01",4213808],["2026-06-24",4530887],["2026-07-06",4531457],["2026-07-13",4531702],["2026-07-20",4531941]]},{"id":"dbsnp","label":"dbSNP reference SNPs","category":"omics-specialized","source":"NCBI / dbSNP","unit":"records","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=snp&term=all%5Bsb%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",1197210835]]},{"id":"sars-cov-2","label":"SARS-CoV-2 sequences","category":"omics-specialized","source":"NCBI Virus","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=nuccore&term=txid2697049%5BOrganism%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2020-12-31",47129],["2021-12-31",3037813],["2022-12-31",6609761],["2023-12-31",8650619],["2024-12-31",9023342],["2026-06-24",9209080],["2026-07-06",9209569],["2026-07-13",9210005],["2026-07-20",9210361]],"sources":[{"label":"NCBI Virus — SARS-CoV-2 sequences by collection year","url":"https://www.ncbi.nlm.nih.gov/labs/virus/vssi/#/virus?taxid=2697049"}],"events":[{"date":"2020-01-11","label":"First genome shared","detail":"The first SARS-CoV-2 genome was publicly released on Virological.org on 11 January 2020, enabling worldwide diagnostics, vaccine design, and the sequencing surge that follows."},{"date":"2020-03-11","label":"WHO declares pandemic","detail":"On 11 March 2020 the WHO declared COVID-19 a pandemic, triggering global surveillance programs that drove the steep early ramp in genome submissions."},{"date":"2020-12-08","label":"First vaccine rollout","detail":"The UK began the world's first authorized COVID-19 vaccinations (Pfizer-BioNTech) on 8 December 2020, intensifying variant surveillance to monitor vaccine escape."},{"date":"2021-05-11","label":"Delta VOC named","detail":"The WHO classified the Delta lineage (B.1.617.2) as a variant of concern on 11 May 2021, fueling the mid-2021 sequencing wave as the more transmissible strain spread globally."},{"date":"2021-11-26","label":"Omicron VOC named","detail":"The WHO designated Omicron (B.1.1.529) a variant of concern on 26 November 2021, prompting a massive spike in sequencing as the heavily mutated variant swept the world."}]},{"id":"ncbi-taxonomy-species","label":"NCBI Taxonomy named species","category":"standards-vocab","source":"NCBI / Taxonomy","unit":"species","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=taxonomy&term=species%5Brank%5D&retmode=json&retmax=0&tool=seqdesk-tracker&email=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2026-06-24",2327771],["2026-07-06",2329845],["2026-07-13",2332354],["2026-07-20",2342699]]},{"id":"envo","label":"ENVO ontology terms","category":"standards-vocab","source":"OLS4 / EMBL-EBI","unit":"terms","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ols4/api/ontologies/envo","method":"GET","parse":{"type":"json","path":"numberOfTerms","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",6906],["2026-07-13",6936]]},{"id":"ena-checklists","label":"ENA sample checklists","category":"standards-vocab","source":"EMBL-EBI / ENA","unit":"checklists","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ena/submit/report/checklists","method":"GET","parse":{"type":"array-length"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",46]]},{"id":"obo-foundry","label":"OBO Foundry active ontologies","category":"standards-vocab","source":"OBO Foundry","unit":"ontologies","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://obofoundry.org/registry/ontologies.jsonld","method":"GET","parse":{"type":"obo-active"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",189]],"note":"Filter ontologies[].activity_status=='active' and count. Wire a custom fetcher."},{"id":"mixs-terms","label":"GSC MIxS terms","category":"standards-vocab","source":"Genomic Standards Consortium","unit":"terms","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://raw.githubusercontent.com/GenomicsStandardsConsortium/mixs/main/src/mixs/schema/mixs.yaml","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",1125]],"note":"Parse the MIxS YAML schema (count slots). Wire a custom fetcher."},{"id":"mixs-paper-citations","label":"MIxS specification paper citations","category":"standards-vocab","source":"OpenAlex (Yilmaz 2011)","unit":"citations","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.openalex.org/works?filter=doi:10.1038/nbt.1823&select=id,cited_by_count,updated_date&mailto=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"results.0.cited_by_count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"note":"Cumulative citations of the MIxS / MIMARKS specification (Yilmaz et al. 2011, Nat Biotechnol) — the adoption proxy for the very metadata standard SeqDesk captures samples against. The FAIR-principles paper ([[fair-citations]]) is the broader companion.","events":[{"date":"2011-05-08","label":"MIxS specification published","detail":"Yilmaz et al. published MIMARKS & MIxS in Nature Biotechnology, defining the minimum contextual fields (geography, environment, host, collection) a sequence record should carry."},{"date":"2016-03-15","label":"FAIR Principles","detail":"Wilkinson et al. (Sci Data) put rich, standardized, machine-readable metadata at the centre of data stewardship — pulling MIxS-style checklists into the FAIR mainstream."},{"date":"2023-03-03","label":"INSDC spatiotemporal standards","detail":"INSDC tightened collection-date & lat/lon reporting and standardised the missing-value vocabulary, reinforcing the MIxS contextual backbone."}],"series":[["2013-01-01",32],["2014-01-01",82],["2015-01-01",129],["2016-01-01",173],["2017-01-01",211],["2018-01-01",259],["2019-01-01",331],["2020-01-01",373],["2021-01-01",448],["2022-01-01",507],["2023-01-01",591],["2024-01-01",645],["2025-01-01",704],["2026-01-01",759],["2026-06-29",809],["2026-07-06",812],["2026-07-13",814],["2026-07-20",815]]},{"id":"ena-sample-fields","label":"ENA sample metadata fields","category":"standards-vocab","source":"EMBL-EBI / ENA","unit":"fields","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/searchFields?result=sample&dataPortal=ena","method":"GET","parse":{"type":"tsv-rows"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",95]],"note":"TSV — count data rows minus header."},{"id":"fair-citations","label":"FAIR Principles paper citations","category":"fair-literature","source":"OpenAlex (Wilkinson 2016)","unit":"citations","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.openalex.org/works?filter=doi:10.1038/sdata.2016.18&select=id,cited_by_count,updated_date&mailto=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"results.0.cited_by_count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2016-01-01",73],["2017-01-01",366],["2018-01-01",947],["2019-01-01",1942],["2020-01-01",3466],["2021-01-01",5489],["2022-01-01",8071],["2023-01-01",10867],["2024-01-01",13642],["2025-01-01",16323],["2026-06-24",17727],["2026-07-06",17866],["2026-07-13",17945],["2026-07-20",18022]],"note":"Cumulative citations of the FAIR Guiding Principles paper — the best single FAIR-adoption proxy.","events":[{"date":"2016-03-15","label":"FAIR Principles paper","detail":"Wilkinson et al. published 'The FAIR Guiding Principles for scientific data management and stewardship' in Scientific Data, coining the FAIR acronym and seeding the citation curve."},{"date":"2016-09-05","label":"G20 Hangzhou endorses FAIR","detail":"The G20 Leaders' Communiqué from the Hangzhou Summit endorsed open science and access to publicly funded research on FAIR principles, lending the framework high-level political backing."},{"date":"2018-09-04","label":"Plan S / cOAlition S","detail":"National funders, backed by the European Commission and ERC, launched cOAlition S and Plan S to mandate immediate open access, amplifying funder-driven demand for FAIR data practices."},{"date":"2021-04-28","label":"Horizon Europe FAIR DMP","detail":"Regulation (EU) 2021/695 established Horizon Europe (2021-2027), making FAIR-compliant data management and a Data Management Plan a non-optional requirement for all funded projects."},{"date":"2023-01-25","label":"NIH DMS Policy","detail":"The NIH Data Management and Sharing Policy took effect, requiring DMS plans and deposit in trusted, FAIR-aligned repositories for all new NIH-funded research generating scientific data."}]},{"id":"openalex-total","label":"Total scholarly works (OpenAlex)","category":"fair-literature","source":"OpenAlex","unit":"works","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.openalex.org/works?per-page=1&select=id&mailto=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"meta.count","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2022-01-03",209000000],["2022-01-04",204557662],["2022-10-13",243553804],["2022-12-25",245547508],["2023-03-31",249384043],["2023-05-30",240678124],["2023-10-16",245161833],["2023-11-03",245728617],["2024-04-13",251568022],["2024-09-10",259074808],["2024-12-03",261625306],["2025-03-17",265074378],["2025-05-02",266577590],["2025-09-23",271095541],["2025-12-31",280109059],["2026-06-24",317500885],["2026-07-06",318967691],["2026-07-13",319512759],["2026-07-20",320788011]],"sources":[{"label":"OpenAlex /works snapshots over time","url":"https://api.openalex.org/works"}]},{"id":"europepmc-oa","label":"Open-access articles (Europe PMC)","category":"fair-literature","source":"EMBL-EBI / Europe PMC","unit":"records","tier":"flagship","flagship":true,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/europepmc/webservices/rest/search?query=OPEN_ACCESS:y&format=json&pageSize=1&resultType=idlist","method":"GET","parse":{"type":"json","path":"hitCount","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2005-01-01",17352],["2006-01-01",41433],["2007-01-01",70346],["2008-01-01",115466],["2009-01-01",175967],["2010-01-01",257672],["2011-01-01",368584],["2012-01-01",519095],["2013-01-01",705677],["2014-01-01",929580],["2015-01-01",1181896],["2016-01-01",1458078],["2017-01-01",1773283],["2018-01-01",2129160],["2019-01-01",2544028],["2020-01-01",3145349],["2021-01-01",3900183],["2022-01-01",4751994],["2023-01-01",5519088],["2024-01-01",6280387],["2025-01-01",7194301],["2026-06-24",7995922],["2026-07-06",8023278],["2026-07-13",8023599],["2026-07-20",8024204]],"sources":[{"label":"Europe PMC — open-access article counts by year (REST search)","url":"https://europepmc.org/RestfulWebService"}],"note":"Cumulative open-access article records in Europe PMC. Historical points are the running sum of OA articles by publication year (which approaches the all-time total); the latest point is the live OPEN_ACCESS:y total."},{"id":"openalex-datasets","label":"OpenAlex dataset works","category":"fair-literature","source":"OpenAlex","unit":"datasets","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://api.openalex.org/works?filter=type:dataset&per-page=1&mailto=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"meta.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-24",24438554],["2026-07-06",24731164],["2026-07-13",24785477],["2026-07-20",24791923]]},{"id":"crossref-datasets","label":"Crossref dataset DOIs","category":"fair-literature","source":"Crossref","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.crossref.org/works?filter=type:dataset&rows=0","method":"GET","parse":{"type":"json","path":"message.total-results","cast":"int"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2010-01-01",62720],["2011-01-01",73812],["2012-01-01",95570],["2013-01-01",256799],["2014-01-01",657913],["2015-01-01",930893],["2016-01-01",1179153],["2017-01-01",1386576],["2018-01-01",1656424],["2019-01-01",1767677],["2020-01-01",1922915],["2021-01-01",2089172],["2022-01-01",2289233],["2023-01-01",2844158],["2024-01-01",3011032],["2025-01-01",3143323],["2026-01-01",3360429],["2026-06-24",3433650],["2026-07-06",3435029],["2026-07-13",3435979],["2026-07-20",3437831]]},{"id":"fair-mentions","label":"Works mentioning “FAIR data”","category":"fair-literature","source":"OpenAlex","unit":"works","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://api.openalex.org/works?search=FAIR%20data&per-page=1&mailto=chrmnch@icloud.com","method":"GET","parse":{"type":"json","path":"meta.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2016-01-01",595830],["2017-01-01",671717],["2018-01-01",754966],["2019-01-01",845375],["2020-01-01",949226],["2021-01-01",1075249],["2022-01-01",1212958],["2023-01-01",1361857],["2024-01-01",1590372],["2025-01-01",1834340],["2026-01-01",2121698],["2026-06-24",2271606],["2026-07-06",2298364],["2026-07-13",2309422],["2026-07-20",2321771]]},{"id":"biorxiv","label":"bioRxiv preprints","category":"fair-literature","source":"bioRxiv / CSHL","unit":"preprints","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.biorxiv.org/details/biorxiv/2013-11-01/{TODAY}/0","method":"GET","parse":{"type":"biorxiv"},"auto":true,"appendPolicy":"always"},"release":null,"series":[["2014-01-01",109],["2015-01-01",995],["2016-01-01",2769],["2017-01-01",7487],["2018-01-01",18826],["2019-01-01",39604],["2020-01-01",68783],["2021-01-01",107500],["2022-01-01",144366],["2023-01-01",180116],["2024-01-01",219270],["2025-01-01",262899],["2026-01-01",312197],["2026-06-24",337398],["2026-07-06",338753],["2026-07-13",340019],["2026-07-20",341476]],"note":"Sum count_new_papers across paginated 'details' pages, or read cumulative from collection stats. Wire a custom fetcher."},{"id":"pangaea-datasets","label":"PANGAEA datasets","category":"earth-environment","source":"PANGAEA (AWI / MARUM)","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://ws.pangaea.de/es/pangaea/panmd/_count","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",445990],["2026-07-06",446061],["2026-07-13",446230],["2026-07-20",446307]],"note":"Curated Earth & environmental-science datasets in PANGAEA, each with a DOI. One record = one published dataset."},{"id":"esgf-cmip6","label":"CMIP6 climate datasets (ESGF)","category":"earth-environment","source":"ESGF / WCRP CMIP6","unit":"datasets","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://esgf-node.ornl.gov/esgf-1-5-bridge?project=CMIP6&limit=0&format=application%2Fsolr%2Bjson","method":"GET","parse":{"type":"json","path":"response.numFound","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",14769319]],"note":"CMIP6 climate-model output datasets distributed through the Earth System Grid Federation — the data behind IPCC-class projections."},{"id":"nasa-cmr-collections","label":"NASA EOSDIS data collections","category":"earth-environment","source":"NASA EOSDIS Common Metadata Repository","unit":"collections","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://cmr.earthdata.nasa.gov/search/collections?page_size=0","method":"GET","parse":{"type":"header","header":"CMR-Hits"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",54950],["2026-07-06",54965],["2026-07-13",54876],["2026-07-20",54862]],"note":"Distinct Earth-observation data collections catalogued in NASA's EOSDIS Common Metadata Repository (satellite, airborne & field datasets)."},{"id":"inspire-hep","label":"INSPIRE-HEP physics records","category":"physics-materials","source":"INSPIRE-HEP (CERN)","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://inspirehep.net/api/literature?q=&size=1&fields=titles","method":"GET","parse":{"type":"json","path":"hits.total","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1865337],["2026-07-06",1868003],["2026-07-13",1869980],["2026-07-20",1871213]],"note":"High-energy-physics literature records curated by INSPIRE (CERN/DESY/Fermilab/SLAC)."},{"id":"cern-opendata","label":"CERN Open Data records","category":"physics-materials","source":"CERN Open Data Portal","unit":"records","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://opendata.cern.ch/api/records","method":"GET","parse":{"type":"json","path":"hits.total","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",80339],["2026-07-06",80347],["2026-07-13",80372],["2026-07-20",80373]],"note":"Open datasets, software & documentation released through the CERN Open Data portal (LHC experiments and more)."},{"id":"simbad-objects","label":"SIMBAD astronomical objects","category":"physics-materials","source":"CDS Strasbourg / SIMBAD","unit":"objects","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://simbad.cds.unistra.fr/simbad/sim-tap/sync?request=doQuery&lang=adql&format=json&query=SELECT+COUNT(*)+FROM+basic","method":"GET","parse":{"type":"json","path":"data.0.0","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",21450272],["2026-07-06",21757670],["2026-07-13",21807528],["2026-07-20",21808058]],"note":"Astronomical objects with cross-identifications and bibliography in the CDS SIMBAD database."},{"id":"gaia-dr3","label":"Gaia DR3 sources","category":"physics-materials","source":"ESA / Gaia DR3 (ESAC TAP)","unit":"sources","tier":"slow-moving","flagship":false,"cadence":"rarely","scale":"linear","fetch":{"url":"https://gea.esac.esa.int/tap-server/tap/sync?REQUEST=doQuery&LANG=ADQL&FORMAT=json&QUERY=SELECT+COUNT(*)+FROM+gaiadr3.gaia_source","method":"GET","parse":{"type":"json","path":"data.0.0","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1811709771]],"note":"Sources (stars, quasars…) in ESA Gaia Data Release 3 — the most detailed map of the Milky Way. Fixed per data release.","events":[{"date":"2022-06-13","label":"Gaia DR3 released","detail":"ESA published Gaia Data Release 3 in June 2022, with ~1.8 billion sources."}]},{"id":"exoplanets","label":"Confirmed exoplanets","category":"physics-materials","source":"NASA Exoplanet Archive (Caltech/IPAC)","unit":"planets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://exoplanetarchive.ipac.caltech.edu/TAP/sync?query=select+count(*)+from+pscomppars&format=json","method":"GET","parse":{"type":"json","path":"0.count(*)","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",6298],["2026-07-06",6316],["2026-07-13",6319],["2026-07-20",6324]],"note":"Confirmed exoplanets in the NASA Exoplanet Archive. Grows as new worlds are validated."},{"id":"vizier-catalogues","label":"VizieR astronomy catalogues","category":"physics-materials","source":"CDS Strasbourg / VizieR","unit":"catalogues","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://tapvizier.cds.unistra.fr/TAPVizieR/tap/sync?REQUEST=doQuery&LANG=ADQL&FORMAT=json&QUERY=SELECT+COUNT(*)+FROM+METAcat","method":"GET","parse":{"type":"json","path":"data.0.0","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",27749],["2026-07-06",27800],["2026-07-13",27829],["2026-07-20",27841]],"note":"Published astronomical catalogues & tables served by the CDS VizieR service — a metadata-richness proxy for the field."},{"id":"materials-project","label":"Materials Project structures","category":"physics-materials","source":"Materials Project (LBNL) · OPTIMADE","unit":"structures","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://optimade.materialsproject.org/v1/structures?page_limit=1","method":"GET","parse":{"type":"json","path":"meta.data_available","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",154387],["2026-07-20",0]],"note":"Computed inorganic materials (mostly DFT-relaxed structures) in the Materials Project, served via OPTIMADE."},{"id":"oqmd","label":"OQMD computed materials","category":"physics-materials","source":"OQMD (Northwestern) · OPTIMADE","unit":"structures","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://oqmd.org/optimade/structures?page_limit=1&response_fields=id","method":"GET","parse":{"type":"json","path":"meta.data_available","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1407395]],"note":"DFT-calculated thermodynamic & structural properties of materials in the Open Quantum Materials Database (OPTIMADE)."},{"id":"nomad","label":"NOMAD calculations","category":"physics-materials","source":"NOMAD Laboratory","unit":"entries","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://nomad-lab.eu/prod/v1/api/v1/entries?page_size=1","method":"GET","parse":{"type":"json","path":"pagination.total","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",19338439],["2026-07-06",19375576],["2026-07-13",19392749],["2026-07-20",19393669]],"note":"Raw & processed computational-materials-science calculations archived in NOMAD."},{"id":"pubchem-compounds","label":"PubChem compounds","category":"chemistry-compounds","source":"NCBI / PubChem","unit":"compounds","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=pccompound&term=all[filt]&retmax=0&retmode=json","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",123920602],["2026-07-06",123991844],["2026-07-13",124001167],["2026-07-20",124004062]],"note":"Unique chemical structures (normalized compounds) in PubChem.","events":[{"date":"2004-09-01","label":"PubChem launches","detail":"NCBI launched PubChem in 2004 as part of the NIH Molecular Libraries Roadmap."}]},{"id":"pubchem-substances","label":"PubChem substances","category":"chemistry-compounds","source":"NCBI / PubChem","unit":"substances","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=pcsubstance&term=all[filt]&retmax=0&retmode=json","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",347098006],["2026-07-06",347160449],["2026-07-13",347271563],["2026-07-20",347276725]],"note":"Depositor-submitted substance records in PubChem — the raw, metadata-bearing layer beneath the deduplicated compounds."},{"id":"pubchem-bioassays","label":"PubChem bioassays","category":"chemistry-compounds","source":"NCBI / PubChem","unit":"bioassays","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://eutils.ncbi.nlm.nih.gov/entrez/eutils/esearch.fcgi?db=pcassay&term=all[filt]&retmax=0&retmode=json","method":"GET","parse":{"type":"json","path":"esearchresult.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1884782],["2026-07-06",1884783],["2026-07-13",1884788]],"note":"Biological assay datasets (screening results) deposited in PubChem."},{"id":"chembl-molecules","label":"ChEMBL bioactive molecules","category":"chemistry-compounds","source":"EMBL-EBI / ChEMBL","unit":"molecules","tier":"secondary","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/chembl/api/data/molecule.json?limit=1","method":"GET","parse":{"type":"json","path":"page_meta.total_count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",2921148]],"note":"Distinct bioactive, drug-like molecules curated in ChEMBL."},{"id":"chebi-terms","label":"ChEBI ontology terms","category":"chemistry-compounds","source":"EMBL-EBI / ChEBI (OLS4)","unit":"terms","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/ols4/api/ontologies/chebi","method":"GET","parse":{"type":"json","path":"numberOfTerms","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",224691],["2026-07-13",237672]],"note":"Ontology terms for chemical entities of biological interest (ChEBI), via EMBL-EBI OLS4."},{"id":"gbif-occurrences","label":"GBIF occurrence records","category":"biodiversity","source":"GBIF — Global Biodiversity Information Facility","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/occurrence/search?limit=0","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",3894470556],["2026-07-06",3901234507],["2026-07-13",3903196936],["2026-07-20",3904814387]],"note":"Species occurrence records (where/when an organism was observed or collected) aggregated globally by GBIF.","events":[{"date":"2001-03-01","label":"GBIF established","detail":"Governments founded the Global Biodiversity Information Facility in 2001 to open access to biodiversity data worldwide."},{"date":"2018-08-01","label":"Passed 1 billion records","detail":"GBIF crossed one billion mediated occurrence records in 2018."}]},{"id":"gbif-datasets","label":"GBIF datasets","category":"biodiversity","source":"GBIF — Global Biodiversity Information Facility","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/dataset?limit=0","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",122876],["2026-07-06",123042],["2026-07-13",123124],["2026-07-20",123222]],"note":"Distinct datasets published through the GBIF network."},{"id":"gbif-dna","label":"GBIF DNA-derived records","category":"biodiversity","source":"GBIF — Global Biodiversity Information Facility","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/occurrence/search?limit=0&dwcaExtension=http://rs.gbif.org/terms/1.0/DNADerivedData","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",152448536],["2026-07-06",152457189],["2026-07-13",152457535],["2026-07-20",152458226]],"note":"GBIF occurrences carrying a DNA-derived-data extension — sequence-based biodiversity records bridging genomics and ecology."},{"id":"gbif-literature","label":"GBIF-citing literature","category":"biodiversity","source":"GBIF — Global Biodiversity Information Facility","unit":"papers","tier":"secondary","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/literature/search?limit=0","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",61648],["2026-07-06",61808],["2026-07-13",61921],["2026-07-20",62055]],"note":"Peer-reviewed works that cite GBIF-mediated data (a research-uptake signal)."},{"id":"obis-records","label":"OBIS marine occurrences","category":"biodiversity","source":"OBIS — Ocean Biodiversity Information System","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.obis.org/v3/statistics","method":"GET","parse":{"type":"json","path":"records","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",200773218],["2026-07-06",200829115],["2026-07-13",201398653],["2026-07-20",202077511]],"note":"Marine species occurrence records aggregated by OBIS."},{"id":"inat-observations","label":"iNaturalist observations","category":"biodiversity","source":"iNaturalist","unit":"observations","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.inaturalist.org/v1/observations?per_page=0","method":"GET","parse":{"type":"json","path":"total_results","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",361629139],["2026-07-06",365062356],["2026-07-13",367304262],["2026-07-20",369499966]],"note":"Community-contributed nature observations on iNaturalist (all quality grades).","events":[{"date":"2008-03-01","label":"iNaturalist founded","detail":"Started as a UC Berkeley master's project in 2008; later a joint initiative of the California Academy of Sciences and National Geographic."}]},{"id":"ctgov-studies","label":"ClinicalTrials.gov studies","category":"clinical-biomed","source":"ClinicalTrials.gov (NLM)","unit":"studies","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://clinicaltrials.gov/api/v2/stats/size","method":"GET","parse":{"type":"json","path":"totalStudies","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",591039],["2026-07-06",592210],["2026-07-13",593334],["2026-07-20",594543]],"note":"Clinical studies registered on ClinicalTrials.gov.","events":[{"date":"2000-02-29","label":"ClinicalTrials.gov launches","detail":"NLM opened ClinicalTrials.gov in February 2000 as the first public registry of clinical studies."},{"date":"2007-09-27","label":"FDAAA 801 mandate","detail":"US law began requiring registration and, later, results reporting for many trials."},{"date":"2017-01-18","label":"Final Rule + NIH policy","detail":"Expanded registration and results-submission requirements took effect, lifting reporting rates."}]},{"id":"gwas-associations","label":"GWAS Catalog associations","category":"clinical-biomed","source":"EMBL-EBI / GWAS Catalog","unit":"associations","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://www.ebi.ac.uk/gwas/api/search/stats","method":"GET","parse":{"type":"json","path":"associations","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1150105],["2026-07-13",1179525]],"note":"Curated SNP–trait associations in the GWAS Catalog."},{"id":"encode-experiments","label":"ENCODE experiments","category":"clinical-biomed","source":"ENCODE Project (Stanford)","unit":"experiments","tier":"secondary","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://www.encodeproject.org/search/?type=Experiment&format=json&limit=0","method":"GET","parse":{"type":"json","path":"total","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",28642]],"note":"Functional-genomics experiments released by the ENCODE project."},{"id":"neuromorpho","label":"NeuroMorpho reconstructions","category":"clinical-biomed","source":"NeuroMorpho.Org (George Mason University)","unit":"reconstructions","tier":"secondary","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://neuromorpho.org/api/neuron?size=1&page=0","method":"GET","parse":{"type":"json","path":"page.totalElements","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",297676],["2026-07-06",297868]],"note":"Digitally reconstructed neuron morphologies in NeuroMorpho.Org."},{"id":"harvard-dataverse","label":"Harvard Dataverse datasets","category":"open-data","source":"Harvard Dataverse (IQSS)","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://dataverse.harvard.edu/api/search?q=*&type=dataset&per_page=1","method":"GET","parse":{"type":"json","path":"data.total_count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",300214],["2026-07-06",300707],["2026-07-13",301046],["2026-07-20",301241]],"note":"Published datasets in the Harvard Dataverse repository."},{"id":"figshare-datasets","label":"figshare datasets","category":"open-data","source":"figshare (via DataCite)","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.datacite.org/dois?query=types.resourceTypeGeneral:Dataset&client-id=figshare.ars&page[size]=0","method":"GET","parse":{"type":"json","path":"meta.total","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2013-01-01",1060],["2014-01-01",6309],["2015-01-01",31403],["2016-01-01",48028],["2017-01-01",152301],["2018-01-01",223479],["2019-01-01",313705],["2020-01-01",414903],["2021-01-01",511170],["2022-01-01",646174],["2023-01-01",783638],["2024-01-01",891547],["2025-01-01",1033657],["2026-01-01",1173629],["2026-06-25",1242870],["2026-07-06",1247183],["2026-07-13",1250400],["2026-07-20",1253504]],"note":"figshare items typed as Dataset, counted via the DataCite REST API."},{"id":"dataeuropa","label":"EU Open Data Portal datasets","category":"open-data","source":"data.europa.eu — EU Open Data Portal","unit":"datasets","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://data.europa.eu/api/hub/search/search?limit=0&filter=dataset","method":"GET","parse":{"type":"json","path":"result.count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1747469],["2026-07-06",1745535],["2026-07-13",1752614],["2026-07-20",1757021]],"note":"Government & public-sector datasets indexed by the EU Open Data Portal."},{"id":"worldbank-indicators","label":"World Bank indicators","category":"open-data","source":"World Bank Open Data","unit":"indicators","tier":"slow-moving","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://api.worldbank.org/v2/indicator?format=json&per_page=1","method":"GET","parse":{"type":"json","path":"0.total","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",29512],["2026-07-06",29536],["2026-07-20",29556]],"note":"Distinct development indicators (each a multi-country time series) in the World Bank Open Data API."},{"id":"gbif-georef","label":"GBIF georeferenced records","category":"metadata-completeness","source":"GBIF — Global Biodiversity Information Facility","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/occurrence/search?limit=0&hasCoordinate=true&hasGeospatialIssue=false","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",3722568085],["2026-07-06",3728908326],["2026-07-13",3730768664],["2026-07-20",3732462762]],"note":"Of all GBIF occurrences, how many carry usable coordinates (no geospatial issues) — a direct measure of spatial-metadata completeness."},{"id":"gbif-images","label":"GBIF records with images","category":"metadata-completeness","source":"GBIF — Global Biodiversity Information Facility","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/occurrence/search?limit=0&mediaType=StillImage","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",291922475],["2026-07-06",293903220],["2026-07-13",295492676],["2026-07-20",296261563]],"note":"GBIF occurrences with at least one still image attached — richer, verifiable evidence beyond a bare record."},{"id":"gbif-types","label":"GBIF type specimens","category":"metadata-completeness","source":"GBIF — Global Biodiversity Information Facility","unit":"records","tier":"secondary","flagship":false,"cadence":"monthly","scale":"linear","fetch":{"url":"https://api.gbif.org/v1/occurrence/search?limit=0&typeStatus=Holotype","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",1103626],["2026-07-06",1104413],["2026-07-13",1072116],["2026-07-20",1069550]],"note":"Type specimens (e.g. holotypes) in GBIF — the highest-provenance, nomenclaturally anchored records."},{"id":"inat-research","label":"iNaturalist research-grade","category":"metadata-completeness","source":"iNaturalist","unit":"observations","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.inaturalist.org/v1/observations?per_page=0&quality_grade=research","method":"GET","parse":{"type":"json","path":"total_results","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2014-01-01",380560],["2015-01-01",777354],["2016-01-01",1454257],["2017-01-01",2775009],["2018-01-01",5497530],["2019-01-01",11053009],["2020-01-01",21238003],["2021-01-01",38484424],["2022-01-01",59296941],["2023-01-01",83124257],["2024-01-01",111743151],["2025-01-01",145614721],["2026-01-01",185152979],["2026-06-25",205029734],["2026-07-06",206790320],["2026-07-13",207955637],["2026-07-20",209138375]],"note":"iNaturalist observations promoted to 'research grade' (community-verified ID + date + location) — the metadata-complete subset usable for science."},{"id":"obis-species","label":"OBIS species-level records","category":"metadata-completeness","source":"OBIS — Ocean Biodiversity Information System","unit":"records","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://api.obis.org/v3/statistics","method":"GET","parse":{"type":"json","path":"specieslevel","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",121923028],["2026-07-06",121969875],["2026-07-13",122237042],["2026-07-20",122786635]],"note":"OBIS records resolved to species level — taxonomic-metadata completeness for marine occurrences."},{"id":"ctgov-results","label":"Trials with posted results","category":"metadata-completeness","source":"ClinicalTrials.gov (NLM)","unit":"studies","tier":"secondary","flagship":false,"cadence":"weekly","scale":"linear","fetch":{"url":"https://clinicaltrials.gov/api/v2/studies?aggFilters=results:with&countTotal=true&pageSize=1&fields=NCTId","method":"GET","parse":{"type":"json","path":"totalCount","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2026-06-25",78797],["2026-07-06",78991],["2026-07-13",79112],["2026-07-20",79230]],"note":"Registered trials that have actually posted results — a transparency / metadata-completeness signal for clinical research."},{"id":"cost-per-genome","label":"Cost to sequence a human genome","category":"sequencing-technology","source":"NHGRI — DNA Sequencing Costs (Genome Sequencing Program)","unit":"USD per genome","tier":"secondary","flagship":false,"cadence":"release","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"headline":{"value":"$525","unit":"per genome · 2022"},"forecast":false,"series":[["2001-09-30",95263072],["2002-03-31",70175437],["2003-03-31",53751684],["2004-01-31",28780376],["2005-01-31",17534970],["2006-01-31",12585659],["2007-01-31",9408739],["2008-01-31",3063820],["2009-01-31",232735],["2010-01-31",46774],["2011-01-31",20963],["2012-01-31",7666],["2013-01-31",5671],["2014-01-31",4008],["2015-01-31",3970],["2016-05-31",1176],["2017-02-28",1015],["2018-02-28",1232],["2019-02-28",993],["2020-02-28",645],["2021-02-28",851],["2022-05-31",525]],"note":"The most-cited chart in genomics: NHGRI's cost to generate one high-quality human whole genome (~30x), from ~$95M in Sept 2001 to ~$525 by May 2022 — a >180,000-fold drop. The famous cliff from 2008 is where second-generation sequencing made the cost fall far faster than computing's Moore's Law. NHGRI updates this dataset on its own (irregular) schedule; the curve is intentionally NOT forecast — cost has roughly plateaued since the $1,000-genome era and straight-line extrapolation would mislead.","sources":[{"label":"NHGRI — DNA Sequencing Costs: Data from the NHGRI Genome Sequencing Program (GSP)","url":"https://www.genome.gov/about-genomics/fact-sheets/DNA-Sequencing-Costs-Data"},{"label":"NHGRI cost data table (May 2022 release, .xls)","url":"https://www.genome.gov/sites/default/files/media/files/2023-05/Sequencing_Cost_Data_Table_May2022.xls"}],"events":[{"date":"2005-09-01","label":"454 pyrosequencing","detail":"Margulies et al. (Nature 437:376) published the 454 picolitre-plate pyrosequencer, the first next-generation platform — roughly 100x the throughput of capillary Sanger instruments."},{"date":"2007-01-01","label":"Illumina/Solexa SBS","detail":"Solexa's sequencing-by-synthesis Genome Analyzer (Illumina acquired Solexa in 2007) brought gigabase-scale short reads, the workhorse chemistry behind most of this cost decline."},{"date":"2008-01-01","label":"Cost breaks Moore's Law","detail":"From 2008 the cost per genome fell far faster than computing's Moore's Law as second-generation sequencing scaled — the steep drop from ~$9.4M (Jan 2007) to ~$233k (Jan 2009)."},{"date":"2014-01-14","label":"$1,000 genome","detail":"Illumina launched the HiSeq X Ten on 14 Jan 2014, the first platform marketed as breaking the $1,000-per-genome barrier for population-scale whole-genome sequencing."},{"date":"2017-01-09","label":"NovaSeq 6000","detail":"Illumina introduced the NovaSeq series in January 2017, pushing per-genome cost toward a few hundred dollars at production scale."},{"date":"2022-09-29","label":"NovaSeq X / sub-$200","detail":"Illumina announced the NovaSeq X Series on 29 Sep 2022, claiming a ~$200 genome at maximum scale — beyond the end of the NHGRI series shown here."}]},{"id":"cost-per-gigabase","label":"Cost per gigabase of sequence","category":"sequencing-technology","source":"NHGRI — DNA Sequencing Costs (Genome Sequencing Program)","unit":"USD per Gb","tier":"secondary","flagship":false,"cadence":"release","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"headline":{"value":"$5.80","unit":"per Gb · 2022"},"forecast":false,"series":[["2001-09-30",5292393],["2002-03-31",3898635],["2003-03-31",2986205],["2004-01-31",1598910],["2005-01-31",974165],["2006-01-31",699203],["2007-01-31",522708],["2008-01-31",102127],["2009-01-31",2586],["2010-01-31",520],["2011-01-31",233],["2012-01-31",85.2],["2013-01-31",63],["2014-01-31",44.5],["2015-01-31",44.1],["2016-05-31",13.1],["2017-02-28",11.3],["2018-02-28",13.7],["2019-02-28",11],["2020-02-28",7.2],["2021-02-28",9.5],["2022-05-31",5.8]],"note":"The raw-sequence view of the same NHGRI data: cost per gigabase (1 Gb = 1,000 Mb) of DNA sequence, from ~$5.3M/Gb in 2001 to ~$5.80/Gb in 2022 (NHGRI publishes this as cost per megabase; shown here per Gb so every point stays above the chart's log floor). This decoupled from cost-per-genome as coverage standards and read counts rose, and is the better measure of pure sequencing-output economics. Not forecast.","sources":[{"label":"NHGRI — DNA Sequencing Costs: Data from the NHGRI Genome Sequencing Program (GSP)","url":"https://www.genome.gov/about-genomics/fact-sheets/DNA-Sequencing-Costs-Data"},{"label":"NHGRI cost data table (May 2022 release, .xls)","url":"https://www.genome.gov/sites/default/files/media/files/2023-05/Sequencing_Cost_Data_Table_May2022.xls"}],"events":[{"date":"2008-01-01","label":"2nd-gen sequencing","detail":"Massively parallel short-read platforms collapsed the cost of raw sequence by orders of magnitude from 2008, the steepest part of this curve."}]},{"id":"ena-platform-illumina","label":"Illumina runs in ENA","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22ILLUMINA%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2014-12-31",781553],["2016-12-31",2453596],["2018-12-31",5586877],["2020-12-31",10523860],["2022-12-31",21582764],["2024-12-31",30955623],["2026-06-25",38482017],["2026-07-06",38670713],["2026-07-13",38754376],["2026-07-20",38840323]],"note":"Cumulative public sequencing runs submitted to the European Nucleotide Archive on Illumina sequencing-by-synthesis platforms — the dominant short-read chemistry. Each record is one submitted run; cumulative, so it only goes up. Counted live from the ENA Portal API by instrument_platform=ILLUMINA.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count by instrument_platform","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22ILLUMINA%22&format=json"},{"label":"ENA advanced/text search","url":"https://www.ebi.ac.uk/ena/browser/text-search?query=ILLUMINA"}]},{"id":"ena-platform-nanopore","label":"Oxford Nanopore runs in ENA","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22OXFORD_NANOPORE%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2014-12-31",13],["2016-12-31",667],["2018-12-31",3111],["2020-12-31",47810],["2022-12-31",547288],["2024-12-31",865708],["2026-06-25",1068585],["2026-07-06",1071454],["2026-07-13",1073438],["2026-07-20",1076471]],"note":"Cumulative public ENA runs on Oxford Nanopore long-read platforms — from just 13 submitted runs by end-2014 (the year the MinION access programme began) to over a million today. The steepest adoption curve of any platform here; counted live from the ENA Portal API by instrument_platform=OXFORD_NANOPORE.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count by instrument_platform","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22OXFORD_NANOPORE%22&format=json"},{"label":"ENA advanced/text search","url":"https://www.ebi.ac.uk/ena/browser/text-search?query=OXFORD_NANOPORE"}]},{"id":"ena-platform-pacbio","label":"PacBio runs in ENA","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22PACBIO_SMRT%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"series":[["2014-12-31",19147],["2016-12-31",40488],["2018-12-31",66908],["2020-12-31",101200],["2022-12-31",666117],["2024-12-31",816427],["2026-06-25",940501],["2026-07-06",942670],["2026-07-13",944488],["2026-07-20",945447]],"note":"Cumulative public ENA runs on Pacific Biosciences SMRT long-read platforms (RS, Sequel, Revio). The jump after 2020 tracks the shift to highly accurate HiFi (CCS) reads and the high-throughput Revio. Counted live from the ENA Portal API by instrument_platform=PACBIO_SMRT.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count by instrument_platform","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22PACBIO_SMRT%22&format=json"},{"label":"ENA advanced/text search","url":"https://www.ebi.ac.uk/ena/browser/text-search?query=PACBIO"}]},{"id":"sequencing-platforms-on-market","label":"Sequencing platform companies on the market","category":"sequencing-technology","source":"SeqDesk — compiled from company histories","unit":"companies","tier":"secondary","flagship":false,"cadence":"curated","scale":"linear","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"headline":{"value":"≈7","unit":"active platform makers · 2023"},"forecast":false,"series":[["1981-01-01",1],["1998-01-01",2],["2005-01-01",4],["2007-01-01",5],["2011-01-01",6],["2016-01-01",4],["2023-01-01",7]],"note":"Approximate count of distinct commercial sequencing-platform makers active over time — an illustrative boom-bust-boom arc, not an exact census. The count climbs through the NGS gold rush (Applied Biosystems Sanger era; Illumina/Solexa; 454; PacBio; Complete Genomics), falls as players are acquired or shut (454 → Roche 2007, line phased out by ~2016; Complete Genomics → BGI 2013; Solexa folded into Illumina 2007), then rises again with a second wave (MGI/BGI outside China, plus Element Biosciences, Ultima and Singular Genomics). Corrections applied during research: ABI's first commercial DNA sequencer (Model 370A) was 1987 — its first commercial instrument was the 470A Protein Sequencer (1982); Complete Genomics' commercial launch was 2010 (2009 was proof-of-concept). Counts are illustrative, so the curve is not forecast.","sources":[{"label":"Illumina, Inc. (history)","url":"https://en.wikipedia.org/wiki/Illumina,_Inc."},{"label":"Complete Genomics","url":"https://en.wikipedia.org/wiki/Complete_Genomics"},{"label":"Applied Biosystems","url":"https://en.wikipedia.org/wiki/Applied_Biosystems"}],"events":[{"date":"1981-01-01","label":"Applied Biosystems","detail":"ABI founded 1981 (Foster City, CA). First commercial instrument: Model 470A Protein Sequencer (1982); first commercial DNA sequencer, the Model 370A, in 1987. Now a Thermo Fisher brand via Life Technologies."},{"date":"1998-04-01","label":"Illumina founded","detail":"Illumina incorporated 1 April 1998 in San Diego; it entered sequencing by acquiring Solexa (completed January 2007), whose sequencing-by-synthesis chemistry became the industry workhorse."},{"date":"2000-01-01","label":"454 founded","detail":"454 Life Sciences founded 2000 by Jonathan Rothberg; its GS20 (2005) was the first commercial next-generation sequencer. Acquired by Roche in 2007 and phased out by ~2016."},{"date":"2004-01-01","label":"PacBio founded","detail":"Pacific Biosciences founded 2004 (originally Nanofluidics); its first single-molecule real-time (SMRT) instrument, the PacBio RS, shipped in 2011."},{"date":"2005-01-01","label":"ONT & Complete Genomics","detail":"Oxford Nanopore (as Oxford Nanolabs) and Complete Genomics were both founded in 2005. Complete Genomics' DNB human-genome service launched commercially in 2010; it was acquired by BGI in 2013, seeding today's MGI."},{"date":"2022-01-01","label":"Second wave","detail":"A new generation of short-read challengers reached the market — Element Biosciences (AVITI), Ultima Genomics, and Singular Genomics — alongside MGI/BGI's global expansion, lifting the count again after the mid-2010s consolidation."}]},{"id":"max-instrument-output-per-run","label":"Max sequencing output per run","category":"sequencing-technology","source":"SeqDesk — compiled from vendor spec sheets","unit":"Gb/run","tier":"secondary","flagship":false,"cadence":"curated","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"headline":{"value":"16 Tb","unit":"per run · NovaSeq X (2023)"},"forecast":false,"series":[["2007-01-01",1],["2010-01-01",600],["2014-01-01",1800],["2017-01-01",6000],["2023-01-01",16000]],"note":"Maximum data output of the highest-throughput instrument available each year, in gigabases per run (log scale). Illumina Genome Analyzer (2007) headlined ~1 Gb/run; HiSeq 2000 reached ~600 Gb with upgraded flow cells (2010); HiSeq X ~1.8 Tb (2014, the '$1,000 genome' machine); NovaSeq 6000 ~6 Tb (2017); NovaSeq X Plus ~16 Tb (2023) — a ~16,000-fold rise in 16 years. For scale, a capillary Sanger run produced ~0.0001 Gb, so the 2007 Genome Analyzer was already ~10,000x more than Sanger. Best-available headline maxima from vendor spec sheets (upgraded-config rather than launch values in places); not forecast, since output jumps at chemistry/flow-cell launches rather than smoothly.","sources":[{"label":"Illumina sequencing platforms & specifications","url":"https://www.illumina.com/systems/sequencing-platforms.html"},{"label":"Illumina sequencing history","url":"https://www.illumina.com/science/technology/next-generation-sequencing/illumina-sequencing-history.html"}],"events":[{"date":"2005-01-01","label":"First NGS instrument","detail":"454 GS20 (2005), the first commercial next-generation sequencer, produced ~25 Mb per run — already well beyond per-run Sanger output."},{"date":"2007-01-01","label":"1 Gb/run","detail":"Illumina/Solexa Genome Analyzer headlined ~1 gigabase per run, the sequencing-by-synthesis breakthrough."},{"date":"2014-01-14","label":"HiSeq X · $1,000 genome","detail":"HiSeq X Ten (announced 14 Jan 2014) delivered ~1.8 Tb per run and the first population-scale $1,000 genome."},{"date":"2023-01-01","label":"NovaSeq X · 16 Tb","detail":"NovaSeq X Plus (shipping from early 2023) reaches ~16 Tb per dual-flow-cell run, the current Illumina output ceiling."}]},{"id":"longest-read-length","label":"Longest sequencing read length","category":"sequencing-technology","source":"SeqDesk — compiled from literature & vendor data","unit":"bp","tier":"secondary","flagship":false,"cadence":"curated","scale":"log","fetch":{"url":"","method":"GET","parse":{"type":"manual"},"auto":false,"appendPolicy":"manual"},"release":null,"headline":{"value":"882 kb","unit":"longest read · nanopore 2018"},"forecast":false,"series":[["1977-01-01",500],["1995-01-01",900],["2008-01-01",900],["2011-01-01",10000],["2014-01-01",64500],["2018-01-01",882000]],"note":"The longest read length practically achievable each year, log scale — and a counter-intuitive story. Sanger reads reached ~500 bp (1977) to ~900 bp (capillary, 1990s) and stayed the ceiling for a decade: when next-generation sequencing arrived it traded length for throughput, so the newest 2006-2010 instruments produced SHORTER reads (Illumina ~35 bp, 454 ~400 bp) than 1977 Sanger — the flat plateau here. Long reads then exploded: PacBio RS ~10 kb (2011), PacBio RS II ~64.5 kb in a published dataset (2014), and an ultra-long Oxford Nanopore read of ~882 kb (2018) — roughly 1,000x beyond Sanger. Several points are dataset maxima or practical ceilings rather than guaranteed specs. Not forecast: read length is platform-defined, not a smooth trend.","sources":[{"label":"PacBio — SMRT sequencing history","url":"https://www.pacb.com/blog/the-evolution-of-dna-sequencing-tools/"},{"label":"Oxford Nanopore — history & ultra-long reads","url":"https://nanoporetech.com/about/history"}],"events":[{"date":"1977-01-01","label":"Sanger method","detail":"Sanger chain-termination sequencing (1977); early reads a few hundred bases, rising to ~900 bp with capillary instruments."},{"date":"2006-01-01","label":"NGS reads got shorter","detail":"The Solexa/Illumina Genome Analyzer (2006) read only ~35 bp — a deliberate step DOWN from Sanger, trading length for massive parallelism. Long-read length would not recover for years."},{"date":"2011-01-01","label":"Long reads arrive","detail":"PacBio RS (2011) brought multi-kilobase single-molecule reads, reversing the downward trend and enabling genome assembly across repeats."},{"date":"2018-01-01","label":"Ultra-long nanopore","detail":"An ~882 kb Oxford Nanopore read (2018) pushed the ceiling roughly three orders of magnitude beyond Sanger, unlocking telomere-to-telomere assembly."}]},{"id":"ena-platform-mgi","label":"MGI / BGI runs in ENA","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"comparisonOnly":true,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22BGISEQ%22%20OR%20instrument_platform%3D%22DNBSEQ%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"forecast":false,"series":[["2016-12-31",38],["2018-12-31",5919],["2020-12-31",31852],["2022-12-31",113787],["2024-12-31",281236],["2026-06-25",646585],["2026-07-06",654399],["2026-07-13",656157],["2026-07-20",659863]],"note":"Cumulative public ENA runs on MGI / BGI DNB platforms (BGISEQ + DNBSEQ instrument_platform values combined). Shown in the platform-mix comparison card; counted live from the ENA Portal API.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count (BGISEQ OR DNBSEQ)","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22BGISEQ%22%20OR%20instrument_platform=%22DNBSEQ%22&format=json"}]},{"id":"ena-platform-iontorrent","label":"Ion Torrent runs in ENA","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"comparisonOnly":true,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22ION_TORRENT%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"forecast":false,"series":[["2016-12-31",28653],["2018-12-31",73966],["2020-12-31",158810],["2022-12-31",375308],["2024-12-31",487513],["2026-06-25",538844],["2026-07-06",539380],["2026-07-13",539654],["2026-07-20",541909]],"note":"Cumulative public ENA runs on Thermo Fisher Ion Torrent semiconductor platforms (PGM, Proton, S5). Shown in the platform-mix comparison card; counted live from the ENA Portal API.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count by instrument_platform","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22ION_TORRENT%22&format=json"}]},{"id":"ena-platform-454","label":"454 runs in ENA (retired)","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"comparisonOnly":true,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22LS454%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"forecast":false,"series":[["2016-12-31",282039],["2018-12-31",343950],["2020-12-31",380269],["2022-12-31",407046],["2024-12-31",416098],["2026-06-25",418573],["2026-07-06",418694],["2026-07-20",419088]],"note":"Cumulative public ENA runs on the retired Roche/454 pyrosequencing platform. The curve is now nearly flat — 454 sequencing was discontinued by ~2016, so this is a legacy archive that barely grows. Shown in the platform-mix comparison card; counted live from the ENA Portal API.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count by instrument_platform","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22LS454%22&format=json"}]},{"id":"ena-platform-sanger","label":"Sanger capillary runs in ENA","category":"sequencing-technology","source":"EMBL-EBI / European Nucleotide Archive","unit":"runs","tier":"secondary","flagship":false,"comparisonOnly":true,"cadence":"weekly","scale":"log","fetch":{"url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform%3D%22CAPILLARY%22&format=json","method":"GET","parse":{"type":"json","path":"count","cast":"int"},"auto":true,"appendPolicy":"on-change"},"release":null,"forecast":false,"series":[["2016-12-31",1010],["2018-12-31",3256],["2020-12-31",335465],["2022-12-31",347570],["2024-12-31",358879],["2026-06-25",365688],["2026-07-06",365722]],"note":"Cumulative public ENA runs on first-generation Sanger capillary instruments (the chemistry the original Human Genome Project used). The step in 2019-2020 reflects a large retrospective deposition. Shown in the platform-mix comparison card; counted live from the ENA Portal API.","sources":[{"label":"EMBL-EBI ENA Portal API — read_run count by instrument_platform","url":"https://www.ebi.ac.uk/ena/portal/api/count?result=read_run&query=instrument_platform=%22CAPILLARY%22&format=json"}]}]}