fe0627da07b9f989bc642d824af381968163d9a8 lrnassar Wed Jul 8 13:40:04 2026 -0700 varFreqs: move ALFA to Global reference panels; add shared-sample bias caveat to description pages. refs #36642 Per Max's post-release feedback on the ticket: ALFA is not a European cohort, it is an NCBI aggregator of dbGaP studies from many sources. Moved from Europe to Global reference panels in the newsarch entry. Also adds a shared-sample bias caveat to the Pooled allele frequency section of both varFreqsAffected.html and varFreqsBackground.html. The background page cites the concrete overlaps (1000 Genomes in both gnomAD HGDP+1kG and HRC; HGDP/SGDP; AllOfUs/TOPMed; ALFA aggregating dbGaP studies used elsewhere). The affected page cites the SPARK WGS cohort being a subset of SPARK WES. Both explain that pooled AN is inflated and pooled AF is skewed toward the frequency in the shared subset, and point users to the per-cohort AC/AF/AN fields for unbiased single-cohort numbers. diff --git src/hg/htdocs/goldenPath/newsarch.html src/hg/htdocs/goldenPath/newsarch.html index 21b1e0272da..85b44b9edd4 100755 --- src/hg/htdocs/goldenPath/newsarch.html +++ src/hg/htdocs/goldenPath/newsarch.html @@ -200,35 +200,35 @@

The container track pulls together cohorts from across the world. A high-level summary of the regions and contributing projects is shown below; a complete table with per-cohort sample counts, data types, sub-populations and download status is on the container track description page.

- + - +
RegionContributing cohorts
East AsiaToMMo, GenomeAsia, NPM, KOVA, WBBC, ChinaMAP, TPMI
South AsiaIndiGen, GenomeIndia
AfricaTishkoff
AmericasAllOfUs, TOPMed, ABraOM, Mexico Biobank
EuropeFinnGen, SweGen, GoNL, HRC, UK Biobank, ALFA
EuropeFinnGen, SweGen, GoNL, HRC, UK Biobank
Middle EastSaudi
OceaniaMGRB
Disease cohortsSFARI SPARK (WES + WGS), SCHEMA, GREGoR, GA4K
Global reference panelsSGDP, gnomAD HGDP+1kG
Global reference panelsSGDP, gnomAD HGDP+1kG, ALFA
Long-read / linked-readGA4K, CoLoRSdb, SVatalog

License restrictions on some sources limit redistribution; see the container track description page for per-cohort details.

We plan to continue updating this track as more population-scale allele-frequency datasets become available. If you are involved with a project that publishes variant frequencies and would like to contribute, please reach out.