The Human Pangenome Project: a global resource to map genomic diversity.


Journal

Nature
ISSN: 1476-4687
Titre abrégé: Nature
Pays: England
ID NLM: 0410462

Informations de publication

Date de publication:
04 2022
Historique:
received: 30 08 2021
accepted: 01 03 2022
entrez: 21 4 2022
pubmed: 22 4 2022
medline: 23 4 2022
Statut: ppublish

Résumé

The human reference genome is the most widely used resource in human genetics and is due for a major update. Its current structure is a linear composite of merged haplotypes from more than 20 people, with a single individual comprising most of the sequence. It contains biases and errors within a framework that does not represent global human genomic variation. A high-quality reference with global representation of common variants, including single-nucleotide variants, structural variants and functional elements, is needed. The Human Pangenome Reference Consortium aims to create a more sophisticated and complete human reference genome with a graph-based, telomere-to-telomere representation of global genomic diversity. Here we leverage innovations in technology, study design and global partnerships with the goal of constructing the highest-possible quality human pangenome reference. Our goal is to improve data representation and streamline analyses to enable routine assembly of complete diploid genomes. With attention to ethical frameworks, the human pangenome reference will contain a more accurate and diverse representation of global genomic variation, improve gene-disease association studies across populations, expand the scope of genomics research to the most repetitive and polymorphic regions of the genome, and serve as the ultimate genetic resource for future biomedical research and precision medicine.

Identifiants

pubmed: 35444317
doi: 10.1038/s41586-022-04601-8
pii: 10.1038/s41586-022-04601-8
pmc: PMC9402379
mid: NIHMS1828150
doi:

Types de publication

Journal Article Review Research Support, N.I.H., Intramural Research Support, N.I.H., Extramural

Langues

eng

Sous-ensembles de citation

IM

Pagination

437-446

Subventions

Organisme : NHGRI NIH HHS
ID : U01 HG010961
Pays : United States
Organisme : NHGRI NIH HHS
ID : U41 HG010972
Pays : United States
Organisme : NHGRI NIH HHS
ID : U01 HG010973
Pays : United States
Organisme : NHGRI NIH HHS
ID : U01 HG010963
Pays : United States
Organisme : NHGRI NIH HHS
ID : U01 HG010971
Pays : United States

Informations de copyright

© 2022. Springer Nature Limited.

Références

International Human Genome Sequencing Consortium. Initial sequencing and analysis of the human genome. Nature 409, 860–921 (2001).
doi: 10.1038/35057062
Venter, J. C. et al. The sequence of the human genome. Science 291, 1304–1351 (2001).
doi: 10.1126/science.1058040 pubmed: 11181995
Gibbs, R. A. The Human Genome Project changed everything. Nat. Rev. Genet. 21, 575–576 (2020).
doi: 10.1038/s41576-020-0275-3 pubmed: 32770171 pmcid: 7413016
Venter, J. C. et al. The sequence of the human genome. Science 291, 1304–1351 (2001).
doi: 10.1126/science.1058040 pubmed: 11181995
Green, R. E. et al. A draft sequence of the Neandertal genome. Science 328, 710–722 (2010).
doi: 10.1126/science.1188021 pubmed: 20448178 pmcid: 5100745
Sherman, R. M. & Salzberg, S. L. Pan-genomics in the human genome era. Nat. Rev. Genet. 21, 243–254 (2020).
doi: 10.1038/s41576-020-0210-7 pubmed: 32034321 pmcid: 7752153
Rhie, A. et al. Towards complete and error-free genome assemblies of all vertebrate species. Nature 592, 737–746 (2021).
doi: 10.1038/s41586-021-03451-0 pubmed: 33911273 pmcid: 8081667
Need, A. C. & Goldstein, D. B. Next generation disparities in human genomics: concerns and remedies. Trends Genet. 25, 489–494 (2009).
doi: 10.1016/j.tig.2009.09.012 pubmed: 19836853
Schneider, V. A. et al. Evaluation of GRCh38 and de novo haploid genome assemblies demonstrates the enduring quality of the reference assembly. Genome Res. 27, 849–864 (2017).
doi: 10.1101/gr.213611.116 pubmed: 28396521 pmcid: 5411779
Bustamante, C. D., Burchard, E. G. & De la Vega, F. M. Genomics for the world. Nature 475, 163–165 (2011). Emphasizes the importance of reference data from ancestral and diverse genomes, as well as stating that researchers should invest time and money into education and outreach to explain why studying global (and local) health is so important.
doi: 10.1038/475163a pubmed: 21753830 pmcid: 3708540
Miga, K. H. & Wang, T. The need for a human pangenome reference sequence. Annu. Rev. Genomics Hum. Genet. 22, 81–102 (2021).
doi: 10.1146/annurev-genom-120120-081921 pubmed: 33929893 pmcid: 8410644
International Human Genome Sequencing Consortium. Finishing the euchromatic sequence of the human genome. Nature 431, 931–945 (2004).
doi: 10.1038/nature03001
Garrison, E. et al. Variation graph toolkit improves read mapping by representing genetic variation in the reference. Nat. Biotechnol. 36, 875–879 (2018). A model for presenting genomes that aims to improve read mapping by representing genetic variation in the reference.
doi: 10.1038/nbt.4227 pubmed: 30125266 pmcid: 6126949
Martiniano, R., Garrison, E., Jones, E. R., Manica, A. & Durbin, R. Removing reference bias and improving indel calling in ancient DNA data analysis by mapping to a sequence variation graph. Genome Biol. 21, 250 (2020).
doi: 10.1186/s13059-020-02160-7 pubmed: 32943086 pmcid: 7499850
Alkan, C., Coe, B. P. & Eichler, E. E. Genome structural variation discovery and genotyping. Nat. Rev. Genet. 12, 363–376 (2011).
doi: 10.1038/nrg2958 pubmed: 21358748 pmcid: 4108431
Sedlazeck, F. J. et al. Accurate detection of complex structural variations using single-molecule sequencing. Nat. Methods 15, 461–468 (2018).
doi: 10.1038/s41592-018-0001-7 pubmed: 29713083 pmcid: 5990442
Sudmant, P. H. et al. An integrated map of structural variation in 2,504 human genomes. Nature 526, 75–81 (2015).
doi: 10.1038/nature15394 pubmed: 26432246 pmcid: 4617611
Chaisson, M. J. P. et al. Multi-platform discovery of haplotype-resolved structural variation in human genomes. Nat. Commun. 10, 1784 (2019).
doi: 10.1038/s41467-018-08148-z pubmed: 30992455 pmcid: 6467913
Li, R. et al. Building the sequence map of the human pan-genome. Nat. Biotechnol. 28, 57–63 (2010).
doi: 10.1038/nbt.1596 pubmed: 19997067
Miga, K. H. et al. Telomere-to-telomere assembly of a complete human X chromosome. Nature 585, 79–84 (2020). The sequence of the first complete human chromosome.
doi: 10.1038/s41586-020-2547-7 pubmed: 32663838 pmcid: 7484160
Logsdon, G. A. et al. The structure, function and evolution of a complete human chromosome 8. Nature 593, 101–107 (2021).
doi: 10.1038/s41586-021-03420-7 pubmed: 33828295 pmcid: 8099727
Nurk, S. et al. The complete sequence of a human genome. Preprint at bioRxiv https://doi.org/10.1101/2021.05.26.445798 (2021). The first complete genome assembly issued from the T2T Consortium, which closed all remaining gaps in the GRCh38, including all acrocentric short arms, segmental duplications and human centromeric regions.
Sirugo, G., Williams, S. M. & Tishkoff, S. A. The missing diversity in human genetic studies. Cell 177, 26–31 (2019).
doi: 10.1016/j.cell.2019.02.048 pubmed: 30901543 pmcid: 7380073
Tettelin, H. et al. Genome analysis of multiple pathogenic isolates of Streptococcus agalactiae: implications for the microbial “pan-genome”. Proc. Natl Acad. Sci. USA 102, 13950–13955 (2005).
doi: 10.1073/pnas.0506758102 pubmed: 16172379 pmcid: 1216834
Vernikos, G., Medini, D., Riley, D. R. & Tettelin, H. Ten years of pan-genome analyses. Curr. Opin. Microbiol. 23, 148–154 (2015).
doi: 10.1016/j.mib.2014.11.016 pubmed: 25483351
Computational Pan-Genomics Consortium. Computational pan-genomics: status, promises and challenges. Brief Bioinform. 19, 118–135 (2018).
Eizenga, J. M. et al. Pangenome graphs. Annu. Rev. Genomics Hum. Genet. 21, 139–162 (2020).
doi: 10.1146/annurev-genom-120219-080406 pubmed: 32453966 pmcid: 8006571
Rehm, H. L. et al. ClinGen—the clinical genome resource. N. Engl. J. Med. 372, 2235–2242 (2015).
doi: 10.1056/NEJMsr1406261 pubmed: 26014595 pmcid: 4474187
Genomes Project Consortium. et al. A global reference for human genetic variation. Nature 526, 68–74 (2015).
doi: 10.1038/nature15393
Popejoy, A. B. et al. The clinical imperative for inclusivity: race, ethnicity, and ancestry (REA) in genomics. Hum. Mutat. 39, 1713–1720 (2018).
doi: 10.1002/humu.23644 pubmed: 30311373 pmcid: 6188707
Popejoy, A. B. et al. Clinical genetics lacks standard definitions and protocols for the collection and use of diversity measures. Am. J. Hum. Genet. 107, 72–82 (2020).
doi: 10.1016/j.ajhg.2020.05.005 pubmed: 32504544 pmcid: 7332657
Bonham, V. L. et al. Physicians’ attitudes toward race, genetics, and clinical medicine. Genet. Med. 11, 279–286 (2009).
doi: 10.1097/GIM.0b013e318195aaf4 pubmed: 19265721 pmcid: 3065019
Race, Ethnicity & Genetics Working Group. The use of racial, ethnic, and ancestral categories in human genetics research. Am. J. Hum. Genet. 77, 519–532 (2005).
doi: 10.1086/491747
Dodson, M. & Williamson, R. Indigenous peoples and the morality of the Human Genome Diversity Project. J. Med. Ethics 25, 204–208 (1999).
doi: 10.1136/jme.25.2.204 pubmed: 10226929 pmcid: 479208
Couzin-Frankel, J. Ethics. DNA returned to tribe, raising questions about consent. Science 328, 558 (2010).
doi: 10.1126/science.328.5978.558 pubmed: 20430983
Dukepoo, F. C. The trouble with the Human Genome Diversity Project. Mol. Med. Today 4, 242–243 (1998).
doi: 10.1016/S1357-4310(98)01282-9 pubmed: 9679240
Fox, K. The illusion of inclusion—the “All of Us” research program and Indigenous peoples’ DNA. N. Engl. J. Med. 383, 411–413 (2020).
doi: 10.1056/NEJMp1915987 pubmed: 32726527
Devaney, S. A., Malerba, L. & Manson, S. M. The “All of Us” program and Indigenous peoples. N. Engl. J. Med. 383, 1892 (2020).
doi: 10.1056/NEJMc2028907 pubmed: 33211939
Hudson, M. et al. Rights, interests and expectations: Indigenous perspectives on unrestricted access to genomic data. Nat. Rev. Genet. 21, 377–384 (2020).
doi: 10.1038/s41576-020-0228-x pubmed: 32251390
Carroll, S. R., Herczog, E., Hudson, M., Russell, K. & Stall, S. Operationalizing the CARE and FAIR principles for Indigenous data futures. Sci. Data 8, 108 (2021).
doi: 10.1038/s41597-021-00892-0 pubmed: 33863927 pmcid: 8052430
Wilkinson, M. D. et al. The FAIR guiding principles for scientific data management and stewardship. Sci. Data 3, 160018 (2016).
doi: 10.1038/sdata.2016.18 pubmed: 26978244 pmcid: 4792175
Genome in a Bottle. NIST https://www.nist.gov/programs-projects/genome-bottle (updated 16 February 2022).
Jarvis, E. D. et al. Automated assembly of high-quality diploid human reference genomes. Preprint at bioRxiv https://doi.org/10.1101/2022.03.06.483034 (2021).
Cheng, H., Concepcion, G. T., Feng, X., Zhang, H. & Li, H. Haplotype-resolved de novo assembly using phased assembly graphs with HiFiasm. Nat. Methods 18, 170–175 (2021). HiFiasm is a haplotype-resolved assembler specifically designed for PacBio HiFi reads that aims to represent haplotype information in a phased assembly graph.
doi: 10.1038/s41592-020-01056-5 pubmed: 33526886 pmcid: 7961889
Nurk, S. et al. HiCanu: accurate assembly of segmental duplications, satellites, and allelic variants from high-fidelity long reads. Genome Res. 30, 1291–1305 (2020).
doi: 10.1101/gr.263566.120 pubmed: 32801147 pmcid: 7545148
Schatz, M. C. et al. Inverting the model of genomics data sharing with the NHGRI Genomic Data Science Analysis, Visualization, and Informatics Lab-space. Cell Genom. 2, 100085 (2022). The AnVIL platform provides scalable solutions for genomic data access, analysis and education.
Li, H., Feng, X. & Chu, C. The design and construction of reference pangenome graphs with Minigraph. Genome Biol. 21, 265 (2020). The Minigraph toolkit has been used to efficiently construct a pangenome graph, which is useful for mapping and constructing graphs that encode structural variation.
doi: 10.1186/s13059-020-02168-z pubmed: 33066802 pmcid: 7568353
Li, H. et al. The Sequence Alignment/Map format and SAMtools. Bioinformatics 25, 2078–2079 (2009).
doi: 10.1093/bioinformatics/btp352 pubmed: 19505943 pmcid: 2723002
Danecek, P. et al. The variant call format and VCFtools. Bioinformatics 27, 2156–2158 (2011).
doi: 10.1093/bioinformatics/btr330 pubmed: 21653522 pmcid: 3137218
Rosen, Y., Eizenga, J. & Paten, B. Modelling haplotypes with respect to reference cohort variation graphs. Bioinformatics 33, i118–i123 (2017).
doi: 10.1093/bioinformatics/btx236 pubmed: 28881971 pmcid: 5870562
Ebert, P. et al. Haplotype-resolved diverse human genomes and integrated analysis of structural variation. Science 372, eabf7117 (2021). The use of long-read data from 64 human genomes to predict structural variants and the patterns of variation across diverse populations.
doi: 10.1126/science.abf7117 pubmed: 33632895 pmcid: 8026704
Abel, H. J. et al. Mapping and characterization of structural variation in 17,795 human genomes. Nature 583, 83–89 (2020).
doi: 10.1038/s41586-020-2371-0 pubmed: 32460305 pmcid: 7547914
Li, H. Minimap2: pairwise alignment for nucleotide sequences. Bioinformatics 34, 3094–3100 (2018).
doi: 10.1093/bioinformatics/bty191 pubmed: 29750242 pmcid: 6137996
Paten, B. et al. Cactus: algorithms for genome multiple sequence alignment. Genome Res. 21, 1512–1528 (2011). Cactus is a highly accurate, reference-free multiple genome alignment program that is useful for studying general rearrangement and copy number variation.
doi: 10.1101/gr.123356.111 pubmed: 21665927 pmcid: 3166836
Pangenome Graph Builder. GitHub https://github.com/pangenome/pggb (2022).
O’Leary, N. A. et al. Reference sequence (RefSeq) database at NCBI: current status, taxonomic expansion, and functional annotation. Nucleic Acids Res. 44, D733–D745 (2016).
doi: 10.1093/nar/gkv1189 pubmed: 26553804
Frankish, A. et al. GENCODE reference annotation for the human and mouse genomes. Nucleic Acids Res. 47, D766–D773 (2019).
doi: 10.1093/nar/gky955 pubmed: 30357393
Spooner, W. et al. Haplosaurus computes protein haplotypes for use in precision drug design. Nat. Commun. 9, 4128 (2018).
doi: 10.1038/s41467-018-06542-1 pubmed: 30297836 pmcid: 6175845
Arita, M., Karsch-Mizrachi, I. & Cochrane, G. The international nucleotide sequence database collaboration. Nucleic Acids Res. 49, D121–D124 (2021).
doi: 10.1093/nar/gkaa967 pubmed: 33166387
Clarke, L. et al. The 1000 Genomes Project: data management and community access. Nat. Methods 9, 459–462 (2012).
doi: 10.1038/nmeth.1974 pubmed: 22543379 pmcid: 3340611
Clarke, L. et al. The International Genome Sample Resource (IGSR): a worldwide collection of genome variation incorporating the 1000 Genomes Project data. Nucleic Acids Res. 45, D854–D859 (2017).
doi: 10.1093/nar/gkw829 pubmed: 27638885
Courtot, M. et al. BioSamples database: an updated sample metadata hub. Nucleic Acids Res. 47, D1172–D1178 (2019).
doi: 10.1093/nar/gky1061 pubmed: 30407529
Vollger, M. R. et al. Segmental duplications and their variation in a complete human genome. Preprint at bioRxiv https://doi.org/10.1101/2021.05.26.445678 (2021).
Aganezov, S. et al. A complete reference genome improves analysis of human genetic variation. Preprint at bioRxiv https://doi.org/10.1101/2021.07.12.452063 (2021). The importance of complete T2T genomes in novel variant discovery and of offering major improvements of variant calls within clinically relevant genes are highlighted.
Miller, D. E. et al. Targeted long-read sequencing identifies missing disease-causing variation. Am. J. Hum. Genet. 108, 1436–1449 (2021).
doi: 10.1016/j.ajhg.2021.06.006 pubmed: 34216551 pmcid: 8387463
Logsdon, G. A., Vollger, M. R. & Eichler, E. E. Long-read human genome sequencing and its applications. Nat. Rev. Genet. 21, 597–614 (2020).
doi: 10.1038/s41576-020-0236-x pubmed: 32504078 pmcid: 7877196
Kim, D. et al. The architecture of SARS-CoV-2 transcriptome. Cell 181, 914–921.e90 (2020).
doi: 10.1016/j.cell.2020.04.011 pubmed: 32330414 pmcid: 7179501
Zhou, P. et al. A pneumonia outbreak associated with a new coronavirus of probable bat origin. Nature 579, 270–273 (2020).
doi: 10.1038/s41586-020-2012-7 pubmed: 32015507 pmcid: 7095418
Toh, C. & Brody, J. P. Evaluation of a genetic risk score for severity of COVID-19 using human chromosomal-scale length variation. Hum. Genomics 14, 36 (2020).
doi: 10.1186/s40246-020-00288-y pubmed: 33036646 pmcid: 7546598
Zeberg, H. & Paabo, S. The major genetic risk factor for severe COVID-19 is inherited from Neanderthals. Nature 587, 610–612 (2020).
doi: 10.1038/s41586-020-2818-3 pubmed: 32998156
Okubo, K., Sugawara, H., Gojobori, T. & Tateno, Y. DDBJ in preparation for overview of research activities behind data submissions. Nucleic Acids Res. 34, D6–D9 (2006).
doi: 10.1093/nar/gkj111 pubmed: 16381940
Kent, W. J. et al. The human genome browser at UCSC. Genome Res. 12, 996–1006 (2002).
doi: 10.1101/gr.229102 pubmed: 12045153 pmcid: 186604
Navarro Gonzalez, J. et al. The UCSC Genome Browser database: 2021 update. Nucleic Acids Res. 49, D1046–D1057 (2021).
doi: 10.1093/nar/gkaa1070 pubmed: 33221922
Stalker, J. et al. The Ensembl web site: mechanics of a genome browser. Genome Res. 14, 951–955 (2004).
doi: 10.1101/gr.1863004 pubmed: 15123591 pmcid: 479125
Howe, K. L. et al. Ensembl 2021. Nucleic Acids Res. 49, D884–D891 (2021).
doi: 10.1093/nar/gkaa942 pubmed: 33137190
Zhou, X. et al. The Human Epigenome Browser at Washington University. Nat. Methods 8, 989–990 (2011).
doi: 10.1038/nmeth.1772 pubmed: 22127213 pmcid: 3552640
Li, D., Hsu, S., Purushotham, D., Sears, R. L. & Wang, T. WashU Epigenome Browser update 2019. Nucleic Acids Res. 47, W158–W165 (2019).
doi: 10.1093/nar/gkz348 pubmed: 31165883 pmcid: 6602459
Popejoy, A. B. & Fullerton, S. M. Genomics is failing on diversity. Nature 538, 161–164 (2016). Analysis of sample descriptions included in the genome-wide association study catalogue indicates that some populations are still under-represented and left behind in studies of genomic medicine.
doi: 10.1038/538161a pubmed: 27734877 pmcid: 5089703
Mills, M. C. & Rahal, C. A scientometric review of genome-wide association studies. Commun. Biol. 2, 9 (2019).
doi: 10.1038/s42003-018-0261-x pubmed: 30623105 pmcid: 6323052
Lieberman-Aiden, E. et al. Comprehensive mapping of long-range interactions reveals folding principles of the human genome. Science 326, 289–293 (2009).
doi: 10.1126/science.1181369 pubmed: 19815776 pmcid: 2858594
Ulahannan, N. et al. Nanopore sequencing of DNA concatemers reveals higher-order features of chromatin structure. Preprint at bioRxiv https://doi.org/10.1101/833590 (2019).
Liu, B., Guo, H., Brudno, M. & Wang, Y. deBGA: read alignment with de Bruijn graph-based seed and extension. Bioinformatics 32, 3224–3232 (2016).
doi: 10.1093/bioinformatics/btw371 pubmed: 27378303
Limasset, A., Cazaux, B., Rivals, E. & Peterlongo, P. Read mapping on de Bruijn graphs. BMC Bioinformatics. 17, 237 (2016).
doi: 10.1186/s12859-016-1103-9 pubmed: 27306641 pmcid: 4910249
Heydari, M., Miclotte, G., Van de Peer, Y. & Fostier, J. BrownieAligner: accurate alignment of Illumina sequencing data to de Bruijn graphs. BMC Bioinformatics 19, 311 (2018).
doi: 10.1186/s12859-018-2319-7 pubmed: 30180801 pmcid: 6122196
1001 Genomes. GenomeMapper. 1001 Genomes https://www.1001genomes.org/software/genomemapper_graph.html (accessed 2021).
Kim, D., Paggi, J. M., Park, C., Bennett, C. & Salzberg, S. L. Graph-based genome alignment and genotyping with HISAT2 and HISAT-genotype. Nat. Biotechnol. 37, 907–915 (2019).
doi: 10.1038/s41587-019-0201-4 pubmed: 31375807 pmcid: 7605509
Hickey, G. et al. Genotyping structural variants in pangenome graphs using the vg toolkit. Genome Biol. 21, 35 (2020).
doi: 10.1186/s13059-020-1941-7 pubmed: 32051000 pmcid: 7017486
Rautiainen, M. & Marschall, T. GraphAligner: rapid and versatile sequence-to-graph alignment. Genome Biol. 21, 253 (2020).
doi: 10.1186/s13059-020-02157-2 pubmed: 32972461 pmcid: 7513500
Jain, C., Misra, S., Zhang, H., Dilthey, A. & Aluru, S. Accelerating sequence alignment to graphs. IEEE Int. Parallel and Distributed Processing Symp. (IPDPS) 451–461 (2019).
Dvorkina, T., Antipov, D., Korobeynikov, A. & Nurk, S. SPAligner: alignment of long diverged molecular sequences to assembly graphs. BMC Bioinformatics 21, 306 (2020).
doi: 10.1186/s12859-020-03590-7 pubmed: 32703258 pmcid: 7379835
Mokveld, T., Linthorst, J., Al-Ars, Z., Holstege, H. & Reinders, M. CHOP: haplotype-aware path indexing in population graphs. Genome Biol. 21, 65 (2020).
doi: 10.1186/s13059-020-01963-y pubmed: 32160922 pmcid: 7066762
Ghaffaari, A. & Marschall, T. Fully-sensitive seed finding in sequence graphs using a hybrid index. Bioinformatics 35, i81–i89 (2019).
doi: 10.1093/bioinformatics/btz341 pubmed: 31510650 pmcid: 6612829
Wick, R. R., Schultz, M. B., Zobel, J. & Holt, K. E. Bandage: interactive visualization of de novo genome assemblies. Bioinformatics 31, 3350–3352 (2015).
doi: 10.1093/bioinformatics/btv383 pubmed: 26099265 pmcid: 4595904
Gonnella, G., Niehus, N. & Kurtz, S. GfaViz: flexible and interactive visualization of GFA sequence graphs. Bioinformatics 35, 2853–2855 (2019).
doi: 10.1093/bioinformatics/bty1046 pubmed: 30596893
Kunyavskaya, O. & Prjibelski, A. D. SGTK: a toolkit for visualization and assessment of scaffold graphs. Bioinformatics 35, 2303–2305 (2019).
doi: 10.1093/bioinformatics/bty956 pubmed: 30475983
Mikheenko, A. & Kolmogorov, M. Assembly Graph Browser: interactive visualization of assembly graphs. Bioinformatics 35, 3476–3478 (2019).
doi: 10.1093/bioinformatics/btz072 pubmed: 30715194
Beyer, W. et al. Sequence tube maps: making graph genomes intuitive to commuters. Bioinformatics 35, 5318–5320 (2019).
doi: 10.1093/bioinformatics/btz597 pubmed: 31368484 pmcid: 6954646
Yokoyama, T. T., Sakamoto, Y., Seki, M., Suzuki, Y. & Kasahara, M. MoMI-G: modular multi-scale integrated genome graph browser. BMC Bioinformatics 20, 548 (2019).
doi: 10.1186/s12859-019-3145-2 pubmed: 31690272 pmcid: 6833150
ODGI. GitHub https://github.com/pangenome/odgi (2021).
Shlemov, A. & Korobeynikov, A. in Algorithms for Computational Biology (eds Holmes, I., Martín-Vide, C. & Vega-Rodríguez, M. A.) 80–94 (Springer, 2019).
Ebler, J. et al. Pangenome-based genome inference. Preprint at bioRxiv https://doi.org/10.1101/2020.11.11.378133 (2020).
Leggett, R. M. et al. Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de Bruijn graphs. PLoS ONE 8, e60058 (2013).
doi: 10.1371/journal.pone.0060058 pubmed: 23536903 pmcid: 3607606
Sibbesen, J. A. et al. Accurate genotyping across variant classes and lengths using variant graphs. Nat. Genet. 50, 1054–1059 (2018).
Chen, S. et al. Paragraph: a graph-based structural variant genotyper for short-read sequence data. Genome Biol. 20, 291 (2019).
doi: 10.1186/s13059-019-1909-7 pubmed: 31856913 pmcid: 6921448
Eggertsson, H. P. et al. GraphTyper2 enables population-scale genotyping of structural variation using pangenome graphs. Nat. Commun. 10, 5402 (2019).
doi: 10.1038/s41467-019-13341-9 pubmed: 31776332 pmcid: 6881350

Auteurs

Ting Wang (T)

Department of Genetics, Washington University School of Medicine, St. Louis, MO, USA. twang@wustl.edu.
Edison Family Center for Genome Sciences and Systems Biology, Washington University School of Medicine, St. Louis, MO, USA. twang@wustl.edu.
McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA. twang@wustl.edu.

Lucinda Antonacci-Fulton (L)

McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Kerstin Howe (K)

Wellcome Sanger Institute, Cambridge, UK.

Heather A Lawson (HA)

Department of Genetics, Washington University School of Medicine, St. Louis, MO, USA.

Julian K Lucas (JK)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA.

Adam M Phillippy (AM)

Genome Informatics Section, National Human Genome Research Institute, Bethesda, MD, USA.

Alice B Popejoy (AB)

Epidemiology Division, Department of Public Health Sciences, University of California, Davis, CA, USA.

Mobin Asri (M)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA.

Caryn Carson (C)

Department of Genetics, Washington University School of Medicine, St. Louis, MO, USA.
Edison Family Center for Genome Sciences and Systems Biology, Washington University School of Medicine, St. Louis, MO, USA.
McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Mark J P Chaisson (MJP)

Department of Quantitative and Computational Biology, University of Southern California, Los Angeles, CA, USA.

Xian Chang (X)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA.

Robert Cook-Deegan (R)

Arizona State University, Barrett & O'Connor Washington Center, Washington DC, USA.

Adam L Felsenfeld (AL)

National Institutes of Health (NIH)-National Human Genome Research Institute, Bethesda, MD, USA.

Robert S Fulton (RS)

McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Erik P Garrison (EP)

Department of Genetics, Genomics and Informatics, University of Tennessee Health Science Center, Memphis, TN, USA.

Nanibaa' A Garrison (NA)

Institute for Society & Genetics, College of Letters and Science, University of California, Los Angeles, Los Angeles, CA, USA.
Institute for Precision Health, David Geffen School of Medicine, University of California, Los Angeles, Los Angeles, CA, USA.
Division of General Internal Medicine & Health Services Research, David Geffen School of Medicine, University of California, Los Angeles, Los Angeles, CA, USA.

Tina A Graves-Lindsay (TA)

McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Hanlee Ji (H)

Department of Medicine, Stanford University, School of Medicine, Stanford, CA, USA.

Eimear E Kenny (EE)

Department of Genetics and Genomic Science, Icahn School of Medicine at Mount Sinai, New York, NY, USA.
Department of Medicine, Icahn School of Medicine at Mount Sinai, New York, NY, USA.
Institute for Genomic Health, Icahn School of Medicine at Mount Sinai, New York, NY, USA.

Barbara A Koenig (BA)

Program in Bioethics and Institute for Human Genetics, University of California, San Francisco, San Francisco, CA, USA.

Daofeng Li (D)

Department of Genetics, Washington University School of Medicine, St. Louis, MO, USA.
Edison Family Center for Genome Sciences and Systems Biology, Washington University School of Medicine, St. Louis, MO, USA.
McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Tobias Marschall (T)

Heinrich Heine University, Medical Faculty, Institute for Medical Biometry and Bioinformatics, Düsseldorf, Germany.

Joshua F McMichael (JF)

McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Adam M Novak (AM)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA.

Deepak Purushotham (D)

Department of Genetics, Washington University School of Medicine, St. Louis, MO, USA.
Edison Family Center for Genome Sciences and Systems Biology, Washington University School of Medicine, St. Louis, MO, USA.
McDonnell Genome Institute, Washington University School of Medicine, St. Louis, MO, USA.

Valerie A Schneider (VA)

National Center for Biotechnology Information (NCBI), National Library of Medicine, Bethesda, MD, USA.

Baergen I Schultz (BI)

National Institutes of Health (NIH)-National Human Genome Research Institute, Bethesda, MD, USA.

Michael W Smith (MW)

National Institutes of Health (NIH)-National Human Genome Research Institute, Bethesda, MD, USA.

Heidi J Sofia (HJ)

National Institutes of Health (NIH)-National Human Genome Research Institute, Bethesda, MD, USA.

Tsachy Weissman (T)

Department of Electrical Engineering, Stanford University, Stanford, CA, USA.

Paul Flicek (P)

European Molecular Biology Laboratory, European Bioinformatics Institute, Cambridge, UK. flicek@ebi.ac.uk.

Heng Li (H)

Department of Biomedical Informatics, Harvard Medical School, Boston, MA, USA. hli@jimmy.harvard.edu.
Department of Data Science, Dana-Farber Cancer Institute, Boston, MA, USA. hli@jimmy.harvard.edu.

Karen H Miga (KH)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA. khmiga@ucsc.edu.

Benedict Paten (B)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA. bpaten@ucsc.edu.

Erich D Jarvis (ED)

Vertebrate Genome Lab and and Laboratory of Neurogenetics of Language, The Rockefeller University, New York, NY, USA. ejarvis@rockefeller.edu.
Howard Hughes Medical Institute, Chevy Chase, MD, USA. ejarvis@rockefeller.edu.

Ira M Hall (IM)

Yale School of Medicine, New Haven, CT, USA. ira.hall@yale.edu.

Evan E Eichler (EE)

Department of Genome Sciences, University of Washington School of Medicine, Seattle, WA, USA. eee@gs.washington.edu.
Howard Hughes Medical Institute, University of Washington, Seattle, WA, USA. eee@gs.washington.edu.

David Haussler (D)

UC Santa Cruz Genomics Institute, University of California, Santa Cruz, CA, USA. haussler@ucsc.edu.
Howard Hughes Medical Institute, University of California, Santa Cruz, CA, USA. haussler@ucsc.edu.

Articles similaires

Genome, Chloroplast Phylogeny Genetic Markers Base Composition High-Throughput Nucleotide Sequencing

[Redispensing of expensive oral anticancer medicines: a practical application].

Lisanne N van Merendonk, Kübra Akgöl, Bastiaan Nuijen
1.00
Humans Antineoplastic Agents Administration, Oral Drug Costs Counterfeit Drugs

Smoking Cessation and Incident Cardiovascular Disease.

Jun Hwan Cho, Seung Yong Shin, Hoseob Kim et al.
1.00
Humans Male Smoking Cessation Cardiovascular Diseases Female
Humans United States Aged Cross-Sectional Studies Medicare Part C

Classifications MeSH