HUPAN: a pan-genome analysis pipeline for human genomes

Zhongqu Duan,Yuyang Qiao,Jinyuan Lu,Huimin Lu,Wenmin Zhang,Fazhe Yan,Chen Sun,Zhiqiang Hu,Zhen Zhang,Guichao Li,Hongzhuan Chen,Zhen Xiang,Zhenggang Zhu,Hongyu Zhao,Yingyan Yu,Chaochun Wei
DOI: https://doi.org/10.1186/s13059-019-1751-y
IF: 17.906
2019-01-01
Genome Biology
Abstract:The human reference genome is still incomplete, especially for those population-specific or individual-specific regions, which may have important functions. Here, we developed a HUman Pan-genome ANalysis (HUPAN) system to build the human pan-genome. We applied it to 185 deep sequencing and 90 assembled Han Chinese genomes and detected 29.5 Mb novel genomic sequences and at least 188 novel protein-coding genes missing in the human reference genome (GRCh38). It can be an important resource for the human genome-related biomedical studies, such as cancer genome analysis. HUPAN is freely available at http://cgm.sjtu.edu.cn/hupan/ and https://github.com/SJTU-CGM/HUPAN .
What problem does this paper attempt to address?