HUPAN: a pan-genome analysis pipeline for human genomes.
Zhongqu DuanYuyang QiaoJinyuan LuHuimin LuWenmin ZhangFazhe YanChen SunZhiqiang HuZhen ZhangGuichao LiHongzhuan ChenZhen XiangZhenggang ZhuHongyu ZhaoYingyan YuChaochun WeiPublished in: Genome biology (2019)
The human reference genome is still incomplete, especially for those population-specific or individual-specific regions, which may have important functions. Here, we developed a HUman Pan-genome ANalysis (HUPAN) system to build the human pan-genome. We applied it to 185 deep sequencing and 90 assembled Han Chinese genomes and detected 29.5 Mb novel genomic sequences and at least 188 novel protein-coding genes missing in the human reference genome (GRCh38). It can be an important resource for the human genome-related biomedical studies, such as cancer genome analysis. HUPAN is freely available at http://cgm.sjtu.edu.cn/hupan/ and https://github.com/SJTU-CGM/HUPAN .