首页    期刊浏览 2024年11月29日 星期五
登录注册

文章基本信息

  • 标题:Highly accurate long-read HiFi sequencing data for five complex genomes
  • 本地全文:下载
  • 作者:Ting Hon ; Kristin Mars ; Greg Young
  • 期刊名称:Scientific Data
  • 电子版ISSN:2052-4463
  • 出版年度:2020
  • 卷号:7
  • 期号:1
  • 页码:1-11
  • DOI:10.1038/s41597-020-00743-4
  • 语种:English
  • 出版社:Nature Publishing Group
  • 摘要:The PacBio庐 HiFi sequencing method yields highly accurate long-read sequencing datasets with read lengths averaging 10鈥?5鈥塳b and accuracies greater than 99.5%. These accurate long reads can be used to improve results for complex applications such as single nucleotide and structural variant detection, genome assembly, assembly of difficult polyploid or highly repetitive genomes, and assembly of metagenomes. Currently, there is a need for sample data sets to both evaluate the benefits of these long accurate reads as well as for development of bioinformatic tools including genome assemblers, variant callers, and haplotyping algorithms. We present deep coverage HiFi datasets for five complex samples including the two inbred model genomes Mus musculus and Zea mays, as well as two complex genomes, octoploid Fragaria鈥壝椻€?i>ananassa and the diploid anuran Rana muscosa. Additionally, we release sequence data from a mock metagenome community. The datasets reported here can be used without restriction to develop new algorithms and explore complex genome structure and evolution. Data were generated on the PacBio Sequel II System.
国家哲学社会科学文献中心版权所有