生物
纳米孔测序
顺序装配
仆从
康蒂格
基因组
遗传学
全基因组测序
酿酒酵母
参考基因组
计算生物学
同源(生物学)
基因座(遗传学)
基因
基因表达
转录组
作者
Alex Salazar,Arthur R. Gorter de Vries,Marcel van den Broek,Melanie Wijsman,Pilar de la Torre Cortés,Anja Brickwedde,Nick Brouwers,Jean‐Marc Daran,Thomas Abeel
标识
DOI:10.1093/femsyr/fox074
摘要
The haploid Saccharomyces cerevisiae strain CEN.PK113-7D is a popular model system for metabolic engineering and systems biology research. Current genome assemblies are based on short-read sequencing data scaffolded based on homology to strain S288C. However, these assemblies contain large sequence gaps, particularly in subtelomeric regions, and the assumption of perfect homology to S288C for scaffolding introduces bias. In this study, we obtained a near-complete genome assembly of CEN.PK113-7D using only Oxford Nanopore Technology's MinION sequencing platform. Fifteen of the 16 chromosomes, the mitochondrial genome and the 2-μm plasmid are assembled in single contigs and all but one chromosome starts or ends in a telomere repeat. This improved genome assembly contains 770 Kbp of added sequence containing 248 gene annotations in comparison to the previous assembly of CEN.PK113-7D. Many of these genes encode functions determining fitness in specific growth conditions and are therefore highly relevant for various industrial applications. Furthermore, we discovered a translocation between chromosomes III and VIII that caused misidentification of a MAL locus in the previous CEN.PK113-7D assembly. This study demonstrates the power of long-read sequencing by providing a high-quality reference assembly and annotation of CEN.PK113-7D and places a caveat on assumed genome stability of microorganisms.
科研通智能强力驱动
Strongly Powered by AbleSci AI