Background <p>The zebrafish has significantly advanced our understanding of human disease and development, with nearly 70% of single-copy protein-coding genes conserved between the species. However, research on zebrafish is limited by gaps in existing genome annotations, which are primarily based on computational predictions and short-read sequencing data.</p> Results <p>To address this issue, we employed the PacBio Sequel II platform to generate a time-series full-length transcriptome landscape of zebrafish embryogenesis, covering 21 time points from embryo to six days post-fertilization. Our analysis uncovered 2113 previously unannotated genes and 33,018 novel isoforms of previously annotated genes, substantially expanding the current zebrafish gene annotations. We verified these findings using various methods, including domain prediction, homology analysis, conservation analysis, transcript quantification with short-read RNA-seq, and promoter position information with H3K4me3 and CAGE-seq. Furthermore, we analyzed the dynamic expression of transcripts across the 21 developmental stages using next-generation sequencing data, identifying variable splicing events throughout these stages.</p> Conclusions <p>Collectively, our study provides a high-resolution and significantly improved transcriptome annotation during zebrafish embryogenesis, offering a valuable resource for the zebrafish research community.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

High resolution of full-length RNA sequencing deciphers massive transcriptome complexity during zebrafish embryogenesis

  • Jing Bo,
  • Wenyu Fang,
  • Jing Wang,
  • Shunping He,
  • Liandong Yang

摘要

Background

The zebrafish has significantly advanced our understanding of human disease and development, with nearly 70% of single-copy protein-coding genes conserved between the species. However, research on zebrafish is limited by gaps in existing genome annotations, which are primarily based on computational predictions and short-read sequencing data.

Results

To address this issue, we employed the PacBio Sequel II platform to generate a time-series full-length transcriptome landscape of zebrafish embryogenesis, covering 21 time points from embryo to six days post-fertilization. Our analysis uncovered 2113 previously unannotated genes and 33,018 novel isoforms of previously annotated genes, substantially expanding the current zebrafish gene annotations. We verified these findings using various methods, including domain prediction, homology analysis, conservation analysis, transcript quantification with short-read RNA-seq, and promoter position information with H3K4me3 and CAGE-seq. Furthermore, we analyzed the dynamic expression of transcripts across the 21 developmental stages using next-generation sequencing data, identifying variable splicing events throughout these stages.

Conclusions

Collectively, our study provides a high-resolution and significantly improved transcriptome annotation during zebrafish embryogenesis, offering a valuable resource for the zebrafish research community.