The long and the short of it: unlocking nanopore long-read RNA sequencing data with short-read differential expression analysis tools

Abstract Application of Oxford Nanopore Technologies’ long-read sequencing platform to transcriptomic analysis is increasing in popularity. However, such analysis can be challenging due to the high sequence error and small library sizes, which decreases quantification accuracy and reduces power for...

Full description

Saved in:

Bibliographic Details
Published in	NAR genomics and bioinformatics Vol. 3; no. 2; p. lqab028
Main Authors	Dong, Xueyi, Tian, Luyi, Gouil, Quentin, Kariyawasam, Hasaru, Su, Shian, De Paoli-Iseppi, Ricardo, Prawer, Yair David Joseph, Clark, Michael B, Breslin, Kelsey, Iminitoff, Megan, Blewitt, Marnie E, Law, Charity W, Ritchie, Matthew E
Format	Journal Article
Language	English
Published	England Oxford University Press 01.06.2021
Subjects	APP Notes
Online Access	Get full text

Cover

Loading…

More Information
Summary:	Abstract Application of Oxford Nanopore Technologies’ long-read sequencing platform to transcriptomic analysis is increasing in popularity. However, such analysis can be challenging due to the high sequence error and small library sizes, which decreases quantification accuracy and reduces power for statistical testing. Here, we report the analysis of two nanopore RNA-seq datasets with the goal of obtaining gene- and isoform-level differential expression information. A dataset of synthetic, spliced, spike-in RNAs (‘sequins’) as well as a mouse neural stem cell dataset from samples with a null mutation of the epigenetic regulator Smchd1 was analysed using a mix of long-read specific tools for preprocessing together with established short-read RNA-seq methods for downstream analysis. We used limma-voom to perform differential gene expression analysis, and the novel FLAMES pipeline to perform isoform identification and quantification, followed by DRIMSeq and limma-diffSplice (with stageR) to perform differential transcript usage analysis. We compared results from the sequins dataset to the ground truth, and results of the mouse dataset to a previous short-read study on equivalent samples. Overall, our work shows that transcriptomic analysis of long-read nanopore data using long-read specific preprocessing methods together with short-read differential expression methods and software that are already in wide use can yield meaningful results.
Bibliography:	ObjectType-Article-1 SourceType-Scholarly Journals-1 ObjectType-Feature-2 content type line 23
ISSN:	2631-9268 2631-9268
DOI:	10.1093/nargab/lqab028