bioRxiv · 10.1101/683425
Evaluation of deep-learning-based lncRNA identification tools
Abstract
Long non-coding RNAs (lncRNAs, length above 200 nt) exert crucial biological roles and have been implicated in cancers1,2. To characterize newly discovered transcripts, one major issue is to distinguish lncRNAs from mRNAs. Since experimental methods are time-consuming and costly, computational methods are preferred for large-scale lncRNA identification. In a recent study, Amin et al.3 evaluated three deep-learning-based lncRNA identification tools (i.e., lncRNAnet4, LncADeep5, and lncFinder6) and concluded \"The LncADeep PR (precision recall) curve is just above the no-skill model and LncADeep showed poor overall performance\". This surprising conclusion is based on the authors use of a non-default setting of LncADeep. Actually, LncADeep has two models, one for full-length transcripts, and the other for transcripts including partial-length. Being aware of the difficulty of assembling full-length transcripts from RNA-seq dataset, LncADeeps default model is for transcripts including partial-length. However, according to the results posted on Amin et al.s website, the authors used LncADeep with full-length model, while they claimed to use the default setting of LncADeep, to identify lncRNAs from GENCODE dataset, which is composed of full- and partial-length transcripts. Thus, in their evaluation, the performance of LncADeep was underestimated. In this correspondence, we have tested LncADeeps default setting (i.e., model for transcripts including partial-length) on the datasets used in Amin et al.3, and LncADeep achieved overall the best performance compared with the other tools results reported by Amin et al.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yang, C., Zhou, M., Xie, H., Zhu, H.. 2019-06-28. Evaluation of deep-learning-based lncRNA identification tools. https://doi.org/10.1101/683425
Cite the original work for its findings. Save a collection to share your selection of sources.