Modeling Circrna Expression Pattern With Integrated Sequence And Epigenetic Features Demonstrates The Potential Involvement Of H3k79me2 In Circrna Expression

BIOINFORMATICS(2020)

引用 5|浏览76
暂无评分
摘要
Motivation: CircRNAs are an abundant class of non-coding RNAs with widespread, cell-/tissue-specific patterns. Previous work suggested that epigenetic features might be related to circRNA expression. However, the contribution of epigenetic changes to circRNA expression has not been investigated systematically. Here, we built a machine learning framework named CIRCScan, to predict circRNA expression in various cell lines based on the sequence and epigenetic features.Results: The predicted accuracy of the expression status models was high with area under the curve of receiver operating characteristic (ROC) values of 0.89-0.92 and the false-positive rates of 0.17-0.25. Predicted expressed circRNAs were further validated by RNA-seq data. The performance of expression-level prediction models was also good with normalized root-mean-square errors of 0.28-0.30 and Pearson's correlation coefficient r over 0.4 in all cell lines, along with Spearman's correlation coefficient rho of 0.33-0.46. Noteworthy, H3K79me2 was highly ranked in modeling both circRNA expression status and levels across different cells. Further analysis in additional nine cell lines demonstrated a significant enrichment of H3K79me2 in circRNA flanking intron regions, supporting the potential involvement of H3K79me2 in circRNA expression regulation.
更多
查看译文
关键词
circrna expression,circrna expression pattern,epigenetic features,h3k79me2
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要