Prediction of the cell-type-specific transcription of non-coding RNAs from genome sequences via machine learning

Masaru Koido, Chung Chau Hon, Satoshi Koyama, Hideya Kawaji, Yasuhiro Murakawa, Kazuyoshi Ishigaki, Kaoru Ito, Jun Sese, Nicholas F. Parrish, Yoichiro Kamatani, Piero Carninci, Chikashi Terao

研究成果: Article査読

11 被引用数 (Scopus)

抄録

Gene transcription is regulated through complex mechanisms involving non-coding RNAs (ncRNAs). As the transcription of ncRNAs, especially of enhancer RNAs, is often low and cell type specific, how the levels of RNA transcription depend on genotype remains largely unexplored. Here we report the development and utility of a machine-learning model (MENTR) that reliably links genome sequence and ncRNA expression at the cell type level. Effects on ncRNA transcription predicted by the model were concordant with estimates from published studies in a cell-type-dependent manner, regardless of allele frequency and genetic linkage. Among 41,223 variants from genome-wide association studies, the model identified 7,775 enhancer RNAs and 3,548 long ncRNAs causally associated with complex traits across 348 major human primary cells and tissues, such as rare variants plausibly altering the transcription of enhancer RNAs to influence the risks of Crohn’s disease and asthma. The model may aid the discovery of causal variants and the generation of testable hypotheses for biological mechanisms driving complex traits.

本文言語English
ページ(範囲)830-844
ページ数15
ジャーナルNature Biomedical Engineering
7
6
DOI
出版ステータスPublished - 2023 6月
外部発表はい

ASJC Scopus subject areas

  • バイオテクノロジー
  • バイオエンジニアリング
  • 医学(その他)
  • 生体医工学
  • コンピュータ サイエンスの応用

フィンガープリント

「Prediction of the cell-type-specific transcription of non-coding RNAs from genome sequences via machine learning」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル