Rich Chromatin Structure Prediction from Hi-C Data

被引:13
|
作者
Malik, Laraib [1 ]
Patro, Rob [1 ]
机构
[1] SUNY Stony Brook, Dept Comp Sci, Stony Brook, NY 11794 USA
基金
美国国家科学基金会;
关键词
Prediction algorithms; Biological cells; Frequency-domain analysis; Indexes; Tools; Genomics; Bioinformatics; Hierarchy; chromatin conformation capture; Hi-C; topologically associating domains; FUNCTIONAL-ORGANIZATION; TOPOLOGICAL DOMAINS; GENOME; PRINCIPLES; TRANSCRIPTION; ARCHITECTURE; MODEL; MAP;
D O I
10.1109/TCBB.2018.2851200
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Recent studies involving the 3-dimensional conformation of chromatin have revealed the important role it has to play in different processes within the cell. These studies have also led to the discovery of densely interacting segments of the chromosome, called topologically associating domains. The accurate identification of these domains from Hi-C interaction data is an interesting and important computational problem for which numerous methods have been proposed. Unfortunately, most existing algorithms designed to identify these domains assume that they are non-overlapping whereas there is substantial evidence to believe a nested structure exists. We present a methodology to predict hierarchical chromatin domains using chromatin conformation capture data. Our method predicts domains at different resolutions, calculated using intrinsic properties of the chromatin data, and effectively clusters these to construct the hierarchy. At each individual level, the domains are non-overlapping in such a way that the intra-domain interaction frequencies are maximized. We show that our predicted structure is highly enriched for actively transcribing housekeeping genes and various chromatin markers, including CTCF, around the domain boundaries. We also show that large-scale domains, at multiple resolutions within our hierarchy, are conserved across cell types and species. We also provide comparisons against existing tools for extracting hierarchical domains. Our software, Matryoshka, is written in C++11 and licensed under GPL v3; it is available at https://github.com/COMBINE-lab/matryoshka.
引用
收藏
页码:1448 / 1458
页数:11
相关论文
共 50 条
  • [1] Rich Chromatin Structure Prediction from Hi-C Data
    Malik, Laraib
    Patro, Rob
    ACM-BCB' 2017: PROCEEDINGS OF THE 8TH ACM INTERNATIONAL CONFERENCE ON BIOINFORMATICS, COMPUTATIONAL BIOLOGY,AND HEALTH INFORMATICS, 2017, : 184 - 193
  • [2] Bayesian Reconstruction of Chromatin Conformation from FISH and Hi-C Data
    Pan, Keyao
    Bathe, Mark
    BIOPHYSICAL JOURNAL, 2013, 104 (02) : 581A - 581A
  • [3] CHROMSTRUCT 4: A Python']Python Code to Estimate the Chromatin Structure from Hi-C Data
    Caudai, Claudia
    Salerno, Emanuele
    Zoppe, Monica
    Merelli, Ivan
    Tonazzini, Anna
    IEEE-ACM TRANSACTIONS ON COMPUTATIONAL BIOLOGY AND BIOINFORMATICS, 2019, 16 (06) : 1867 - 1878
  • [4] Efficient Hi-C inversion facilitates chromatin folding mechanism discovery and structure prediction
    Schuette, Greg
    Ding, Xinqiang
    Zhang, Bin
    BIOPHYSICAL JOURNAL, 2023, 122 (17) : 3425 - 3438
  • [5] Extracting multi-way chromatin contacts from Hi-C data
    Liu, Lei
    Zhang, Bokai
    Hyeon, Changbong
    PLOS COMPUTATIONAL BIOLOGY, 2021, 17 (12)
  • [6] Fine mapping chromatin contacts in capture Hi-C data
    Christiaan Q Eijsbouts
    Oliver S Burren
    Paul J Newcombe
    Chris Wallace
    BMC Genomics, 20
  • [7] Fine mapping chromatin contacts in capture Hi-C data
    Eijsbouts, Christiaan Q.
    Burren, Oliver S.
    Newcombe, Paul J.
    Wallace, Chris
    BMC GENOMICS, 2019, 20 (1)
  • [8] The shape of chromatin: insights from computational recognition of geometric patterns in Hi-C data
    Raffo, Andrea
    Paulsen, Jonas
    BRIEFINGS IN BIOINFORMATICS, 2023, 24 (05)
  • [9] Reconstruction of the chromatin 3D conformation from single cell Hi-C data
    Kos, Pavel I.
    Galitsyna, Aleksandra A.
    Ulianov, Sergey V.
    Gelfand, Mikhail S.
    Razin, Sergey V.
    Chertovich, Alexander V.
    PROCEEDINGS 2018 IEEE INTERNATIONAL CONFERENCE ON BIOINFORMATICS AND BIOMEDICINE (BIBM), 2018, : 2476 - 2476
  • [10] Identifying statistically significant chromatin contacts from Hi-C data with FitHiC2
    Kaul, Arya
    Bhattacharyya, Sourya
    Ay, Ferhat
    NATURE PROTOCOLS, 2020, 15 (03) : 991 - 1012