A systematic analyses of different bioinformatics pipelines for genomic data and its impact on deep learning models for chromatin loop prediction
Anup Kumar Halder, Abhishek Agarwal, Karolina Jodkowska, Dariusz Plewczyński
Warsaw University of Technology Institut thématique Génétique, génomique et bioinformatique University of Warsaw
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Genomic data analysis has witnessed a surge in complexity and volume, primarily driven by the advent of high-throughput technologies. In particular, studying chromatin loops and structures has become pivotal in understanding gene regulation and genome organization. This systematic investigation explores the realm of specialized bioinformatics pipelines designed specifically for the analysis of chromatin loops and structures. Our investigation incorporates two protein (CTCF and Cohesin) factor-specific loop interaction datasets from six distinct pipelines, amassing a comprehensive collection of 36 diverse datasets. Through a meticulous review of existing literature, we offer a holistic perspective on the methodologies, tools and algorithms underpinning the analysis of this multifaceted genomic feature. We illuminate the vast array of approaches deployed, encompassing pivotal aspects such as data preparation pipeline, preprocessing, statistical features and modelling techniques. Beyond this, we rigorously assess the strengths and limitations inherent in these bioinformatics pipelines, shedding light on the interplay between data quality and the performance of deep learning models, ultimately advancing our comprehension of genomic intricacies.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
生物医学Genomics and Chromatin Dynamics
Genetic Mapping and Diversity in Plants and Animals · Chromosomal and Genetic Variations
参考文献 91
此处列出前 3 条
引用本文 3
按被引量排序,此处列出前 3 条