Bioinformatics for the 3D Genome: An Introduction to Analyzing and Interpreting Hi-C Data

Bioinformatics for the 3D Genome: An Introduction to Analyzing and Interpreting Hi-C Data

🎙 Sophia Nomuk 👥 1K 📅 September 28, 2023 ⏱ 59 min 👁 21K 📄 tutorial 🧭 2026-08-18
Available in: English (current) Français

Keywords

Hi-C3D genomebioinformaticschromatinJuicer

Summary

This webinar, presented by Sophia Nomuk from Arima Genomics, provides an introduction to analyzing and interpreting Hi-C data. It begins by explaining the importance of 3D genomics in understanding genome structure and gene regulation, highlighting applications such as identifying disease risk variants and structural variants. The talk then outlines the bioinformatics pipeline for Hi-C data, covering steps like aligning and pairing reads, filtering low-quality and duplicate reads, binning the genome into windows, and normalizing the contact matrix. The presenter emphasizes the use of the Juicer pipeline and its associated tools, such as Juicebox for visualization. Key metrics for quality control are discussed, including the percentage of chimeric reads, unmapped reads, and long-range interactions. The webinar also covers how to choose bin sizes for different analyses (e.g., 100 kb for compartments, 50 kb for TADs, 5 kb for loops) and provides sequencing depth recommendations. Finally, it touches on downstream analyses like loop calling and differential loop analysis, using a study on T-cell acute lymphoblastic leukemia as an example of how Hi-C can reveal enhancer-promoter interactions driving oncogene expression.

177 words

Critical Evaluation

Value of the Information & Strength of the Argument

The webinar provides a clear and structured introduction to Hi-C data analysis, valuable for researchers new to the field. The argumentation is solid, based on established methods and a specific example from the literature. The presenter explains concepts logically, building from the basics of 3D genomics to the specifics of the bioinformatics pipeline. The use of a real-world example (MYC oncogene in leukemia) effectively illustrates the biological relevance of Hi-C. The recommendations for sequencing depth and quality metrics are practical and based on Arima’s experience, though they are presented as guidelines rather than universally validated standards.

Scientific Rigor, Source Quality, Title Accuracy

The scientific rigor is high: the presenter is a computational biologist with a PhD, and the content aligns with standard practices in the field. The webinar references the Juicer pipeline, which is publicly available and widely used, and mentions a specific study on T-ALL, though the exact citation is not provided. The title accurately reflects the content, which is an introduction to Hi-C bioinformatics. The presentation is well-structured and avoids overclaiming, acknowledging the complexity of the field and the need for context-dependent choices. No comments were provided for analysis.

200 words

Title / Content Match

The title accurately reflects the content, which is an introduction to analyzing and interpreting Hi-C data.

Quality & Reliability

8/10

The webinar is presented by a computational biologist with a PhD, providing a structured overview of Hi-C data analysis. It covers key steps and metrics, referencing a publicly available pipeline (Juicer) and a specific study. The information is accurate and well-organized, though it is introductory and does not delve into code or advanced details.

Key Moments

Cited Sources

  • Juicer — Mentioned as the recommended pipeline for Hi-C data analysis
  • Juicebox — Mentioned as the visualization tool associated with Juicer

Concurring Sources

  • Juicer — The pipeline described in the webinar is publicly available and widely used.
  • Hi-C (genomics) — General information on Hi-C technique.

Contribution & Novelties

The webinar provides a clear, high-level introduction to Hi-C data analysis, demystifying the bioinformatics steps and offering practical recommendations for quality control and resolution. It is particularly useful for researchers new to 3D genomics, as it bridges the gap between experimental design and computational analysis. The example of MYC regulation in leukemia illustrates the power of Hi-C in understanding disease mechanisms.

Pour aller plus loin :

89 words

Radar Profile

The radar profile shows high scores in quantity and quality of information, with a moderate technical level, indicating a well-balanced introductory tutorial. The reliability is high, reflecting the expertise of the presenter and the use of established tools.

Reliability 8/10