Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • liaojinyue
    Junior Member
    • Oct 2012
    • 8

    #1

    How to estimate representation of RRBS library

    Dear all,

    We have recently generated RRBS libraries and sequenced them. Although the CpG site coverage is acceptable (~1M with coverage > 10), we are puzzled by the alignment result. Enclosed figure shows the alignment results from two of our samples covering one fragment with MspI sites at both ends. However, we only detect reads deriving from positive strand for the first sample (coverage = 186 reads) and reads from negative strand for the second sample (coverage = 58reads). My understanding is that the positive and negative strand should be more or less equally amplified and we should be able to detect reads from both strands in each sample. The second figure shows the methylation value density for the CpG sites. It looks like the methylation levels are either 0% or 100%. Our concern is that reads mapped to each MspI digested fragment might all come from one template. Does it mean that we do not have sufficient representation due to sample loss during library preparation? I will be grateful for your help.


    Best,

    Jason
    Attached Files
  • dpryan
    Devon Ryan
    • Jul 2011
    • 3478

    #2
    In my experience this sort of thing ends up being fairly common in RRBS. My guess is that PCR amplification leads to this sort of bias.

    Comment

    • nucacidhunter
      Jafar Jabbari
      • Jan 2013
      • 1250

      #3
      All current RRBS methods cover both strands unless you have used a novel method that targets only one strand. You have not mentioned if observed strand specificity in libraries is for all fragments or just for a subset of them. If your data analysis is correct and all fragments in a library come from a particular strand, I would suspect that library prep has not been optimal.

      If your adapters did not include UMIs it will be difficult to know if reads for a fragments are PCR duplicates or originate from different restriction fragments.

      Comment

      • ahlinyaobatherbal
        Junior Member
        • May 2016
        • 1

        #4
        Why my backlink can not be read ?
        cordyceps plus capsule

        Comment

        • Diagenode
          Junior Member
          • Apr 2016
          • 3

          #5
          I think we are dealing with two issues here.
          The first one is the question of extreme values (all 0 or 100% methylation), which I think is normal, or at least usual in RRBS-seq. Especially if you work with cell lines, the cells have no reason to have different methylation profiles, but even if you prepare biopsies from a fairly homogene tissue like liver, I would not expect a lot of cell-to-cell variation, which means that a certain CpG site is either methylated or unmethylated in all cells uniformly, so the sequencing results should reflect this. We have recently done RRBS-seq on a batch of samples, I quickly checked the methylation ratios of CpGs, and I also noticed that over 90% of all the covered CpGs are methylated either >90% or <10%.
          And remember that RRBS mostly capture the most prominent CpG islands, that have a crucial role in gene expression regulation. Other CpGs, like the ones in CpG shores and shelves and further away from the islands are less represented, especially if you have a low coverage, and I think it is these CpGs where we can expect a bigger variation, as they are likely to have a less essential role. There are also studies that prove that C methylation plays a role in other context than CpG, eg. CHG (see for example http://nar.oxfordjournals.org/content/42/5/3009.full) in these contexts probably there are more variations too. But CpG islands seem to be quite consistent (within a tissue or a cell line at least).

          The other issue (getting reads only from a single strand) is far from normal. It must be some library preparation issue; as nucacidhunter said, both strands should be covered more or less equally. I can only recommend this kit: Premium RRBS, which we use routinely for our RRBS-seq service, and all of our customers are very satisfied with it so far, we haven't observed the strand bias you described. It generates a pool of several samples for a cost-effective sequencing, and it's very user friendly, easy to use, difficult to make mistakes, not to mention the great CpG representation and coverage we get.

          Comment

          • Niels Wagemaker
            Junior Member
            • Mar 2016
            • 6

            #6
            We had some bias in the watson crick # reads but when we starter doing epigbs with 5mc insensitive restriction enzymen this was solved...

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
              by SEQadmin2



              CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

              Despite this, “CRISPR helped turn genome editing from a specialized technique into
              ...
              Today, 11:01 AM
            • SEQadmin2
              Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
              by SEQadmin2


              Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

              The systematic characterization of the human proteome has
              ...
              07-20-2026, 11:48 AM
            • SEQadmin2
              Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
              by SEQadmin2



              Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
              ...
              07-09-2026, 11:10 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, Today, 02:55 AM
            0 responses
            7 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-24-2026, 12:17 PM
            0 responses
            11 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-23-2026, 11:41 AM
            0 responses
            12 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-20-2026, 11:10 AM
            0 responses
            24 views
            0 reactions
            Last Post SEQadmin2  
            Working...