Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • NM_010117
    Member
    • Oct 2010
    • 14

    Bowtie and Tophat in disagreement with alignment

    Hi all,

    I had a question about running bowtie and tophat. If I run bowtie (flags: -p 4 -S --un -v 3; paired-end) independently of tophat, I get very different results when looking at what % of the reads were mapped to my reference genome (hg18). Bowtie by itself aligns 38,474,274 of 87,783,858 reads to hg18, whereas when tophat is run (flags: -p 4 -r 20; paired-end) it maps 69,190,587 reads. Does anyone have an idea why they would differ so greatly?

    The data is 76bp paired end from cancer patients done on the GAIIx.
  • ZhengXia
    Member
    • Feb 2009
    • 10

    #2
    Tophat will map more reads than Bowtie because Tophat detects the splicing junction first which will make the reads spanning junctions mappable.
    Originally posted by NM_010117 View Post
    Hi all,

    I had a question about running bowtie and tophat. If I run bowtie (flags: -p 4 -S --un -v 3; paired-end) independently of tophat, I get very different results when looking at what % of the reads were mapped to my reference genome (hg18). Bowtie by itself aligns 38,474,274 of 87,783,858 reads to hg18, whereas when tophat is run (flags: -p 4 -r 20; paired-end) it maps 69,190,587 reads. Does anyone have an idea why they would differ so greatly?

    The data is 76bp paired end from cancer patients done on the GAIIx.

    Comment

    • NM_010117
      Member
      • Oct 2010
      • 14

      #3
      That makes sense, but surely it shouldn't have such a large impact that it doubles the number of reads aligned to the genome?

      Comment

      • kmcarr
        Senior Member
        • May 2008
        • 1181

        #4
        Originally posted by NM_010117 View Post
        That makes sense, but surely it shouldn't have such a large impact that it doubles the number of reads aligned to the genome?
        You observations sound perfectly in agreement with what I would expect. The average size of a exon in the human genome is 200bp. Your reads are 76nt. If the leftmost mapping position of a read is < 76bp from the end of the exon it will fail to map using bowtie. This 76bp "dead window" if you will represents 38% of the average exon length. You observed a reduction of 44%; you're in the ballbark.

        Comment

        • NM_010117
          Member
          • Oct 2010
          • 14

          #5
          Originally posted by kmcarr View Post
          You observations sound perfectly in agreement with what I would expect. The average size of a exon in the human genome is 200bp. Your reads are 76nt. If the leftmost mapping position of a read is < 76bp from the end of the exon it will fail to map using bowtie. This 76bp "dead window" if you will represents 38% of the average exon length. You observed a reduction of 44%; you're in the ballbark.
          Ah ha! Ok, that makes perfect sense. So tophat and bowtie are more or less in agreement with eachother; I'm just interpreting the results incorrectly. Well, at least it's promising to know that the mapping was done correctly at least.

          Thanks for the help.

          Comment

          Latest Articles

          Collapse

          • SEQadmin2
            Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
            by SEQadmin2


            Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

            The systematic characterization of the human proteome has
            ...
            07-20-2026, 11:48 AM
          • SEQadmin2
            Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
            by SEQadmin2



            Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
            ...
            07-09-2026, 11:10 AM
          • SEQadmin2
            Cancer Drug Resistance: The Lingering Barrier to Rising Survival
            by SEQadmin2



            Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

            There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
            07-08-2026, 05:17 AM

          ad_right_rmr

          Collapse

          News

          Collapse

          Topics Statistics Last Post
          Started by SEQadmin2, 07-24-2026, 12:17 PM
          0 responses
          16 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 07-23-2026, 11:41 AM
          0 responses
          18 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 07-20-2026, 11:10 AM
          0 responses
          24 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 07-13-2026, 10:26 AM
          0 responses
          37 views
          0 reactions
          Last Post SEQadmin2  
          Working...