Unconfigured Ad

**chadn737** · 01-13-2013, 10:30 PM

1) Do not convert the reads to fasta. That is unnecessary, Tophat takes fastq files....your references will always be in fasta, this is not an issue.

2) It sounds like you did have a lot of unmapped reads, but its impossible to diagnose the issue from what you have described. Try providing some information like, the exact commands you ran Tophat with or some examples of unmapped reads.

**amarth** · 01-13-2013, 11:19 PM

1) Thanks for the advice,

2) Could it be the seq Quality?

The Tophat2 command i ran was:

tophat2 --sequence-length 100 --max-insertion-length 3 --max-deletion-length 3 reference index3.fasta

so much thanks

**EGrassi** · 01-14-2013, 01:00 AM

What does samtools flagstat on the accepted_hits.bam and rejected looks like?

Just to mention the fact: I had some paired data with fastqc results similar to yours and quality trimming lead me to nothing, the problem was about the -r/--mate-std-dev parameters (I had paired data): it seems that in some cases tophat really needs them (-r 300 --mate-std-dev 50 gave me a 60% percentage of properly paired reads against the 5% without them...I will never trust the FAQ/manual again

).

**kmcarr** · 01-14-2013, 10:01 AM

Originally posted by amarth View Post

2) Could it be the seq Quality?

The quality of your reads looks fine.

The Tophat2 command i ran was:

Code:

tophat2 --sequence-length 100 --max-insertion-length 3 --max-deletion-length 3 reference index3.fasta

so much thanks

There is no '--sequence-length' option for tophat. Did you mean '--segment-length'? If so 100 (presumably the full length of your read) is not an appropriate setting. The default value for --segment-length (25) is appropriate for most cases.

To diagnose the problem start using only default options and then work out from there.

Topics	Statistics	Last Post
New AI Model Captures Long-Range Genomic Signals to Improve RNA Splice Site Prediction by SEQadmin2 Started by SEQadmin2, 06-30-2026, 05:37 AM	0 responses 11 views 0 reactions	Last Post by SEQadmin2 06-30-2026, 05:37 AM
Large-Scale Protein Screen Uncovers Hidden Regulators of Alternative Polyadenylation by SEQadmin2 Started by SEQadmin2, 06-26-2026, 11:10 AM	0 responses 18 views 0 reactions	Last Post by SEQadmin2 06-26-2026, 11:10 AM
Whole-Genome Sequencing Traces Faroe Islands Ancestry to a North Atlantic Founder Population by SEQadmin2 Started by SEQadmin2, 06-17-2026, 06:09 AM	0 responses 52 views 0 reactions	Last Post by SEQadmin2 06-17-2026, 06:09 AM
Sequencing the Two-Toed Sloth Genome Reveals Jumping Genes Tied to Its Extreme Metabolism by SEQadmin2 Started by SEQadmin2, 06-09-2026, 11:58 AM	0 responses 111 views 0 reactions	Last Post by SEQadmin2 06-09-2026, 11:58 AM

Unconfigured Ad

Getting just a few alignments with Tophat2

Comment

Comment

Comment

Comment

Latest Articles

ad_right_rmr

News