Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • Cuffset number of genes too high

    Hi,

    So I'm implementing a pretty standard tuxedo pipeline on paired-end mouse data. I went along as follows.

    Using the standard mouse build (37) from Ensemble and the associated annotation data I did the following.

    cutadapt -> tophat2 -> cufflinks -> cuffmerge -> cuffquant -> cuffdiff -> cummerbund.

    Tophat2 is giving me roughly 85% overall mapping, 5-10% multimap. Cufflinks was run on around 100 files and those were merged with cuffmerge. All of the bam files were then run with cuffquant using the gtf from cuffmerge.

    To test out an initial dataset I just used two sample conditions (2 replicates each) to do cuffdiff. 8 files. (4 control, 4 corresponding experimental case).


    Now, I load this up in cummeRbund and see the following:
    CuffSet instance with:
    2 samples
    52615 genes
    220501 isoforms
    106928 TSS
    49476 CDS
    52615 promoters
    106928 splicing
    21223 relCDS
    So this is my first time working with mouse or cufflinks pipeline, but somehow these numbers don't feel right. So I checked it out and at least I found that mouse has only around 23,000 genes, so thats wrong for sure.

    Could someone explain to me what sort of numbers I should be seeing here and why at least cuffdiff/cummRbund is showing approximately double the number of genes that should exist in the genome I'm looking at?

    I appreciate any feedback anyone can provide.

Latest Articles

Collapse

  • seqadmin
    Techniques and Challenges in Conservation Genomics
    by seqadmin



    The field of conservation genomics centers on applying genomics technologies in support of conservation efforts and the preservation of biodiversity. This article features interviews with two researchers who showcase their innovative work and highlight the current state and future of conservation genomics.

    Avian Conservation
    Matthew DeSaix, a recent doctoral graduate from Kristen Ruegg’s lab at The University of Colorado, shared that most of his research...
    03-08-2024, 10:41 AM
  • seqadmin
    The Impact of AI in Genomic Medicine
    by seqadmin



    Artificial intelligence (AI) has evolved from a futuristic vision to a mainstream technology, highlighted by the introduction of tools like OpenAI's ChatGPT and Google's Gemini. In recent years, AI has become increasingly integrated into the field of genomics. This integration has enabled new scientific discoveries while simultaneously raising important ethical questions1. Interviews with two researchers at the center of this intersection provide insightful perspectives into...
    02-26-2024, 02:07 PM

ad_right_rmr

Collapse

News

Collapse

Topics Statistics Last Post
Started by seqadmin, 03-14-2024, 06:13 AM
0 responses
32 views
0 likes
Last Post seqadmin  
Started by seqadmin, 03-08-2024, 08:03 AM
0 responses
71 views
0 likes
Last Post seqadmin  
Started by seqadmin, 03-07-2024, 08:13 AM
0 responses
80 views
0 likes
Last Post seqadmin  
Started by seqadmin, 03-06-2024, 09:51 AM
0 responses
68 views
0 likes
Last Post seqadmin  
Working...
X