Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • rasel@barj
    Junior Member
    • Nov 2013
    • 4

    #1

    de novo assembly of PROTON transcriptome data

    Hi all,

    We performed a plant whole transcriptome sequencing in Ion PROTON. We have about 10 Gbp data of 74 million reads in a single run using v2 proton kits. Average read length is 129 bp. Which assembler will give us better output for de nove assembly. Could anyone suggest me?

    Thanks
    Rasel
  • GenoMax
    Senior Member
    • Feb 2008
    • 7142

    #2
    Have a look at this thread: http://seqanswers.com/forums/showthread.php?t=1610

    There are several other threads available (look at the top panel in "similar threads").

    Short answer may be that you may have to try a few assemblers out. If you have a similar species for which a transcriptome/genome is available then you can use that as a rough reference to do some comparisons to the assembly that you build.

    Comment

    • rasel@barj
      Junior Member
      • Nov 2013
      • 4

      #3
      Thanks GenoMax for your suggestion.

      We tried newbler but it was failed to assemble the data. Now MIRA is running for 13 days but still not completed. We don;t know why it is taking too much time although we only took 5 Gbp data for MIRA assembly running now. Last time in case of 2 Gbp PROTON transcriptome data MIRA took 12 days to complete and the metrics were not good.

      Actually I am worried about the data type of PROTON which may cause problem in assembler. Unfortunately there is no sufficient past example or I couldn't find specific post related to PROTON transcriptome (plant) data de novo assembly.


      Thanks again!
      Rasel

      Comment

      • colindaven
        Senior Member
        • Oct 2008
        • 417

        #4
        Soapdenovo or in your case the -trans version has been pretty capable or at least quick in my hands.
        CLC assembly cell (commercial) or CLC genomics workbench might also be worth a look.

        Velvet might be an option too.

        Comment

        • rasel@barj
          Junior Member
          • Nov 2013
          • 4

          #5
          Thanks colindaven for your suggestion

          Comment

          • jonathanjacobs
            Member
            • Apr 2011
            • 23

            #6
            CLC's assembler is quite good (and very fast) - so I would second that suggestion. Also, you may want to try IDBA's transcriptome assembler: http://i.cs.hku.hk/~alse/hkubrg/proj...ran/index.html

            That being said -- 129 bp reads are not going to give you great assemblies - no matter what assembler you choose. Proton RNA-seq data is going to be much better for read mapping and mutation/SNP mapping. For de novo assembly of transcriptomes, you would do much better using PE reads.
            @bioinformer
            http://www.linkedin.com/in/jonathanjacobs

            Comment

            • super0925
              Senior Member
              • Feb 2014
              • 206

              #7
              Originally posted by rasel@barj View Post
              Thanks colindaven for your suggestion
              Hi rasel@barj ,
              I have similar issue , average reads length is 114 bp at Proton , 28 M reads per run.
              Proton use MIRA as their GUI plugin, but it seems that it is very slow from your post.
              So which assembler did you use eventually ?!

              Comment

              • super0925
                Senior Member
                • Feb 2014
                • 206

                #8
                Originally posted by jonathanjacobs View Post
                CLC's assembler is quite good (and very fast) - so I would second that suggestion. Also, you may want to try IDBA's transcriptome assembler: http://i.cs.hku.hk/~alse/hkubrg/proj...ran/index.html

                That being said -- 129 bp reads are not going to give you great assemblies - no matter what assembler you choose. Proton RNA-seq data is going to be much better for read mapping and mutation/SNP mapping. For de novo assembly of transcriptomes, you would do much better using PE reads.
                Hi jonathanjacobs,
                (1)Why do you say that? (i.e. 129 bp reads are not going to give you great assemblies)
                (2)Why do you say you would prefer use PE reads ?
                (3) Is the IDBA's transcriptome assembler suitable for Proton data as well? how does it compare with Abyss, Velvet or Soapdenovo?
                Last edited by super0925; 07-22-2015, 05:55 AM.

                Comment

                • super0925
                  Senior Member
                  • Feb 2014
                  • 206

                  #9
                  How about Trinity?

                  Comment

                  • rasel@barj
                    Junior Member
                    • Nov 2013
                    • 4

                    #10
                    Hi super0925

                    We tried several assemblers but no one can build high quality assembly. Among all the assemblers Trinity gives much better metrics than others.

                    Comment

                    Latest Articles

                    Collapse

                    • SEQadmin2
                      Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
                      by SEQadmin2



                      CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

                      Despite this, “CRISPR helped turn genome editing from a specialized technique into
                      ...
                      07-31-2026, 11:01 AM
                    • SEQadmin2
                      Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
                      by SEQadmin2


                      Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

                      The systematic characterization of the human proteome has
                      ...
                      07-20-2026, 11:48 AM

                    ad_right_rmr

                    Collapse

                    News

                    Collapse

                    Topics Statistics Last Post
                    Started by SEQadmin2, 08-13-2026, 12:22 PM
                    0 responses
                    25 views
                    0 reactions
                    Last Post SEQadmin2  
                    Started by SEQadmin2, 08-11-2026, 10:35 AM
                    0 responses
                    20 views
                    0 reactions
                    Last Post SEQadmin2  
                    Started by SEQadmin2, 08-06-2026, 07:41 AM
                    0 responses
                    33 views
                    0 reactions
                    Last Post SEQadmin2  
                    Started by SEQadmin2, 08-03-2026, 10:13 AM
                    0 responses
                    51 views
                    0 reactions
                    Last Post SEQadmin2  
                    Working...