Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • KBlackwell
    Junior Member
    • Jun 2016
    • 1

    #1

    Hypervariable region(s) of choice for Illumina MiSeq

    I am having trouble determining why certain hypervariable regions are targeted. The papers I have reviewed all seem to say it depends on your study, such as whether you're interested in the rare biosphere or specific types of bacteria or archaea.

    However, I am not necessarily sure what should be used for environmental samples from the deep ocean, which is what I am looking at in my project. I will also be using Illumina MiSeq. Some papers I have looked at use V1-V2 and others use V6 or V6-V8, however a reason isn't provided.

    Any advice would be greatly appreciated!
  • Brian Bushnell
    Super Moderator
    • Jan 2014
    • 2709

    #2
    I think the region targeting has more to do with available read lengths than anything else. Personally... I think it's best to use the region that has the most comprehensive databases, to enhance your ability to classify whatever you are sequencing while also maximizing your contributions. Another factor to consider is the conservation of primer sites; the more highly-conserved, the more likely your data will have some kind of relationship with the actual abundance. If you are interested in archaea, be sure to design your primers appropriately...

    Comment

    • thermophile
      Senior Member
      • Apr 2015
      • 243

      #3
      Agreed, length is the biggest consideration. This is why v6 was first chosen, it's the shortest. Then JGI chose v6-9 for their standard for 454 (though they sequenced from 3' so rarely were you able to keep v6). I liked v1-3 for 454 soil communities (maximized diversity recovered) but that's too long for current Illumina, so now I use v4 like everyone else. Unless you are targeting a specific group of organisms, I'd suggest sticking with v4.
      Microbial ecologist, running a sequencing core. I have lots of strong opinions on how to survey communities, pretty sure some are even correct.

      Comment

      • nucacidhunter
        Jafar Jabbari
        • Jan 2013
        • 1250

        #4
        I agree with above comments. Following would affect identified taxa or species:
        1- sampling, storage and handling and DNA extraction
        2- targeted hyper variable region
        3- library prep method
        4- database used for identification

        It is important to note that in each case one gets only one view of the community composition and other views are possible as well.

        I have not seen any publication trying all variable regions and then choosing one for the bulk of their material.

        Comment

        • thermophile
          Senior Member
          • Apr 2015
          • 243

          #5
          Originally posted by nucacidhunter View Post
          I have not seen any publication trying all variable regions and then choosing one for the bulk of their material.
          I saw posters of JGI's efforts on that front back in the day-probably 2009ish. Not sure if they ever published it.
          Microbial ecologist, running a sequencing core. I have lots of strong opinions on how to survey communities, pretty sure some are even correct.

          Comment

          • strawbaubz
            Junior Member
            • Apr 2015
            • 5

            #6
            Hi guys sorry for the random poat but not sure where I could ask this question, whihc seems a little stupid but here goes.

            I sequenced from the V1V3 and V3V4 regions. I looked at the quality profile of my sequences and found that my V1V3 reads had way poorer per base quality that V3V4.

            I used the Illumina Miseq 2X300bp to sequence my reads. I am trying to come up with a reason for this but can't seem to grasp the concept.

            Is possible that the sequences within the V3V4 can trigger more errors and thus poorer quality profiles? HELP??

            I apologize if this is a stupid question.

            Comment

            • thermophile
              Senior Member
              • Apr 2015
              • 243

              #7
              those products are way too long, you should be aiming for near complete overlap. Additionally the error profiles for the v3 2x300 kits are worse than the v2 2x250.
              Microbial ecologist, running a sequencing core. I have lots of strong opinions on how to survey communities, pretty sure some are even correct.

              Comment

              • strawbaubz
                Junior Member
                • Apr 2015
                • 5

                #8
                Thanks for reply.

                I meant to say my V1V3 is worst than V3V4. You mention that the error profiles are worst for V3 kits than V2. I see the opposite?

                I read a paper that said some sequence patterns can trigger more errors than others. Can this be happening to my data? I am writing up results and not sure if I have to try explain why it is that my V1V3 results did worst?

                I know that the overlap was small but I am talking about the reads not the merged sequences?

                Comment

                • Brian Bushnell
                  Super Moderator
                  • Jan 2014
                  • 2709

                  #9
                  If you have very high homogeneity for a while at the beginning of the read, it will reduce the quality there. The effects can carry over to some extent for the remainder of the read, even when you get into variable regions. This effect can be reduced with staggered adapters. It's also possible that this is sample-specific, if you are running different organisms, or amplification-specific, if different organisms are amplifying well.

                  Comment

                  • nucacidhunter
                    Jafar Jabbari
                    • Jan 2013
                    • 1250

                    #10
                    Originally posted by strawbaubz View Post
                    Hi guys sorry for the random poat but not sure where I could ask this question, whihc seems a little stupid but here goes.

                    I sequenced from the V1V3 and V3V4 regions. I looked at the quality profile of my sequences and found that my V1V3 reads had way poorer per base quality that V3V4.

                    I used the Illumina Miseq 2X300bp to sequence my reads. I am trying to come up with a reason for this but can't seem to grasp the concept.

                    Is possible that the sequences within the V3V4 can trigger more errors and thus poorer quality profiles? HELP??

                    I apologize if this is a stupid question.
                    Possible reasons:

                    1- batch to batch variation of sequencing kits
                    2- higher cluster density of V1-V3 run
                    3- differences in PhiX spike in
                    4- differences in sequence diversity of regions in your samples, lower sequence variation along reads will result in lower quality

                    I do not think sequence error will reduce Q score. A base with high Q score can be erroneous.
                    Last edited by nucacidhunter; 12-05-2016, 12:59 PM.

                    Comment

                    • thermophile
                      Senior Member
                      • Apr 2015
                      • 243

                      #11
                      When you look at the individual reads the qscores are worse v1v3 vs v3v4 or merged reads are worse. v1v3 is longer so you have less overlap therefor I'd expect poorer quality in the merged reads.

                      Base homogeneity across bacteria is greater in the beginning of v1 than v3 so i don't think that's your answer.
                      Last edited by thermophile; 12-06-2016, 07:09 AM. Reason: clarifying
                      Microbial ecologist, running a sequencing core. I have lots of strong opinions on how to survey communities, pretty sure some are even correct.

                      Comment

                      • nucacidhunter
                        Jafar Jabbari
                        • Jan 2013
                        • 1250

                        #12
                        Originally posted by thermophile View Post

                        Base homogeneity across bacteria is greater in the beginning of v1 than v3 so i don't think that's your answer.
                        Low sequence diversity at the start of read will affect number of reads passing filter with no or little impact on the quality of bases past the low diversity region. However, a low diversity region, for instance, in the middle of the read will reduce the Q scores for those bases.

                        Sequence heterogeneity of the target region for some samples can be very low if they are composed of bacteria mainly from one particular phyla even though that region might be very heterogeneous if considered globally among all bacteria.

                        Comment

                        Latest Articles

                        Collapse

                        • SEQadmin2
                          Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
                          by SEQadmin2



                          CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

                          Despite this, “CRISPR helped turn genome editing from a specialized technique into
                          ...
                          07-31-2026, 11:01 AM
                        • SEQadmin2
                          Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
                          by SEQadmin2


                          Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

                          The systematic characterization of the human proteome has
                          ...
                          07-20-2026, 11:48 AM
                        • SEQadmin2
                          Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
                          by SEQadmin2



                          Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
                          ...
                          07-09-2026, 11:10 AM

                        ad_right_rmr

                        Collapse

                        News

                        Collapse

                        Topics Statistics Last Post
                        Started by SEQadmin2, 07-31-2026, 02:55 AM
                        0 responses
                        17 views
                        0 reactions
                        Last Post SEQadmin2  
                        Started by SEQadmin2, 07-24-2026, 12:17 PM
                        0 responses
                        15 views
                        0 reactions
                        Last Post SEQadmin2  
                        Started by SEQadmin2, 07-23-2026, 11:41 AM
                        0 responses
                        13 views
                        0 reactions
                        Last Post SEQadmin2  
                        Started by SEQadmin2, 07-20-2026, 11:10 AM
                        0 responses
                        25 views
                        0 reactions
                        Last Post SEQadmin2  
                        Working...