Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • griffon42
    Member
    • Jan 2009
    • 23

    #1

    Interpretation of columns in MAQ SNP output?

    Hi all-

    Apologies if this information has been posted elsewhere - I wasn't able to find what I was looking for searching the forums.

    I'm using MAQ cns2SNP (and SNPfilter) to call SNPs on short read alignments, mouse genome. Though things have been going well, I'm having a hard time interpreting some of the output.

    The MAQ manual states, for cns2SNP output:
    "Each line consists of chromosome, position, reference base, consensus base, Phred-like consensus quality, read depth, the average number of hits of reads covering this position, the highest mapping quality of the reads covering the position, the minimum consensus quality in the 3bp flanking regions at each side of the site (6bp in total), the second best call, log likelihood ratio of the second best and the third best call, and the third best call. "

    I'm having trouble interpreting "the average number of hits of reads covering this position." Can anyone translate this into something more well-defined?

    Also, for "Phred-like consensus quality" - though I realize this is not a true Phred score, what is the best way to think about this? Overall consensus base quality at a given position? "Confidence score" for calling a SNP?
    In general, using SNPfilter and some arbitrary filters (discussed on the forums), I've been ignoring called SNPs with a "Phred-like" score of <40. I'd like to know what this actually means and whether it makes good sense.

    Thanks for you help!
  • der_eiskern
    Member
    • Jul 2009
    • 46

    #2
    re: Griffon interpreting SNPfilter outputs

    Originally posted by griffon42 View Post
    Hi all-

    Apologies if this information has been posted elsewhere - I wasn't able to find what I was looking for searching the forums.

    I'm using MAQ cns2SNP (and SNPfilter) to call SNPs on short read alignments, mouse genome. Though things have been going well, I'm having a hard time interpreting some of the output.

    The MAQ manual states, for cns2SNP output:
    "Each line consists of chromosome, position, reference base, consensus base, Phred-like consensus quality, read depth, the average number of hits of reads covering this position, the highest mapping quality of the reads covering the position, the minimum consensus quality in the 3bp flanking regions at each side of the site (6bp in total), the second best call, log likelihood ratio of the second best and the third best call, and the third best call. "

    I'm having trouble interpreting "the average number of hits of reads covering this position." Can anyone translate this into something more well-defined?
    ...
    In general, using SNPfilter and some arbitrary filters (discussed on the forums), I've been ignoring called SNPs with a "Phred-like" score of <40. I'd like to know what this actually means and whether it makes good sense.
    What I can gather is that the "average number of hits of reads covering this position" refers to the repetitiveness of the reads mapped to that locus. My impression was that many reads will map to more than 1 genomic position. Anything greater than ~1.1 I consider repetitive and ignore (if i can). This was suggested by Shen et al. According to them 1.1 "represents a conservative cutoff to avoid repeats and alleviate the mapping issues with shorter...reads."

    as for phred-like quality scores, i believe heng li talks about what that means in section 4.5 of the first MAQ paper: http://genome.cshlp.org/content/18/11/1851.full.
    quickly, (mathematics omitted)
    Before consensus calling, MAQ first combines mapping quality and base quality. If a read is incorrectly mapped, any sequence differences inferred from the read cannot be reliable. Therefore, the base quality used in SNP calling cannot exceed the mapping quality of the read. MAQ reassigns the quality of each base as the smaller value between the read mapping quality and the raw sequencing base quality.
    hope this helps.

    Comment

    • griffon42
      Member
      • Jan 2009
      • 23

      #3
      Thanks. That makes things considerably more clear.

      Comment

      • mixter
        Member
        • May 2010
        • 22

        #4
        Allele counts from MAQ snp output

        Hello,

        I have an additional question, after having read this thread and the documentation.

        What I need would be the direct counts for the reference and variant allele, like in VarScan, for example, but in MAQ output there is just "read depth"?

        Is there any way I can get or derive these 2 counts (reference and variant) from MAQ SNP output?

        Thanks!

        Comment

        Latest Articles

        Collapse

        • SEQadmin2
          Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
          by SEQadmin2



          CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

          Despite this, “CRISPR helped turn genome editing from a specialized technique into
          ...
          07-31-2026, 11:01 AM
        • SEQadmin2
          Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
          by SEQadmin2


          Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

          The systematic characterization of the human proteome has
          ...
          07-20-2026, 11:48 AM

        ad_right_rmr

        Collapse

        News

        Collapse

        Topics Statistics Last Post
        Started by SEQadmin2, Yesterday, 10:35 AM
        0 responses
        9 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 08-06-2026, 07:41 AM
        0 responses
        27 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 08-03-2026, 10:13 AM
        0 responses
        45 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-31-2026, 02:55 AM
        0 responses
        48 views
        0 reactions
        Last Post SEQadmin2  
        Working...