Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • Lien
    Member
    • Dec 2009
    • 47

    #1

    Interpretation of Picard's MEDIAN_CV_COVERAGE

    Dear all,

    We just performed an RNA-seq experiment from human samples. To get some idea of the quality, I ran Picard's CollectRNASeqMetrics.
    In the output, I find 'MEAN_CV_COVERAGE'. The explanation for this value is also on the website: the median CV of coverage of the 1000 most highly expressed transcripts.

    I'm not really sure what CV means. And how can I interpret this value? The value I get is 0.48. If I think of this as 48x coverage, this seems really a lot to me. Especially since visualisation with IGV shows me a lower coverage.

    Any help will be appreciated!

    Thanks
    Lien
  • Bukowski
    Senior Member
    • Jan 2010
    • 388

    #2
    Originally posted by Lien View Post
    Dear all,

    We just performed an RNA-seq experiment from human samples. To get some idea of the quality, I ran Picard's CollectRNASeqMetrics.
    In the output, I find 'MEAN_CV_COVERAGE'. The explanation for this value is also on the website: the median CV of coverage of the 1000 most highly expressed transcripts.

    I'm not really sure what CV means. And how can I interpret this value? The value I get is 0.48. If I think of this as 48x coverage, this seems really a lot to me. Especially since visualisation with IGV shows me a lower coverage.

    Any help will be appreciated!

    Thanks
    Lien
    I always assumed CV referred to the coefficient of variation:



    As in 'mean coefficient of variation of coverage'

    Although happy to be proved wrong.

    Comment

    • jstjohn
      Member
      • Jun 2010
      • 35

      #3
      from the javadoc

      According to the javadoc:
      "The median CV of coverage of the 1000 most highly expressed transcripts. Ideal value = 0."

      What is confusing me though is it looks like this involves the coefficient of variation (sample sd)/(sample mean). Since this number is reported for a single RNA-seq bam file, I am wondering where the population is coming from? It looks like the only way to get at an sd or mean from the above description is by looking at the coverage of the top 1000 transcripts. That would give you a single number, not something you could take the median of. Pretty sure I am missing something here.


      http://picard.sourceforge.net/javadoc/net/sf/picard/analysis/RnaSeqMetrics.html#MEDIAN_CV_COVERAGE

      Comment

      • pengchy
        Senior Member
        • Feb 2009
        • 116

        #4
        What use of MEDIAN_CV_COVERAGE

        Originally posted by jstjohn View Post
        According to the javadoc:
        "The median CV of coverage of the 1000 most highly expressed transcripts. Ideal value = 0."

        What is confusing me though is it looks like this involves the coefficient of variation (sample sd)/(sample mean). Since this number is reported for a single RNA-seq bam file, I am wondering where the population is coming from? It looks like the only way to get at an sd or mean from the above description is by looking at the coverage of the top 1000 transcripts. That would give you a single number, not something you could take the median of. Pretty sure I am missing something here.


        http://picard.sourceforge.net/javadoc/net/sf/picard/analysis/RnaSeqMetrics.html#MEDIAN_CV_COVERAGE
        According to the definition, MEDIAN_CV_COVERAGE is "The median CV of coverage of the 1000 most highly expressed transcripts. Ideal value = 0." I think the variation should be very large, because there will be several gene expressed at very high value. Then, why the ideal value is "0"? "0" means the variation is zero, it not consistent with the fact.

        Comment

        Latest Articles

        Collapse

        • SEQadmin2
          Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
          by SEQadmin2



          CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

          Despite this, “CRISPR helped turn genome editing from a specialized technique into
          ...
          07-31-2026, 11:01 AM
        • SEQadmin2
          Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
          by SEQadmin2


          Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

          The systematic characterization of the human proteome has
          ...
          07-20-2026, 11:48 AM

        ad_right_rmr

        Collapse

        News

        Collapse

        Topics Statistics Last Post
        Started by SEQadmin2, 08-06-2026, 07:41 AM
        0 responses
        16 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 08-03-2026, 10:13 AM
        0 responses
        32 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-31-2026, 02:55 AM
        0 responses
        42 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-24-2026, 12:17 PM
        0 responses
        26 views
        0 reactions
        Last Post SEQadmin2  
        Working...