Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • fastqc kmer relative enrichment

    I was wondering how to interpret the Kmer content graph in fastqc. I have attached it. It seems from my graph that 50% of all reads have the 6 kmers listed ? Is this a normal graph ?
    Attached Files
    Last edited by mattanswers; 08-13-2013, 10:29 AM. Reason: need to attach .png file

  • #2
    Is this RNA-seq data?

    Following thread may be useful (though it refers to a MiSeq run the issue is applicable to illumina sequencing in general): http://seqanswers.com/forums/showthread.php?t=30448

    One more: http://seqanswers.com/forums/showthread.php?t=17219
    Last edited by GenoMax; 08-13-2013, 11:57 AM.

    Comment


    • #3
      Thank you very much, GenoMax, for the links. They are very informative.

      This is RNA-Seq. It seems the first 12 or so bases are due to 'random' priming, but I was also wondering about why the lines on the graph stay up at ~50% for the length of the graph ? Random priming would explain the first 12 or so bases, but why the steady % for the rest of the sequence ?

      Comment


      • #4
        Originally posted by mattanswers View Post
        Thank you very much, GenoMax, for the links. They are very informative.

        This is RNA-Seq. It seems the first 12 or so bases are due to 'random' priming, but I was also wondering about why the lines on the graph stay up at ~50% for the length of the graph ? Random priming would explain the first 12 or so bases, but why the steady % for the rest of the sequence ?

        Comment


        • #5
          Thanks again for your help, GenoMax.

          My sequence length is only 50 bases and the quality is very good.

          From what I read on the linked site, it seems that I have 6 kmers that are 50-fold enriched throughout the length of my sequence. But what does this mean in terms of sample quality ?

          If I have 25-30 million reads and there is a 50 fold enrichment of these kmers (most likely I would guess from the adaptor) then how many sequences does that affect ? So, if there were 100,000 sequences in which had adaptor sequence at various positions other than the end of the sequence what would the fold-enrichment be ? 100,000 affected sequences may be enough to make the fold-enrichment high, but they are only a small percentage of the total. On the other hand, if I had a much smaller number of total sequences, then the fold-enrichment may be a problem. So, I guess I want to know how to relate fold-enrichment and total number of sequences in order to tell if the fold-enrichment is a problem or just from an insignificant part of the total.

          Comment

          Latest Articles

          Collapse

          • seqadmin
            Latest Developments in Precision Medicine
            by seqadmin



            Technological advances have led to drastic improvements in the field of precision medicine, enabling more personalized approaches to treatment. This article explores four leading groups that are overcoming many of the challenges of genomic profiling and precision medicine through their innovative platforms and technologies.

            Somatic Genomics
            “We have such a tremendous amount of genetic diversity that exists within each of us, and not just between us as individuals,”...
            Yesterday, 01:16 PM
          • seqadmin
            Recent Advances in Sequencing Analysis Tools
            by seqadmin


            The sequencing world is rapidly changing due to declining costs, enhanced accuracies, and the advent of newer, cutting-edge instruments. Equally important to these developments are improvements in sequencing analysis, a process that converts vast amounts of raw data into a comprehensible and meaningful form. This complex task requires expertise and the right analysis tools. In this article, we highlight the progress and innovation in sequencing analysis by reviewing several of the...
            05-06-2024, 07:48 AM

          ad_right_rmr

          Collapse

          News

          Collapse

          Topics Statistics Last Post
          Started by seqadmin, Yesterday, 07:15 AM
          0 responses
          12 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 05-23-2024, 10:28 AM
          0 responses
          15 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 05-23-2024, 07:35 AM
          0 responses
          16 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 05-22-2024, 02:06 PM
          0 responses
          10 views
          0 likes
          Last Post seqadmin  
          Working...
          X