Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • fastqc kmer relative enrichment

    I was wondering how to interpret the Kmer content graph in fastqc. I have attached it. It seems from my graph that 50% of all reads have the 6 kmers listed ? Is this a normal graph ?
    Attached Files
    Last edited by mattanswers; 08-13-2013, 10:29 AM. Reason: need to attach .png file

  • #2
    Is this RNA-seq data?

    Following thread may be useful (though it refers to a MiSeq run the issue is applicable to illumina sequencing in general): http://seqanswers.com/forums/showthread.php?t=30448

    One more: http://seqanswers.com/forums/showthread.php?t=17219
    Last edited by GenoMax; 08-13-2013, 11:57 AM.

    Comment


    • #3
      Thank you very much, GenoMax, for the links. They are very informative.

      This is RNA-Seq. It seems the first 12 or so bases are due to 'random' priming, but I was also wondering about why the lines on the graph stay up at ~50% for the length of the graph ? Random priming would explain the first 12 or so bases, but why the steady % for the rest of the sequence ?

      Comment


      • #4
        Originally posted by mattanswers View Post
        Thank you very much, GenoMax, for the links. They are very informative.

        This is RNA-Seq. It seems the first 12 or so bases are due to 'random' priming, but I was also wondering about why the lines on the graph stay up at ~50% for the length of the graph ? Random priming would explain the first 12 or so bases, but why the steady % for the rest of the sequence ?

        Comment


        • #5
          Thanks again for your help, GenoMax.

          My sequence length is only 50 bases and the quality is very good.

          From what I read on the linked site, it seems that I have 6 kmers that are 50-fold enriched throughout the length of my sequence. But what does this mean in terms of sample quality ?

          If I have 25-30 million reads and there is a 50 fold enrichment of these kmers (most likely I would guess from the adaptor) then how many sequences does that affect ? So, if there were 100,000 sequences in which had adaptor sequence at various positions other than the end of the sequence what would the fold-enrichment be ? 100,000 affected sequences may be enough to make the fold-enrichment high, but they are only a small percentage of the total. On the other hand, if I had a much smaller number of total sequences, then the fold-enrichment may be a problem. So, I guess I want to know how to relate fold-enrichment and total number of sequences in order to tell if the fold-enrichment is a problem or just from an insignificant part of the total.

          Comment

          Latest Articles

          Collapse

          • seqadmin
            Recent Advances in Sequencing Technologies
            by seqadmin



            Innovations in next-generation sequencing technologies and techniques are driving more precise and comprehensive exploration of complex biological systems. Current advancements include improved accessibility for long-read sequencing and significant progress in single-cell and 3D genomics. This article explores some of the most impactful developments in the field over the past year.

            Long-Read Sequencing
            Long-read sequencing has seen remarkable advancements,...
            12-02-2024, 01:49 PM
          • seqadmin
            Genetic Variation in Immunogenetics and Antibody Diversity
            by seqadmin



            The field of immunogenetics explores how genetic variations influence immune responses and susceptibility to disease. In a recent SEQanswers webinar, Oscar Rodriguez, Ph.D., Postdoctoral Researcher at the University of Louisville, and Ruben Martínez Barricarte, Ph.D., Assistant Professor of Medicine at Vanderbilt University, shared recent advancements in immunogenetics. This article discusses their research on genetic variation in antibody loci, antibody production processes,...
            11-06-2024, 07:24 PM

          ad_right_rmr

          Collapse

          News

          Collapse

          Topics Statistics Last Post
          Started by seqadmin, 12-02-2024, 09:29 AM
          0 responses
          152 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 12-02-2024, 09:06 AM
          0 responses
          51 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 12-02-2024, 08:03 AM
          0 responses
          43 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 11-22-2024, 07:36 AM
          0 responses
          76 views
          0 likes
          Last Post seqadmin  
          Working...
          X