Seqanswers Leaderboard Ad

**BAMseek** · 07-05-2011, 11:46 AM

If your data is microRNA or something where the sequence length is less than the read length, then you might end up reading into adapter or primer sequences. This would possibly show up as over-represented k-mers on the 3' end of the sequences. In this case you might want to trim the adapters before alignment.

**simonandrews** · 07-05-2011, 12:36 PM

If they're in the overrepresented sequences then it probably means that your library was contaminated with primer dimers. If you read through into adapter you'd get slightly different sequences so they would appear in the Kmer plot rather than the overrepresented sequences.

There's not much you can do about the dimers for the data you've already collected. They're easy enough to filter out if you want to remove them, but they probably won't map to whichever genome you're using anyway.

You can normally spot this kind of contamination by doing a BioAnalyser run on your library before sequencing. Others here are better qualified than me to advise on how you might avoid them in the first place.

**DZhang** · 07-05-2011, 06:18 PM

Another source contributing to the over-represented sequences is reads from rRNA genes. Even with rRNA removal, oftentimes you still have high-level rRNA reads.

Douglas

https://www.contigexpress.com

Topics	Statistics	Last Post
Genetic Variants and Diabetes Risk in Childhood Cancer Survivors by seqadmin Started by seqadmin, Today, 08:47 AM	0 responses 12 views 0 likes	Last Post by seqadmin Today, 08:47 AM
Cancer Metastasis: A Deep Dive into Cellular Plasticity by seqadmin Started by seqadmin, 04-11-2024, 12:08 PM	0 responses 60 views 0 likes	Last Post by seqadmin 04-11-2024, 12:08 PM
Proteogenomic Profiles Offer New Clues in Prostate Cancer by seqadmin Started by seqadmin, 04-10-2024, 10:19 PM	0 responses 59 views 0 likes	Last Post by seqadmin 04-10-2024, 10:19 PM
Novel Diagnostic Assay Enhances Ovarian Cancer Detection by seqadmin Started by seqadmin, 04-10-2024, 09:21 AM	0 responses 54 views 0 likes	Last Post by seqadmin 04-10-2024, 09:21 AM

Seqanswers Leaderboard Ad

Announcement

fastqc - overrepresented sequences

Comment

Comment

Comment

Latest Articles

ad_right_rmr

News