Seqanswers Leaderboard Ad

**GenoMax** · 06-18-2015, 07:06 AM

bbduk.sh from BBMap package can do this. If that sequence is at the end of the reads then,

Code:

$ bbduk.sh -Xmx1g in=reads.fq outm=matched.fq outu=unmatched.fq restrictleft=19 k=19 literal=ATTTTTTTTAGAAAAAAAA

In this case, all reads starting with "ATTTTTTTTAGAAAAAAAA" will end up in "matched.fq" and all other reads will end up in "unmatched.fq". Specifically, the command means "look for 19-mers in the leftmost 19 bp of the read", which will require an exact prefix match, though you can relax that if you want.

So you could bin all the reads with your known sequence, then look at the remaining reads to see what they have in common. You can do the same thing with the tail of the read using "restrictright" instead, though you can't use both restrictions at the same time.

**umamayil** · 06-18-2015, 08:18 AM

Hi,
Thanks. The sequence will be either in the middle or in the end. How to separate if the interested sequence is in the middle.

Thanks again
Mayil

**GenoMax** · 06-18-2015, 08:25 AM

Just remove the "restrictleft/right" directive and the entire sequence will be searched.

**umamayil** · 06-19-2015, 05:29 AM

Hi,

Thanks a lot. I will try the commands you have given to me.

Thanks again and have a nice weekend.

Mayil

Topics	Statistics	Last Post
Expanding the Horizons of Cellular Research with the Single Cell Atlas by seqadmin Started by seqadmin, 04-25-2024, 11:49 AM	0 responses 19 views 0 likes	Last Post by seqadmin 04-25-2024, 11:49 AM
Genetic Variants and Diabetes Risk in Childhood Cancer Survivors by seqadmin Started by seqadmin, 04-24-2024, 08:47 AM	0 responses 20 views 0 likes	Last Post by seqadmin 04-24-2024, 08:47 AM
Cancer Metastasis: A Deep Dive into Cellular Plasticity by seqadmin Started by seqadmin, 04-11-2024, 12:08 PM	0 responses 62 views 0 likes	Last Post by seqadmin 04-11-2024, 12:08 PM
Proteogenomic Profiles Offer New Clues in Prostate Cancer by seqadmin Started by seqadmin, 04-10-2024, 10:19 PM	0 responses 61 views 0 likes	Last Post by seqadmin 04-10-2024, 10:19 PM

Seqanswers Leaderboard Ad

Announcement

RNA_seq read separation help

Comment

Comment

Comment

Comment

Latest Articles

ad_right_rmr

News