Hello, I have a multi-sample VCF file, with about 4000 contigs/loci represented. Many of these contigs contain multiple SNPs, meaning the SNPs are linked to one another in this case. However, some of my downstream analyses do not handle linked markers well, so I would like to be able to filter my VCF file so that I am left with a single SNP per contig. If there were a way to randomly select one SNP per contig, that would be great. Any ideas about how to achieve this? As far as I can tell, VCFtools does not allow this.
Seqanswers Leaderboard Ad
Collapse
Announcement
Collapse
No announcement yet.
X
-
vcftools does have a --thin option, and if you set it to the maximum contig size then there will be only 1 SNP per contig (--thin 100000, for example).
It may select the first SNP in the contig, though.Providing nextRAD genotyping and PacBio sequencing services. http://snpsaurus.com
-
Great, thanks SNPsaurus, this is working nicely. I would still like to do a 'random SNP per locus' option, and compare results to what I'm getting with this 'first SNP per locus' method, but this might require some more advanced programming skills.
Comment
Latest Articles
Collapse
-
by seqadmin
Technological advances have led to drastic improvements in the field of precision medicine, enabling more personalized approaches to treatment. This article explores four leading groups that are overcoming many of the challenges of genomic profiling and precision medicine through their innovative platforms and technologies.
Somatic Genomics
“We have such a tremendous amount of genetic diversity that exists within each of us, and not just between us as individuals,”...-
Channel: Articles
05-24-2024, 01:16 PM -
-
by seqadmin
The sequencing world is rapidly changing due to declining costs, enhanced accuracies, and the advent of newer, cutting-edge instruments. Equally important to these developments are improvements in sequencing analysis, a process that converts vast amounts of raw data into a comprehensible and meaningful form. This complex task requires expertise and the right analysis tools. In this article, we highlight the progress and innovation in sequencing analysis by reviewing several of the...-
Channel: Articles
05-06-2024, 07:48 AM -
ad_right_rmr
Collapse
News
Collapse
Topics | Statistics | Last Post | ||
---|---|---|---|---|
Started by seqadmin, 06-03-2024, 06:55 AM
|
0 responses
12 views
0 likes
|
Last Post
by seqadmin
06-03-2024, 06:55 AM
|
||
Started by seqadmin, 05-30-2024, 03:16 PM
|
0 responses
26 views
0 likes
|
Last Post
by seqadmin
05-30-2024, 03:16 PM
|
||
Comprehensive Sequencing of Great Ape Sex Chromosomes Yields Insights into Evolution and Genetic Variability
by seqadmin
Started by seqadmin, 05-29-2024, 01:32 PM
|
0 responses
29 views
0 likes
|
Last Post
by seqadmin
05-29-2024, 01:32 PM
|
||
Started by seqadmin, 05-24-2024, 07:15 AM
|
0 responses
215 views
0 likes
|
Last Post
by seqadmin
05-24-2024, 07:15 AM
|
Comment