Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • A heretically simple approach to variant calling

    Hi everyone,

    we had a look at the distribution of heterozygous allele frequencies in NGS datasets and found that their variance is larger than expected by a bionomial distribution (http://www.ncbi.nlm.nih.gov/pubmed?term=22127862). For every variant caller this means a binomial prior distribution is not the right choice and might lead to false negative calls. We also found that a simple frequency classifier (heterozygous if covered by more the 20 reads and variant allele between 14% and 86%) is more sensitive at comparable specificity for high quality data, compared to default setting of most standard calling tools.

    Is anyone aware of a fast tool, that allows to apply such a frequency filter directly on a .bam file?

    cheers,

    peter

  • #2
    You can ask samtools mpileup to print out the nucleotide pileups for each position in bam file. Parsing that should be fairly simple with a script.

    Comment


    • #3
      the samtools mpileup output can be piped into VarScan to apply a coverage and frequency filter:
      samtools pileup -f reference.fasta myData.bam | java -jar VarScan.v2.2.jar pileup2snp --min-coverage 20 --min-var-freq 0.14
      see: http://varscan.sourceforge.net/using-varscan.html

      Comment


      • #4
        Very nice paper.

        Sometimes simpler is better.

        Comment


        • #5
          This is awesome. Moving away from big fancy well-established tools to something like the "14-86%" rule is scary though.

          Comment

          Latest Articles

          Collapse

          • seqadmin
            Recent Advances in Sequencing Technologies
            by seqadmin



            Innovations in next-generation sequencing technologies and techniques are driving more precise and comprehensive exploration of complex biological systems. Current advancements include improved accessibility for long-read sequencing and significant progress in single-cell and 3D genomics. This article explores some of the most impactful developments in the field over the past year.

            Long-Read Sequencing
            Long-read sequencing has seen remarkable advancements,...
            12-02-2024, 01:49 PM
          • seqadmin
            Genetic Variation in Immunogenetics and Antibody Diversity
            by seqadmin



            The field of immunogenetics explores how genetic variations influence immune responses and susceptibility to disease. In a recent SEQanswers webinar, Oscar Rodriguez, Ph.D., Postdoctoral Researcher at the University of Louisville, and Ruben Martínez Barricarte, Ph.D., Assistant Professor of Medicine at Vanderbilt University, shared recent advancements in immunogenetics. This article discusses their research on genetic variation in antibody loci, antibody production processes,...
            11-06-2024, 07:24 PM

          ad_right_rmr

          Collapse

          News

          Collapse

          Topics Statistics Last Post
          Started by seqadmin, 12-02-2024, 09:29 AM
          0 responses
          153 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 12-02-2024, 09:06 AM
          0 responses
          51 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 12-02-2024, 08:03 AM
          0 responses
          43 views
          0 likes
          Last Post seqadmin  
          Started by seqadmin, 11-22-2024, 07:36 AM
          0 responses
          76 views
          0 likes
          Last Post seqadmin  
          Working...
          X