Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • lg36
    Member
    • Mar 2012
    • 12

    #1

    Please Help - SNP Calling in relation to Consensus Sequence Quality

    Dear All,

    I've posted on this before and attracted no interest at all. So I'd be very grateful to get your opinions. I'm relatively new to NGS sequence analysis so apologies for any naive assumptions.

    As I understand (perhaps misunderstand) if a SNP is deemed to be present (after all the applied SNP filters, scores etc) then (most of the time) it will appear in the consensus sequence???

    Is it therefore right to think if the consensus sequence is of a 'higher quality', the SNP's too will be of a 'higher quality'??

    I'm defining quality of the consensus by the ability to reflect the presence or absence of certain oligonucleotides that we know to be there from PCR studies in the lab.

    What I'm doing at the moment is trying to set the parameters correctly (paired end width, etc) in order to reach a consensus sequence that is correct according to what we know in the lab. Then I was planning on calling the SNPs based on the mapped reads we used for the best consensus sequence. Is there any validity at all in this approach, or are the two things SNP calling and consensus sequence production too 'unlinked' to make it invalid????

    Many thanks in advance, apologies for any misconceptions

    lg36
  • pag
    Member
    • May 2012
    • 72

    #2
    Most base callers either call "n" at SNP location or pick one or the other (combination of sharpest peak and highest peak as scaled to average intensity for that base). Very few use IUPAC ambiguity codes to convey the potential difference. Quality for those locations, assuming an unbiased amplification level should be approximately 1/2 of the surrounding bases (I believe it's linearly scaled at least). So, if scripts don't already exist, you should be able to compare quality from one base to the neighboring two bases and see if the qual is significantly lower. Then you look at the trace itself (if available) to see if this is a potential SNP.

    It's more difficult if it's a "chimera" or extended polymorphism or if there's a "frameshift."

    I'd be interested to know if there are any chimera/SNP deconvolver programs or scripts out there for use that I can supply the reference sequence and it can let me know the alternate read. I was staring at at two overlapping reads that were significantly out-of-phase with each other yesterday and made about 50 base calls manually over an hour or two. Hardly the model of efficiency

    Comment

    • lg36
      Member
      • Mar 2012
      • 12

      #3
      Thanks pag for your help. So, is it true to say if I can replicate what we know to be the reality of the consensus sequence (from our lab findings), then the SNPs that are output from the same mapped reads will be more reliable?

      Comment

      Latest Articles

      Collapse

      • SEQadmin2
        Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
        by SEQadmin2








        CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

        Despite this, “CRISPR helped turn genome editing
        ...
        07-31-2026, 11:01 AM
      • SEQadmin2
        Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
        by SEQadmin2


        Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

        The systematic characterization of the human proteome has
        ...
        07-20-2026, 11:48 AM
      • SEQadmin2
        Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
        by SEQadmin2



        Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
        ...
        07-09-2026, 11:10 AM

      ad_right_rmr

      Collapse

      News

      Collapse

      Topics Statistics Last Post
      Started by SEQadmin2, 07-31-2026, 02:55 AM
      0 responses
      20 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 07-24-2026, 12:17 PM
      0 responses
      16 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 07-23-2026, 11:41 AM
      0 responses
      16 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 07-20-2026, 11:10 AM
      0 responses
      26 views
      0 reactions
      Last Post SEQadmin2  
      Working...