Seqanswers Leaderboard Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • jorgebm
    Member
    • Feb 2010
    • 18

    Doubts about GATK "raw data processing" step for SOliD exome data

    Hello,


    I just received SOLiD 75bp fragment (SE) exome data. I haven't previosly worked with SOLiD data and I'm not sure the suitable way to handle them for a further GATK based variant detection. In concrete I have doubts about:

    - filtering PCR duplicated

    I'm undecided to filter or not them using picard markduplicates. Being SingleEnd data I'm inclined not to do it to avoid lose "true" reads (not duplicates) at expense of generate "dirty" variant calls. What's you opinion?.

    - base recalibration

    Due to special caracteristics of colorspace read seems to be unnecessary to make base recalibration. I checked "1000 Genomes" paper (http://www.nature.com/nature/journal...ture09534.html) where based recalibration was avoided for SOLiD data. By other hand, in GATK documentation (http://www.broadinstitute.org/gsa/wi..._recalibration) "PrimerRoundCovariate" covariate is described as "The primer round for this base (only meaningful for SOLiD reads).". I'm confused and do not know if the recalibration of bases is a recommended step for SOLiD.
    Any advice?

    I would appreciate your help.


    Thank you
  • szilva
    Member
    • Aug 2009
    • 16

    #2
    Actually I also have similar doubts (though working with paired data). Do you have any progress or hints on this topic?

    Comment

    • jorgebm
      Member
      • Feb 2010
      • 18

      #3
      Unfortunately, don't. But I think If you've PairEnd data you could safetly remove PCR duplicates. However, I neved did it for SOLiD.

      About "base reclaibration" I'm going to try it. I'm waiting for finishing alignment step to try it. I'll let you know then.

      By other hand, I'm using BFAST to align to colospace but it is extremely slow. What alignerare you using?

      Comment

      Latest Articles

      Collapse

      • seqadmin
        New Genomics Tools and Methods Shared at AGBT 2025
        by seqadmin


        This year’s Advances in Genome Biology and Technology (AGBT) General Meeting commemorated the 25th anniversary of the event at its original venue on Marco Island, Florida. While this year’s event didn’t include high-profile musical performances, the industry announcements and cutting-edge research still drew the attention of leading scientists.

        The Headliner
        The biggest announcement was Roche stepping back into the sequencing platform market. In the years since...
        03-03-2025, 01:39 PM
      • seqadmin
        Investigating the Gut Microbiome Through Diet and Spatial Biology
        by seqadmin




        The human gut contains trillions of microorganisms that impact digestion, immune functions, and overall health1. Despite major breakthroughs, we’re only beginning to understand the full extent of the microbiome’s influence on health and disease. Advances in next-generation sequencing and spatial biology have opened new windows into this complex environment, yet many questions remain. This article highlights two recent studies exploring how diet influences microbial...
        02-24-2025, 06:31 AM

      ad_right_rmr

      Collapse

      News

      Collapse

      Topics Statistics Last Post
      Started by seqadmin, 03-20-2025, 05:03 AM
      0 responses
      17 views
      0 reactions
      Last Post seqadmin  
      Started by seqadmin, 03-19-2025, 07:27 AM
      0 responses
      18 views
      0 reactions
      Last Post seqadmin  
      Started by seqadmin, 03-18-2025, 12:50 PM
      0 responses
      19 views
      0 reactions
      Last Post seqadmin  
      Started by seqadmin, 03-03-2025, 01:15 PM
      0 responses
      186 views
      0 reactions
      Last Post seqadmin  
      Working...