Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • tusharbiot
    Junior Member
    • Feb 2012
    • 7

    Pre processing of Illumina data

    Hi all!

    Though there have been many similar posts, but to have an updated (and more recent) thought, hence I am posting it again.

    I have some Illumina GAIIX paired end data in two separate files (with their suffix as _1 and _2). Before putting them for assembly, I need to perform certain pre processing of these file to alleviate its quality (probably FASTQ MASKER by quality control on GLAXY web tools). My query is " Should I join these paired end data first and then go on with the quality treatments OR should I just perform the quality treatment and then join the two files.

    Thanks in advance
    Any help will be deeply appreciated.
  • ians
    Member
    • Aug 2011
    • 53

    #2
    Originally posted by tusharbiot View Post
    My query is " Should I join these paired end data first and then go on with the quality treatments OR should I just perform the quality treatment and then join the two files.
    what do you mean when you say, "join"?

    Comment

    • tusharbiot
      Junior Member
      • Feb 2012
      • 7

      #3
      With join I meant should I combine the two files into one (basically the tool I am using just concatenates the two sequences from a single locus/index).

      Comment

      • faozhi
        Junior Member
        • Dec 2011
        • 5

        #4
        I know I didn't combine them during trimming based on qual...

        Comment

        • tusharbiot
          Junior Member
          • Feb 2012
          • 7

          #5
          Originally posted by faozhi View Post
          I know I didn't combine them during trimming based on qual...
          faozhi, so that would mean you combined them after the quality screening.

          Comment

          • SES
            Senior Member
            • Mar 2010
            • 275

            #6
            Originally posted by tusharbiot View Post
            Hi all!

            Though there have been many similar posts, but to have an updated (and more recent) thought, hence I am posting it again.

            I have some Illumina GAIIX paired end data in two separate files (with their suffix as _1 and _2). Before putting them for assembly, I need to perform certain pre processing of these file to alleviate its quality (probably FASTQ MASKER by quality control on GLAXY web tools). My query is " Should I join these paired end data first and then go on with the quality treatments OR should I just perform the quality treatment and then join the two files.

            Thanks in advance
            Any help will be deeply appreciated.
            It actually makes no difference at all, but you don't want to create extra work for yourself. With that in mind, you should probably trim the files separately, that way it will be easy to identify which read pairs were retained after the trimming and which reads should be treated as single-end.

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
              by SEQadmin2


              Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

              The systematic characterization of the human proteome has
              ...
              07-20-2026, 11:48 AM
            • SEQadmin2
              Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
              by SEQadmin2



              Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
              ...
              07-09-2026, 11:10 AM
            • SEQadmin2
              Cancer Drug Resistance: The Lingering Barrier to Rising Survival
              by SEQadmin2



              Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

              There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
              07-08-2026, 05:17 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, Yesterday, 12:17 PM
            0 responses
            13 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-23-2026, 11:41 AM
            0 responses
            13 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-20-2026, 11:10 AM
            0 responses
            23 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-13-2026, 10:26 AM
            0 responses
            37 views
            0 reactions
            Last Post SEQadmin2  
            Working...