Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • Amative
    Member
    • Dec 2011
    • 45

    RNA-Seq alignment issue

    Hello all,

    I have been given paired-end RNA-Seq files to align against a couple of references. I used Bowtie2 to do the job. The alignment results were very low in most of the cases (less than 5% overall alignment rate).Now, we are thinking this might be caused by either contamination or mix up samples.

    Any suggestion what to do in such case?
    Thank you in advance
  • swbarnes2
    Senior Member
    • May 2008
    • 910

    #2
    Bioinformatically, there's nothing you can do, other than help the people to know what went wrong.

    For starters, spot-check some random high quality reads, BLAST them against nr, see if you can determine what they are.

    Try aligning to the whole genome to see how much of the library was genomic.

    See if there are certain highly repetitive reads (like Illumina adapters) taking up a lot of reads.

    And of course see if the run overall was of good enough quality for you to believe that your reads are accurate.

    Comment

    • rboettcher
      Member
      • Oct 2010
      • 71

      #3
      Hi Amative,

      what kind of reference did you provide? Bowtie2 is not splicing aware, so it is not able to deal with reads spanning splice junctions. Therefore, it can only be used to align against the transcriptome (for RNAseq). This is why TopHat was created to align against the whole genome.

      Regards

      Comment

      • Amative
        Member
        • Dec 2011
        • 45

        #4
        Thanks swbarnes2 & rboettcher,

        @swbarnes2
        • I tried to blast the first ten reads from one of the samples I have, blast results were not that good. I tried to align against the available sequences of the two of the top blast hits. Same low alignment rate.
        • I checked for adapters, sequences are already trimmed.


        @rboettcher
        Yes, I am aligning against the transcriptome sequences.

        Comment

        • GenoMax
          Senior Member
          • Feb 2008
          • 7142

          #5
          Originally posted by Amative View Post
          I tried to blast the first ten reads from one of the samples I have, blast results were not that good.
          You probably want to go into the file some ways. With illumina atleast the first hundred (or more) sequences may not represent the best of the lot since they are generally from the edge of the flowcell/start of the lane.

          You may also want to use this tool to do some screening: http://www.bioinformatics.babraham.a.../fastq_screen/

          Comment

          • Amative
            Member
            • Dec 2011
            • 45

            #6
            Thanks GenoMax, for the suggestion I am working on it.

            I like the fastq_screen It saves some time!

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
              by SEQadmin2


              Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

              The systematic characterization of the human proteome has
              ...
              Yesterday, 11:48 AM
            • SEQadmin2
              Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
              by SEQadmin2



              Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
              ...
              07-09-2026, 11:10 AM
            • SEQadmin2
              Cancer Drug Resistance: The Lingering Barrier to Rising Survival
              by SEQadmin2



              Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

              There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
              07-08-2026, 05:17 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, Yesterday, 11:10 AM
            0 responses
            9 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-13-2026, 10:26 AM
            0 responses
            30 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-09-2026, 10:04 AM
            0 responses
            41 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-08-2026, 10:08 AM
            0 responses
            25 views
            0 reactions
            Last Post SEQadmin2  
            Working...