Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • AJenkins
    Member
    • Nov 2013
    • 13

    RNA-seq not showing splice junctions

    I have two RNA-seq data sets. When one is run through Cufflinks I see isoforms but when I run the other I see no isoforms, only single exons/transcripts, none with multiple exons. Further investigating I see that my alignments don't show spanning splice junctions using IGV to visualize. I have attached two png files to show what it should look like (normal.png) and what it does look like (broad.png). Any insights onto why this may be happening?



  • dpryan
    Devon Ryan
    • Jul 2011
    • 3478

    #2
    Cufflinks won't be remapping anything, just estimating abundances and looking for new transcripts (among other related things). Have you tried looking at the same area in both alignments? Also, what about just running "samtools view alignments_lacking_apparent_splicing.bam | cut -f 6 | grep N" to double check things, as that'll check for splicing in the whole file instead of you needing to spot-check things.

    Comment

    • dpryan
      Devon Ryan
      • Jul 2011
      • 3478

      #3
      BTW, did you use the same aligner in both cases?

      Comment

      • AJenkins
        Member
        • Nov 2013
        • 13

        #4
        dpryan, I just contacted the institute that aligned the files and the didn't use any splice junction aligned like tophat or PASTA, which I foolishly assumed. As they won't send me the read files as they are "too large" in their words, what is the best way to extract sequences from .bam files that I can run on tophat?

        Comment

        • dpryan
          Devon Ryan
          • Jul 2011
          • 3478

          #5
          Wow, they sounds like assholes. I bet the original reads had crap alignment rates or something, suggesting that the data may be bad to begin with (though you wouldn't be able to see this from the alignments). You can just use the SamToFastq command from picard tools to extract the raw reads for realignment. The bad news is that, since they didn't use a splice-aware aligner (if these really are RNAseq reads, then these people are morons), a lot of the spliced reads probably just didn't align, so they probably aren't included in the BAM file.

          Comment

          • AJenkins
            Member
            • Nov 2013
            • 13

            #6
            I have found the culprit of the problem. It seems that the sequencing institute aligned the reads to the genome, which makes us think that they received reads and then aligned them genome. In actuality the sequencer is automated to align the reads and report a BAM file to indicate possible alignment sites, therefore consolidating data. Using picard to get the sequences to run through the proper Tophat-cufflinks protocol. Thanks for all the help everyone

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
              by SEQadmin2


              Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

              The systematic characterization of the human proteome has
              ...
              07-20-2026, 11:48 AM
            • SEQadmin2
              Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
              by SEQadmin2



              Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
              ...
              07-09-2026, 11:10 AM
            • SEQadmin2
              Cancer Drug Resistance: The Lingering Barrier to Rising Survival
              by SEQadmin2



              Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

              There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
              07-08-2026, 05:17 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, 07-24-2026, 12:17 PM
            0 responses
            25 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-23-2026, 11:41 AM
            0 responses
            20 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-20-2026, 11:10 AM
            0 responses
            27 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-13-2026, 10:26 AM
            0 responses
            38 views
            0 reactions
            Last Post SEQadmin2  
            Working...