Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • rdsqc22
    Junior Member
    • Nov 2013
    • 7

    Odd statistical differences in Cuffdiff output?

    Hi,

    I've been aligning and counting some RNA-seq reads with SHRiMP and Cuffdiff, doung the same analysis with both an older genome assembly and a newer one, and I found an interesting possible discrepancy in my Cuffdiff output. If anyone could help explain it would be much appreciated.

    Basically, I noticed a number of different genes where the expression levels was similar between the two assemblies, yet for some reason Cuffdiff was reporting wildly different significance results between the two. For example:

    gene locus sample_1 sample_2 status value_1 value_2 log2(fold_change) test_stat p_value q_value significant
    Asmb 5: Gfap 10:90763148-90771847 lineA lineN OK 484.283 11.1909 -5.43545 1.62926 0.103258 0.394632 no
    Asmb 4: Gfap 10:92059880-92068555 lineA lineN OK 526.67 12.77 -5.36606 4.09233 4.27058E-005 0.00052085 yes

    Both were run with the same cuffdiff binary (Cuffdiff 2.0.2), with the exact same command (adjusted for the appropriate assembly), with an FDR of 0.05. It would stand to reason that the results are similar between the binaries- Line A is much more upregulated than line N in both cases, and the only statistical difference I can see that might have an effect is that the size of the gene in the assembly changed by 24 nucleotides, out of just under 10000.

    If the gene size, fold change, and FPKM values are so similar, why are the statistical values so wildly different? This does not make sense to me.

    Thanks!
  • Wallysb01
    Senior Member
    • Feb 2011
    • 286

    #2
    Agree this is odd.

    Can you post your commands? That might help us get a little more information.

    Also, how much changed as far as number of genes in your annotation, or what percent of reads are mapping to each genome?

    You should probably update your version of cufflinks too. Even though these are the same version, we are a long way from v2.0.2 now.

    Have you tried doing this with DESeq2? Might be worth seeing if this is something that is cufflinks specific or more broadly true about something going on in your new genome and genome annotation.

    Comment

    • rdsqc22
      Junior Member
      • Nov 2013
      • 7

      #3
      My command, used in both cases, is simply:

      cuffdiff --FDR 0.05 -u -b genome.fa -p 4 -L lineA,lineN -o cuffdiffout genes.gtf lineA.bam lineN.bam

      We downgraded to 2.0.2 because we had run into trouble with version 2.1.1, which is what had been installed previously- our sequencing center uses 2.0.2, which is why that version was chosen. I'm currently running another run with 2.2.0.

      This is the older assembly used: http://www.ncbi.nlm.nih.gov/assembly/237618/
      And the newer one: http://www.ncbi.nlm.nih.gov/assembly/382928

      I'm familiar with Cuffdiff, which is why it was used. I'll try using DeSeq2, though.

      Comment

      Latest Articles

      Collapse

      • SEQadmin2
        Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
        by SEQadmin2


        Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

        The systematic characterization of the human proteome has
        ...
        07-20-2026, 11:48 AM
      • SEQadmin2
        Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
        by SEQadmin2



        Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
        ...
        07-09-2026, 11:10 AM
      • SEQadmin2
        Cancer Drug Resistance: The Lingering Barrier to Rising Survival
        by SEQadmin2



        Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

        There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
        07-08-2026, 05:17 AM

      ad_right_rmr

      Collapse

      News

      Collapse

      Topics Statistics Last Post
      Started by SEQadmin2, 07-24-2026, 12:17 PM
      0 responses
      26 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 07-23-2026, 11:41 AM
      0 responses
      21 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 07-20-2026, 11:10 AM
      0 responses
      210 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 07-13-2026, 10:26 AM
      0 responses
      78 views
      0 reactions
      Last Post SEQadmin2  
      Working...