Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • vcorby
    Junior Member
    • Jan 2012
    • 1

    Using Cufflinks Cuffdiff output for 2 factor analysis

    I have an experiment where I am testing for the effect of two factors (age and diet) on gene expression. I have three biological replicates for each age*diet combination.


    I would like to test for both main effects and an interaction term on gene expression using a negative binomial regression model, but I see that others prefer a Poisson model. At any rate, I have been using Cufflinks and Cuffdiff because I like the concept of looking at differences in isoform abundance across treatments, however, I see that the best analysis is only a pairwise comparison of FPKM in one condition compared to another. I used cuffdiff for a first look, indicating that I had three biological replicates. I entered the following to compare gene expression in diets A vs B, for age A only (edited, of course so people see where I'm going):
    > cuffdiff -p 8 -o outputdirectory ReferenceGTF -L DietA,DietB AgeADietARep1.bam,AgeADietARep2.bam,AgeADietARep3.bam AgeADietBRep1.bam,AgeADietBRep2.bam,AgeADietBRep3.bam

    This gives me an output comparing, in pairwise fashion, whether Diet A has significantly different gene expression compared to Diet B, all at Age A. However, I would like to do a proper 2-factor analysis on these data.

    So my question is the following: if I run a cuffdiff analysis and say that each age*diet combination is essentially its own replicate by entering this:
    > cuffdiff -p 8 -o outputdirectory ReferenceGTF -L AgeADietARep1,AgeADietARep2,AgeADietARep3,AgeADietBRep1,AgeADietBRep2,AgeADietBRep3,AgeBDietARep1,AgeBDietARep2,AgeBDietARep2,AgeBDietBRep1,AgeBDietBRep2,AgeBDietBRep3 AgeADietARep1.bam AgeADietARep2.bam AgeADietARep3.bam AgeADietBRep1.bam AgeADietBRep2.bam AgeADietBRep3.bam AgeBDietARep1.bam AgeBDietARep2.bam AgeBDietARep2.bam AgeBDietBRep1.bam AgeBDietBRep2.bam AgeBDietBRep3.bam

    can I get an estimate for isoform abundance (FPKM) in each library in the cuffdiff file labeled 'genes.fpkm_tracking' and then input those values into a downstream analysis that tests for significant effects of age, diet, or age*diet using a Negative Binomial regression? It appears (from the cufflinks documentation) that the FPKMs in this 'genes.fpkm_tracking' file are normalized to account for, say, differences in library size and overdispersion, however, DESeq does not account for isoform abundance, which is appealing to me.

    Thanks for any comments on this.

    Vanessa

Latest Articles

Collapse

  • SEQadmin2
    From Collection to Sequencing: Why Sample Preparation and Preservation Define Sequencing Data
    by SEQadmin2


    Data variability is still an issue in sequencing technologies despite the advances in reproducibility and accuracy of these platforms. But the problem does not originate in the sequencing itself, but in the previous steps, before the sample reaches the sequencer.


    The first step is collection, followed by preservation and sample preparation for analysis. Most scientists overlook those steps, but not being careful might just be skewing the experiment’s results.
    ...
    06-02-2026, 10:05 AM
  • SEQadmin2
    Single-Cell Sequencing at an Inflection Point: Early Impacts of New Platforms and Emerging Trends
    by SEQadmin2


    With the launch of new single-cell sequencing platforms in 2026, the field stands at an exciting inflection point. This article surveys the most impactful advances in the field and discusses how they’re reshaping research in cancer, immunology, and beyond.


    Introduction

    Single-cell sequencing technologies have undergone remarkable advances over the past decade, transitioning from low-throughput experimental approaches to highly scalable platforms capable of...
    05-22-2026, 06:42 AM
  • SEQadmin2
    Environmental Genomics in the Age of NGS: From Microbes to Conservation Strategies
    by SEQadmin2

    Studying ecosystems means dealing with complex, multi-species communities that are hard to observe at scale. This complexity, however, hides many important questions to be answered, from how biogeochemical cycles work and how climate change can affect species distribution to how conservation strategies can work best.


    Genomics, particularly since the expansion of NGS, has transformed ecosystem ecology. By sequencing environmental DNA, we can now assess biodiversity without direct...
    05-06-2026, 09:04 AM

ad_right_rmr

Collapse

News

Collapse

Topics Statistics Last Post
Started by SEQadmin2, Yesterday, 08:59 AM
0 responses
14 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 06-02-2026, 12:03 PM
0 responses
22 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 06-02-2026, 11:40 AM
0 responses
19 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 05-28-2026, 11:40 AM
0 responses
32 views
0 reactions
Last Post SEQadmin2  
Working...