Acceptable Sp/Sn output from cufflinks and problems with Homo_sapiens.GRCh37.60.gtf

nat

Junior Member

Join Date: Jul 2009

Posts: 8
- Share
- Tweet
#1

Acceptable Sp/Sn output from cufflinks and problems with Homo_sapiens.GRCh37.60.gtf

12-02-2010, 10:58 PM

Hi there,

I was wondering what levels of Sp/Sn people are seeing from Cufflinks output, and what you would consider as acceptable.

I am working with 36bp Illumina data and processing using the TopHat/Cufflinks pipeline.

The latest dataset I have run through Cufflinks (v0.9.1) gives the following result in the stats file:

#= Summary for dataset:
# Query mRNAs : 67342 in 66642 loci (21607 multi-exon transcripts)
# (635 multi-transcript loci, ~1.0 transcripts per locus)
# Reference mRNAs : 52287 in 14604 loci (50514 multi-exon)
# Corresponding super-loci: 13723
#--------------------| Sn | Sp | fSn | fSp
Base level: 53.5 59.2 - -
Exon level: 7.7 16.1 13.4 27.8
Intron level: 18.6 87.1 18.9 88.3
Intron chain level: 0.7 1.5 1.3 2.9
Transcript level: 0.0 0.0 0.0 0.0
Locus level: 2.2 0.5 3.6 0.8
Missed exons: 110793/219160 ( 50.6%)
Wrong exons: 24548/105391 ( 23.3%)
Missed introns: 133243/179883 ( 74.1%)
Wrong introns: 2273/38400 ( 5.9%)
Missed loci: 0/14604 ( 0.0%)
Wrong loci: 18822/66642 ( 28.2%)

I have seen worse results on previous reslts, and always see 0 for the Transcript level - is this something I should focus on? Or rather just the % values for Missed and wrong exons? (I have seen other posters focus on this)

Another point is - surely I should see total number of reference loci around ~23,000 ?? (number of human protein coding genes, I have seen a greater number of loci with previous versions of Cufflinks and the exact same Ensembl reference .gtf file)

What are other people seeing for Ensembl Human reference.

Another problem that I am having, Cufflinks is producing 'u' matches (and no error messages) for every transcripts when I use the Homo_sapiens.GRCh37.60.gtf reference file, but gave good results ( a range of match types) when I used Homo_sapiens.GRCh37.55.gtf (both files have been formatted so that 'chr' is in front of the chromosome number). I can't see any obvious formatting difference between the two files, so I am a bit stumped.

Cheers
Tags: cufflinks, reference annotation, rna-seq, tophat rna-seq

Previous template Next

Pathogen Surveillance with Advanced Genomic Tools

by seqadmin

The COVID-19 pandemic highlighted the need for proactive pathogen surveillance systems. As ongoing threats like avian influenza and newly emerging infections continue to pose risks, researchers are working to improve how quickly and accurately pathogens can be identified and tracked. In a recent SEQanswers webinar, two experts discussed how next-generation sequencing (NGS) and machine learning are shaping efforts to monitor viral variation and trace the origins of infectious...
- Channel: Articles
03-24-2025, 11:48 AM
New Genomics Tools and Methods Shared at AGBT 2025

by seqadmin

This year’s Advances in Genome Biology and Technology (AGBT) General Meeting commemorated the 25^th anniversary of the event at its original venue on Marco Island, Florida. While this year’s event didn’t include high-profile musical performances, the industry announcements and cutting-edge research still drew the attention of leading scientists.

The Headliner
The biggest announcement was Roche stepping back into the sequencing platform market. In the years since...
- Channel: Articles
03-03-2025, 01:39 PM

Topics	Statistics	Last Post
New Software Simplifies 3D Gene Expression Mapping by seqadmin Started by seqadmin, Yesterday, 10:17 AM	0 responses 7 views 0 reactions	Last Post by seqadmin Yesterday, 10:17 AM
AI Tool Creates High-Resolution 3D Maps of the Mouse Brain by seqadmin Started by seqadmin, 03-20-2025, 05:03 AM	0 responses 49 views 0 reactions	Last Post by seqadmin 03-20-2025, 05:03 AM
Studying Microbial Gene Transfer with RNA Barcoding by seqadmin Started by seqadmin, 03-19-2025, 07:27 AM	0 responses 59 views 0 reactions	Last Post by seqadmin 03-19-2025, 07:27 AM
Mapping the snoRNAome in Zebrafish to Advance Disease Research by seqadmin Started by seqadmin, 03-18-2025, 12:50 PM	0 responses 50 views 0 reactions	Last Post by seqadmin 03-18-2025, 12:50 PM

Seqanswers Leaderboard Ad

Acceptable Sp/Sn output from cufflinks and problems with Homo_sapiens.GRCh37.60.gtf

Latest Articles

ad_right_rmr

News