Originally posted by TEFA
View Post
Unconfigured Ad
Collapse
X
-
Was this by any chance a TopHat output? Somebody alerted me recently to a bug that causes mishandling of SAM lines that contain the "NH" flag, and that seems to appear in newer tophat output. I'll try to fix this ASAP.Originally posted by EGrassi View PostWoa, I used samtools -n and htseq-count gave me all counts of zero...that's strange, I will check with IGV or similar tools but cuffdiff gave me fpkg values with the same gtf and bam file...
Simon
Comment
-
Tophat 2.0 output, yes!Originally posted by Simon Anders View PostWas this by any chance a TopHat output? Somebody alerted me recently to a bug that causes mishandling of SAM lines that contain the "NH" flag, and that seems to appear in newer tophat output. I'll try to fix this ASAP.
Thank you,
E.
Comment
-
Ok, on the whole data it falls back to all 0:
samtools view S1.sorted.bam.bam | sed 's/NM:i:\d//g' | htseq-count --stranded=no - /rogue/bioinfotree/prj/ewing-rnaseq/local/share/data/Homo_sapiens/UCSC/hg19/Annotation/Genes/genes.gtf > counts_S1_htseqq
I'm really a little surprised by the amount of errors and other glitches in rna seq software (this is a generic rant, not focused on htseq!)...it's a really complicated and new area, that's true, but right now using any of the existing tool seems a bet and to tell the truth I never feel confident on the results, as long as with every version results changes, and when something runs without errors I never know if that means that it worked or that an exception just had not screamed enough :/
Comment
-
The bug with the NH flags, which was introduced during some refactoring in the version of a week ago or so is now fixed in version 0.5.3p8. Sorry that the recent changes turned out to be so volatile.
Comment
-
Hello, I have performed the mapping of miRNA reads to the mirbase precursor file by SHRiMP. But I'm facing a probem in HTSeq while taking the read counts from the output.sam file.Originally posted by Simon Anders View PostSorry, it seems we made some mix-up between version 0.5.3p4 and 0.5.3p5. Essentially, p5 undid some fixes in p4, including the one for "*" qualities. Now, there is version 0.5.3p6, which should clean up this mess. Please let me know if you still have problems.
The error which I'm trying to solve is;
Error occured when processing SAM input (line 2305 of file HBI_10.sam): ("'seq' and 'qualstr' do not have the same length.", 'line 2305 of file HBI_10.sam') [Exception type: ValueError, raised in _HTSeq.pyx:808]
Has anyone of you come across this error? What might be the solution?
Comment
-
Actually, I did see that SAM file. The problem is that my sequence string & its respective sequence quality string is not of equal length.Originally posted by maubp View PostSushant - what does line 2305 of your SAM file look like? Unless the quality entry is the special value '*' only then it is a likely an unrelated issue.
However, I converted my fastq files to fasta and then performed the alignment, which solved this problem.
But the mapping percentage I'm getting with SHRiMP is too low(2.77%), for the mirna aligning to mature.
What parameters should I consider?
Comment
Latest Articles
Collapse
-
by SEQadmin2
CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).
Despite this, “CRISPR helped turn genome editing from a specialized technique into...-
Channel: Articles
07-31-2026, 11:01 AM -
-
by SEQadmin2
Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.
The systematic characterization of the human proteome has...-
Channel: Articles
07-20-2026, 11:48 AM -
-
by SEQadmin2
Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
...-
Channel: Articles
07-09-2026, 11:10 AM -
ad_right_rmr
Collapse
News
Collapse
| Topics | Statistics | Last Post | ||
|---|---|---|---|---|
|
Started by SEQadmin2, 07-31-2026, 02:55 AM
|
0 responses
15 views
0 reactions
|
Last Post
by SEQadmin2
07-31-2026, 02:55 AM
|
||
|
Started by SEQadmin2, 07-24-2026, 12:17 PM
|
0 responses
15 views
0 reactions
|
Last Post
by SEQadmin2
07-24-2026, 12:17 PM
|
||
|
Started by SEQadmin2, 07-23-2026, 11:41 AM
|
0 responses
13 views
0 reactions
|
Last Post
by SEQadmin2
07-23-2026, 11:41 AM
|
||
|
Started by SEQadmin2, 07-20-2026, 11:10 AM
|
0 responses
24 views
0 reactions
|
Last Post
by SEQadmin2
07-20-2026, 11:10 AM
|
Comment