Originally posted by Markiyan
View Post
Unconfigured Ad
Collapse
X
-
Hey, I did try metaSPAdes, less than 1% of total reads assembled. A lot of people tried alternative methods, joined paired-ends and get long reads, but don't assembled reads. Then, use the long merged reads to do BLAST or other annotations.Originally posted by fanli View PostOut of curiosity, why are you joining the read pairs? A lot of the metagenomics software out there now supports paired end reads as input. The metaSPAdes assembler @GenoMax mentioned requires paired end data IIRC.
Comment
-
-
Would something like kraken or CLARK not be helpful? Are you trying to assemble and annotate de novo genomes? Or trying to figure out the microbial composition and functional content? I guess my point is you would discard ~40% of your data in the joining process, which may not be necessary depending on your task of interest.
Comment
-
-
I am not interested in a specific genome in the soil community. Basically, I am just interested in the microbial composition and functional content. There is another way that people usually do. They don't assemble or merge pairs and they just blast using raw data. However, I think blast using ~150bp reads is worse. Some publication shows the intermediate length (merged pair) is better than assembled longer reads or unasembled short reads for the question that I am asking for.Originally posted by fanli View PostWould something like kraken or CLARK not be helpful? Are you trying to assemble and annotate de novo genomes? Or trying to figure out the microbial composition and functional content? I guess my point is you would discard ~40% of your data in the joining process, which may not be necessary depending on your task of interest.
Comment
-
-
You might find this benchmark to be helpful:
My understanding is that you are better off using newer k-mer based approaches as opposed to BLAST. I've had reasonable success with kraken, although the memory requirements are somewhat onerous. You also have to be extremely careful about removing contaminant (aka human) sequences as these tend to get misclassified.
Another option is kallisto (https://github.com/pachterlab/metakallisto) but I have yet to be able to even build a database due to memory constraints.
Comment
-
-
I will try. Can you gimme the Kraken link? Thanks.Originally posted by fanli View PostYou might find this benchmark to be helpful:
My understanding is that you are better off using newer k-mer based approaches as opposed to BLAST. I've had reasonable success with kraken, although the memory requirements are somewhat onerous. You also have to be extremely careful about removing contaminant (aka human) sequences as these tend to get misclassified.
Another option is kallisto (https://github.com/pachterlab/metakallisto) but I have yet to be able to even build a database due to memory constraints.
Comment
-
Latest Articles
Collapse
-
by SEQadmin2
Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.
The systematic characterization of the human proteome has...-
Channel: Articles
07-20-2026, 11:48 AM -
-
by SEQadmin2
Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
...-
Channel: Articles
07-09-2026, 11:10 AM -
-
by SEQadmin2
Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.
There is no single reason why many patients don’t respond to treatment as expected. Cancer is...-
Channel: Articles
07-08-2026, 05:17 AM -
ad_right_rmr
Collapse
News
Collapse
| Topics | Statistics | Last Post | ||
|---|---|---|---|---|
|
Started by SEQadmin2, Today, 11:41 AM
|
0 responses
9 views
0 reactions
|
Last Post
by SEQadmin2
Today, 11:41 AM
|
||
|
Started by SEQadmin2, 07-20-2026, 11:10 AM
|
0 responses
21 views
0 reactions
|
Last Post
by SEQadmin2
07-20-2026, 11:10 AM
|
||
|
Started by SEQadmin2, 07-13-2026, 10:26 AM
|
0 responses
35 views
0 reactions
|
Last Post
by SEQadmin2
07-13-2026, 10:26 AM
|
||
|
Started by SEQadmin2, 07-09-2026, 10:04 AM
|
0 responses
44 views
0 reactions
|
Last Post
by SEQadmin2
07-09-2026, 10:04 AM
|
Comment