Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts

  • lcollado
    replied
    Hello,

    I know that this is an old thread, but I'm curious to know how Swift compares vs the most recent SolexaPipeline versions.

    Thank you!
    Leonardo

    Leave a comment:


  • bioinfosm
    replied
    Are there any updates on SWIFT? data sizes, number of files generated, comparison with Illumina pipeline results..

    Leave a comment:


  • new300
    replied
    Originally posted by lparsons View Post
    I'm quite interested in using open-source software for scientific work. We have recently acquired an Illumina GAII machine, and are trying to come up with data management solutions. Right now we are planning to throw away the images after the primary analysis (base-calling) is completed. We are saving the intensity and noise files, but not the images, which seems to be fairly common. However, it seems that this software requires the original images, which makes sense, but would limit our ability to use it on past experiments.
    Are you using iPar to process the images and then mirroring off the intensity files? Swift will process from intensity files (as produced by the UNIX pipeline). I've heard the iPar intensity format is different from that used by the UNIX pipeline if someone wants to send me a sample file I'll write a parser for it.

    Originally posted by lparsons View Post
    Would it be feasible to use swift on the Firecrest output (intensity and noise)?
    Yes it's feasible, I would hope the results would be comparable with the Illumina pipeline.

    Originally posted by lparsons View Post

    Do many labs actually save the image files?

    It seems like an ideal initial setup would be to process the images with both the Illumina pipeline and Swift. Has anyone yet set this up?
    As mentioned Sanger save the images while they do QC, the images are mirrored off as the run progresses and processed using the UNIX pipeline on a separate cluster.

    If you're interested in trying out Swift drop me an email at new at sgenomics dot org. It's in ``active development'' at the moment and I'm happy to work with people on any issues that come up.

    Leave a comment:


  • clivey
    replied
    sanger have it set up - talk to Tom Skelley.

    Images are still very diagnostic of any issue with your sample or sequencer (or run). Looking at images allowed sanger to optimise their pipeline. For example, when your flowcell quality goes down, or an operator gets oil on the flowcell etc., or your focusing is off and you suddenly get lots of strange new 'contaminants' in your output file as a result, or your base qualities all drop halfway through your project, youe data goes bad and you look and your clusters look wierd coz of an issue with your cluster station, or theres stuff growing in your reagents appearing as blobs on the images (but not visible to the naked eye), or your flowcell surface isnt there etc etc. You should keep them for QC - then throw them. Generally (but not in all cases) higher throughput labs with big projects indulge in some image retention for some period.

    Leave a comment:


  • lparsons
    replied
    I'm quite interested in using open-source software for scientific work. We have recently acquired an Illumina GAII machine, and are trying to come up with data management solutions. Right now we are planning to throw away the images after the primary analysis (base-calling) is completed. We are saving the intensity and noise files, but not the images, which seems to be fairly common. However, it seems that this software requires the original images, which makes sense, but would limit our ability to use it on past experiments.

    Would it be feasible to use swift on the Firecrest output (intensity and noise)?

    Do many labs actually save the image files?

    It seems like an ideal initial setup would be to process the images with both the Illumina pipeline and Swift. Has anyone yet set this up?

    Leave a comment:


  • new300
    replied
    Originally posted by timread View Post
    BTW - the link: http://swiftng.sourceforge.net appears to be broken.

    The connection seems to be a problem only from my desktop at work (which is behind a US government firewall). From other locations i can get through OK.
    Odd, you can try: http://sgenomics.org/swift/ which should also work.

    Leave a comment:


  • new300
    replied
    Originally posted by iris42 View Post
    Is it normal to see different output when running the same binary version of swift on the same computer for multiple times and running it on different computers? I observed both. It looks like most of the differences in the fastq output is the quality scores.
    Running on different computers it's quite likely that the output will vary slightly as they are likely to have different floating point implementations.

    On the same computer is a little odd, how different are the results? If it's a small difference then this could be down to the FFTW implementation we are using which sometimes employs a non-deterministic algorithm.

    Leave a comment:


  • iris42
    replied
    Is it normal to see different output when running the same binary version of swift on the same computer for multiple times and running it on different computers? I observed both. It looks like most of the differences in the fastq output is the quality scores.

    Leave a comment:


  • cgb
    replied
    works for me

    Leave a comment:


  • timread
    replied
    BTW - the link: http://swiftng.sourceforge.net appears to be broken.

    The connection seems to be a problem only from my desktop at work (which is behind a US government firewall). From other locations i can get through OK.
    Last edited by timread; 11-18-2008, 12:44 PM. Reason: clarification of connection problem

    Leave a comment:


  • new300
    replied
    In terms of memory usage we're trying to stay within a 2Gb limit. A 37Gb paired end peaks at around 1Gb.

    Leave a comment:


  • new300
    replied
    Originally posted by dvh View Post
    Could you maybe share some stats as to how Swift performs vs the current version of Bustard?
    E.g. amount of data/reads mapped, error rate for the same lane analysed both ways.
    thanks
    david
    I'm still in the process of validating it on non-phiX data. For the phiX data I've looked at, against the 1.0 pipeline I've seen 20% more PF reads at a similar error rate.

    In terms of runtime, a GA1 single end takes around 10mins end to end. GA2 37 cycles paired end takes around an hour end to end.

    Leave a comment:


  • dvh
    replied
    Could you maybe share some stats as to how Swift performs vs the current version of Bustard?
    E.g. amount of data/reads mapped, error rate for the same lane analysed both ways.
    thanks
    david

    Leave a comment:


  • new300
    replied
    Originally posted by cgb View Post
    i wonder if if can be put onto a boot DVD and run on the iPar computers - data mirrored in real time using the sanger mirroring scripts ?
    Yes, this absolutely should be possible and is something we'd like to look in to. Users interested in doing this are encouraged to make contract.

    Leave a comment:


  • cgb
    replied
    cool

    i wonder if if can be put onto a boot DVD and run on the iPar computers - data mirrored in real time using the sanger mirroring scripts ?

    Leave a comment:

Latest Articles

Collapse

  • SEQadmin2
    Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
    by SEQadmin2



    CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

    Despite this, “CRISPR helped turn genome editing from a specialized technique into
    ...
    07-31-2026, 11:01 AM
  • SEQadmin2
    Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
    by SEQadmin2


    Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

    The systematic characterization of the human proteome has
    ...
    07-20-2026, 11:48 AM
  • SEQadmin2
    Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
    by SEQadmin2



    Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
    ...
    07-09-2026, 11:10 AM

ad_right_rmr

Collapse

News

Collapse

Topics Statistics Last Post
Started by SEQadmin2, 07-31-2026, 02:55 AM
0 responses
18 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-24-2026, 12:17 PM
0 responses
16 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-23-2026, 11:41 AM
0 responses
15 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-20-2026, 11:10 AM
0 responses
26 views
0 reactions
Last Post SEQadmin2  
Working...