Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • Santi Català
    Junior Member
    • May 2012
    • 5

    #1

    Metagenomics analysis with CLOTU: problems with input data

    Hi to all,

    I'm doing a metagenomic approach of environmental DNA with an amplicon library constructed with the ITS1 of one genera of fungi, with a total of four different libraries (4 MID's). I'm trying to analyze the data with CLOTU but I obtain the next log file:

    Make-Cluster: -i [accepted.fas] -p b -s [n] -c 50 -m 97 -l y -a 8
    make-FAS_FILE-EACH-Cluser: y accepted.fas blastclust_out.txt y
    Make-Matrix: -f METADATA.txt -t TPA.txt -g y -o [matrix]

    Manage file step completed

    Illegal division by zero at /usit/titan/u1/globus/CLOTU/VER-1.1/filter.pl line 444, <DATA> line 444216.
    Cannot open file "accepted.fas".
    Cannot open accepted.fas to READ
    Cannot open file "blastclust_out.txt".


    Seqs per file: 200
    Couldn't open 'cluster_out.fas': No such file or directory at /usit/titan/u1/globus/CLOTU/VER-1.1/clotu.pl line 225.

    Who know what is the problem?

    Another question is related to methodology. My library is very closed to the genera that we are studying, and we think that we don't have more than 20 species per sample (library). What is the best methodology that we can use? Assembly with Mira or Newbler is a good option?

    Thanks in advance

    Santi Català
    Valencia (Spain)
  • Santi Català
    Junior Member
    • May 2012
    • 5

    #2
    I forgot that my amplicons (300 bp) were generated on a Junior GS pyrosequencer.

    Comment

    • miguelangel
      Member
      • Jun 2012
      • 16

      #3
      hello!

      I'm having similar problems, how do you finally resolve them?

      By the way, what about the assembly for fungi?

      I'm from Madrid, so if you want we can speak in Spanish

      Comment

      • Santi Català
        Junior Member
        • May 2012
        • 5

        #4
        Hola Miguel Ángel,

        Mi principal problema era de que disponía de un PC arcaico por lo que me veía obligado a hacer uso de una plataforma de este tipo. Hace unos meses conseguí hacer uso de un equipo de un grupo de unos colegas de otra Universidad, con lo que el Pipeline nos lo hicimos nosotros. Tras extraer los datos con el sff_extract hicimos el clustering con el BlastClust, y con la ayuda de un script en Python exporté a Fasta el Listfile, para poder trabajar con las secuencias de cada OTU. Me vi con el problema de una cantidad de errores, aparte de homopolímeros, increible, probablemente debida al uso de una polimerasa convencional tras preparar la librería por Nested-PCR, con lo que en lugar de coger a random sequence de cada cluster, hice un alineamiento para exportar una consenso, que fue la que utilicé para blastear. Éste último paso no tendría sentido si tuviera un número excesivo de OTU's, pero con la alta especificidad de los cebadores me quité muchos organismos de enmedio.

        Probé un ensamblaje con el MIRA, pero creo que no tenía mucho sentido hacerlo de esta manera, a pesar de que a nivel de diversidad de especies obtuve similares resultados.

        Si te interesa la metodología te la puedo pasar vía e-mail más detallada.

        ¿Qué tal tus secuencias? ¿Qué bichos buscas?

        Saludos cordiales

        Comment

        • miguelangel
          Member
          • Jun 2012
          • 16

          #5
          Muchas gracias por la info!

          Nosotros estamos realizando los análisis de secuencias de bacterias, hongos y algas, obtenidas desde un 454 Titanium de Roche.

          Por ahora las secuencias parece que están bastante bien, lo que nos está suponiendo algo más de problema es su cribado y, sobre todo, su asignación taxonómica.
          Hemos usado por ahora unos scripts de BioPython y el SOP de mothur de Schloss et al, pero a la hora de asignar cada secuencia a cada bicho nos hemos encontrado con problemas: para Hongos no hay una buena base de datos pre-alineada con la que comparar, para algas ídem y para bacterias, la que hay (SILVA) casi no tiene cianobacterias (y las que tiene las confunde con DNA de cloroplastos), que son las que más esperamos tener.

          Por ahora estamos haciendo unos pasos con mothur y queremos probar que tal la asignación con CLOTU, que es bastante sencillito una vez he conseguido enterarme de cómo tienen que ir los archivos iniciales.

          Muchas Gracias

          Saludos!

          Comment

          • Santi Català
            Junior Member
            • May 2012
            • 5

            #6
            Bueno, pues igual para la siguiente placa te pediré consejo para meter los datos en el Clotu :-). Aunque con la poca diversidad que tengo me apaño bastante bien haciéndolo a mano.

            Por mucha casualidad, tus muestras no serán del Machupichu?

            Saludos!

            Comment

            • miguelangel
              Member
              • Jun 2012
              • 16

              #7
              Encantado de poder ayudarte.

              Mis muestras no son de Machu Pichu, pero... ¿por qué lo preguntas?

              Comment

              • Santi Català
                Junior Member
                • May 2012
                • 5

                #8
                Nada, era mucha casualidad jeje. Hace unos meses estuve hablando con un colega de Madrid y me comentó exactamente la misma metodología que has utilizado tú para la asignación taxonómica, y él andaba trabajando con esas muestras, con todo tipo de bichos como tú jeje.

                Comment

                Latest Articles

                Collapse

                • SEQadmin2
                  Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
                  by SEQadmin2



                  CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

                  Despite this, “CRISPR helped turn genome editing from a specialized technique into
                  ...
                  07-31-2026, 11:01 AM
                • SEQadmin2
                  Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
                  by SEQadmin2


                  Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

                  The systematic characterization of the human proteome has
                  ...
                  07-20-2026, 11:48 AM
                • SEQadmin2
                  Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
                  by SEQadmin2



                  Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
                  ...
                  07-09-2026, 11:10 AM

                ad_right_rmr

                Collapse

                News

                Collapse

                Topics Statistics Last Post
                Started by SEQadmin2, 08-03-2026, 10:13 AM
                0 responses
                16 views
                0 reactions
                Last Post SEQadmin2  
                Started by SEQadmin2, 07-31-2026, 02:55 AM
                0 responses
                32 views
                0 reactions
                Last Post SEQadmin2  
                Started by SEQadmin2, 07-24-2026, 12:17 PM
                0 responses
                23 views
                0 reactions
                Last Post SEQadmin2  
                Started by SEQadmin2, 07-23-2026, 11:41 AM
                0 responses
                21 views
                0 reactions
                Last Post SEQadmin2  
                Working...