Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • How to combine all the seqences in a fasta file to one sequence?

    Hi every one,

    I have a fasta file like this like this:
    >GL132456
    ABCDEDERFSE
    >GL123456
    ABDFDRAGDTGEGAGFDAS
    >GL125436
    AFGHSRSGFGSHSFG
    I want to combine all the sequences to one sequence, like this:
    ABCDEDERFSEABDFDRAGDTGEGAGFDASAFGHSRSGFGSHSFG

    Thanks a lot!

    Richard

  • #2
    Maybe:

    Code:
    grep -v ">" your.fasta | sed -e 's/\n//g' -e 's/ //g' > youredited.fasta
    Select all lines than don't have ">", remove newlines, remove spaces, direct output to a new file.
    Last edited by rhinoceros; 10-26-2013, 06:03 PM.
    savetherhino.org

    Comment


    • #3
      Originally posted by rhinoceros View Post
      Maybe:

      Code:
      grep -v ">" your.fasta | sed -e 's/\n//g' -e 's/ //g' > youredited.fasta
      Select all lines than don't begin with ">", remove newlines, remove spaces, direct output to a new file.
      Just a minor update of above would create the required output:

      Code:
      grep -v ">" your.fasta|tr '\n' ' ' | sed -e 's/ //g' > youredited.fasta
      Output:

      ABCDEDERFSEABDFDRAGDTGEGAGFDASAFGHSRSGFGSHSFG

      Comment


      • #4
        use cat command

        Comment


        • #5
          I have checked all the commands. No one worked for my data.
          Thanks!

          Comment


          • #6
            Originally posted by wmseq View Post
            I have checked all the commands. No one worked for my data.
            Thanks!
            You're doing it wrong then because they do work:

            Code:
            $ cat file.fasta 
            >GL132456
            ABCDEDERFSE
            >GL123456
            ABDFDRAGDTGEGAGFDAS
            >GL125436
            AFGHSRSGFGSHSFG
            
            $ grep -v ">" file.fasta | tr '\n' ' ' | sed -e 's/ //g' > youredited.fasta
            
            $ cat youredited.fasta 
            ABCDEDERFSEABDFDRAGDTGEGAGFDASAFGHSRSGFGSHSFG
            savetherhino.org

            Comment


            • #7
              I am sorry for my mistake, because my sequence is as follows.
              >GL132456
              ABCDEDERFSE
              DGA
              >GL123456
              ABDFDRAGDTGEGAGFDAS
              >GL125436
              AFGHSRSGFGSHSFG

              After I used the commands, I got the following sequence with white space in it.
              ABCDEDERFSE
              DGA
              ABDFDRAGDTGEGAGFDASAFGHSRSGFGSHSFG

              thank you for your patience!
              Last edited by wmseq; 10-28-2013, 07:05 AM.

              Comment


              • #8
                perl -pe 'chomp;s/>.+//' input.fasta >out.txt

                Comment

                Latest Articles

                Collapse

                • seqadmin
                  Latest Developments in Precision Medicine
                  by seqadmin



                  Technological advances have led to drastic improvements in the field of precision medicine, enabling more personalized approaches to treatment. This article explores four leading groups that are overcoming many of the challenges of genomic profiling and precision medicine through their innovative platforms and technologies.

                  Somatic Genomics
                  “We have such a tremendous amount of genetic diversity that exists within each of us, and not just between us as individuals,”...
                  05-24-2024, 01:16 PM
                • seqadmin
                  Recent Advances in Sequencing Analysis Tools
                  by seqadmin


                  The sequencing world is rapidly changing due to declining costs, enhanced accuracies, and the advent of newer, cutting-edge instruments. Equally important to these developments are improvements in sequencing analysis, a process that converts vast amounts of raw data into a comprehensible and meaningful form. This complex task requires expertise and the right analysis tools. In this article, we highlight the progress and innovation in sequencing analysis by reviewing several of the...
                  05-06-2024, 07:48 AM

                ad_right_rmr

                Collapse

                News

                Collapse

                Topics Statistics Last Post
                Started by seqadmin, Today, 06:55 AM
                0 responses
                8 views
                0 likes
                Last Post seqadmin  
                Started by seqadmin, 05-30-2024, 03:16 PM
                0 responses
                23 views
                0 likes
                Last Post seqadmin  
                Started by seqadmin, 05-29-2024, 01:32 PM
                0 responses
                27 views
                0 likes
                Last Post seqadmin  
                Started by seqadmin, 05-24-2024, 07:15 AM
                0 responses
                214 views
                0 likes
                Last Post seqadmin  
                Working...
                X