Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • BWA - input files

    Hi,

    I am trying to align paired end reads (Illumina) to a reference genome using BWA. I have 10 reads files, 5 for each direction.

    In this old post 'totalnew' says to align the files separately:
    Code:
    bwa aln database.fasta 4_1.fq > 1_1.fq.sai
    bwa aln database.fasta 4_2.fq > 1_2.fq.sai
    bwa aln database.fasta 5_1.fq > 2_1.fq.sai
    bwa aln database.fasta 5_2.fq > 2_2.fq.sai
    [...]
    I have been reading for quite a while now so sorry if I missed something totally obvious but how do I run sampe from there on? Can I just say
    Code:
    bwa sampe database.fa 1_1.fq.sai 2_1.fq.sai 3_1... 1_2.fq.sai 2_2.fq.sai 3_2... 1_1.fq 2_1.fq 3_1... 1_2.fq 2_2.fq 3_2... > alignment.sam
    or do I need to run sampe for each file separately?
    What if I concatenate all reads files into s_1_sequence.txt and s_2_sequences.txt and then run bwa aln twice and bwa samse once?

    cheers!

    ps (offtopic) @lh3: I tried downloading bwa 0.5.8a but it seems as though there are files missing. Here what I get:
    Code:
    wget http://sourceforge.net/projects/bio-bwa/files/bwa-0.5.8a.tar.bz2/download
    bunzip2 bwa-0.5.8a.tar.bz2
    tar -xf bwa-0.5.8a.tar
    ls bwa-0.5.8a
    bntseq.c  bwase.c     bwtaln.c  bwtgap.h    bwtio.c     bwtsw2_aux.c    bwtsw2_main.c  is.c     kstring.c  main.h        simple_dp.c     utils.c
    bntseq.h  bwase.h     bwtaln.h  bwt_gen     bwt_lite.c  bwtsw2_chain.c  ChangeLog      khash.h  kstring.h  Makefile      solid2fastq.pl  utils.h
    bwa.1     bwaseqio.c  bwt.c     bwt.h       bwt_lite.h  bwtsw2_core.c   COPYING        kseq.h   kvec.h     NEWS          stdaln.c
    bwape.c   bwa.txt     bwtgap.c  bwtindex.c  bwtmisc.c   bwtsw2.h        cs2nt.c        ksort.h  main.c     qualfa2fq.pl  stdaln.h
    I downloaded 0.5.7 and it runs fine.

  • #2
    I think you'd better to run sampe for each PE files separately:

    bwa sampe [options] <prefix> <in1.sai> <in2.sai> <in1.fq> <in2.fq>

    Although you concatenate all read files into two PE seq is okay, I recommend you split them (for example: 100M reads per files) and run in different CPU cores or computer nodes (MPI mode), so that it would be more faster.

    Comment


    • #3
      Hi,

      Thanks for your reply.

      So you suggest running sampe at least five times and then merge the resulting sam files?

      I will try this.

      Cheers!

      *** edit
      Thanks, works fine.
      Last edited by Bruins; 07-12-2010, 01:47 AM.

      Comment

      Latest Articles

      Collapse

      • seqadmin
        Recent Innovations in Spatial Biology
        by seqadmin


        Spatial biology is an exciting field that encompasses a wide range of techniques and technologies aimed at mapping the organization and interactions of various biomolecules in their native environments. As this area of research progresses, new tools and methodologies are being introduced, accompanied by efforts to establish benchmarking standards and drive technological innovation.

        3D Genomics
        While spatial biology often involves studying proteins and RNAs in their...
        01-01-2025, 07:30 PM
      • seqadmin
        Advancing Precision Medicine for Rare Diseases in Children
        by seqadmin




        Many organizations study rare diseases, but few have a mission as impactful as Rady Children’s Institute for Genomic Medicine (RCIGM). “We are all about changing outcomes for children,” explained Dr. Stephen Kingsmore, President and CEO of the group. The institute’s initial goal was to provide rapid diagnoses for critically ill children and shorten their diagnostic odyssey, a term used to describe the long and arduous process it takes patients to obtain an accurate...
        12-16-2024, 07:57 AM

      ad_right_rmr

      Collapse

      News

      Collapse

      Topics Statistics Last Post
      Started by seqadmin, 01-09-2025, 04:04 PM
      0 responses
      11 views
      0 likes
      Last Post seqadmin  
      Started by seqadmin, 01-09-2025, 09:42 AM
      0 responses
      18 views
      0 likes
      Last Post seqadmin  
      Started by seqadmin, 01-08-2025, 03:17 PM
      0 responses
      28 views
      0 likes
      Last Post seqadmin  
      Started by seqadmin, 01-03-2025, 11:18 AM
      1 response
      45 views
      1 like
      Last Post Tonia
      by Tonia
       
      Working...
      X