Question: nr database Diamond
0
gravatar for Dave Th
7 weeks ago by
Dave Th0
Dave Th0 wrote:

Hi guys,

I want to use DIAMOND for my metagenome Functional analysis. As the instruction, I have to download NCBI nr database (ftp://ftp.ncbi.nlm.nih.gov/blast/db/). Unfortunately, my internet connection is not very stable, so I have to download a multiple nr file nr.**.tar.gz instead of a nr single gz file using these code:

wget 'ftp://ftp.ncbi.nih.gov/blast/db/nr.01.tar.gz';
cat nr.**.tar.gz | tar -zxvi -f - -C

After that, I got a lot of file in my output directory (~180Gb). I wonder how I can combine all these file into a single nr.faa just like in the DIAMOND manual.

Thank you all.

Dave

protein alignment • 186 views
ADD COMMENTlink modified 7 weeks ago by h.mon24k • written 7 weeks ago by Dave Th0
1
gravatar for h.mon
7 weeks ago by
h.mon24k
Brazil
h.mon24k wrote:

DIAMOND needs its own database, it does not work with blast databases - which is what you are downloading. You have to download the NR fasta file, then:

wget ftp://ftp.ncbi.nlm.nih.gov/blast/db/FASTA/nr.gz
diamond makedb --in nr.gz -d nr
ADD COMMENTlink modified 7 weeks ago • written 7 weeks ago by h.mon24k
1

DIAMOND also needs more RAM than BLAST+. Something to keep in mind.

ADD REPLYlink written 7 weeks ago by genomax65k
Please log in to add an answer.

Help
Access

Use of this site constitutes acceptance of our User Agreement and Privacy Policy.
Powered by Biostar version 2.3.0
Traffic: 1858 users visited in the last hour