parva strain HT, DSM 21527,

parva strain HT, DSM 21527, download catalog was grown in semisolid DSMZ medium 1113 (Leptospira medium) [43] at 30��C. DNA was isolated from 1-1.5 g of cell paste using MasterPure Gram-positive DNA purification kit (Epicentre MGP04100) following the standard protocol as recommended by the manufacturer with modification st/DL for cell lysis as described in Wu et al. 2009 [41]. DNA is available through the DNA Bank Network [44]. Genome sequencing and assembly The genome was sequenced using a combination of Illumina and 454 sequencing platforms. All general aspects of library construction and sequencing can be found at the JGI website [45]. Pyrosequencing reads were assembled using the Newbler assembler (Roche).

The initial Newbler assembly consisting of 217 contigs in 1 scaffold was converted into a phrap [46] assembly by making fake reads from the consensus, to collect the read pairs in the 454 paired end library. Illumina GAii sequencing data (8,018.4 Mb) was assembled with Velvet [47] and the consensus sequences were shredded into 1.5 kb overlapped fake reads (shreds) and assembled together with the 454 data. The 454 draft assembly was based on 200.6 Mb 454 draft data and all of the 454 paired end data. Newbler parameters are -consed -a 50 -l 350 -g -m -ml 21. The Phred/Phrap/Consed software package [46] was used for sequence assembly and quality assessment in the subsequent finishing process. After the shotgun stage, reads were assembled with parallel phrap (High Performance Software, LLC). Possible mis-assemblies were corrected with gapResolution [45], Dupfinisher [48], or sequencing cloned bridging PCR fragments with subcloning.

Gaps between contigs were closed by editing in Consed, by PCR and by Bubble PCR primer walks (J.-F. Chang, unpublished). A total of 361 additional reactions and 11 shatter library were necessary to close some gaps and to raise the quality of the final contigs. Illumina reads were also used to correct potential base errors and increase consensus quality using a software Polisher developed at JGI [49]. The error rate of the final genome sequence is less than 1 in 100,000. Together, the combination of the Illumina and 454 sequencing platforms provided 1,722.1 �� coverage of the genome. The final assembly contained 348,698 pyrosequence and 97,925,368 Illumina reads.

Genome annotation Genes were identified using Prodigal [50] as part of the DOE-JGI annotation pipeline [51], followed by a round of manual curation using the JGI GenePRIMP pipeline [52]. The predicted Anacetrapib CDSs were translated and used to search the National Center for Biotechnology Information (NCBI) nonredundant database, UniProt, TIGR-Fam, Pfam, PRIAM, KEGG, COG, and InterPro databases. Additional gene prediction analysis and functional annotation was performed within the Integrated Microbial Genomes – Expert Review (IMG-ER) platform [53]. Genome properties The genome statistics are provided in Table 3 and Figure 3.

Leave a Reply

Your email address will not be published. Required fields are marked *

*

You may use these HTML tags and attributes: <a href="" title=""> <abbr title=""> <acronym title=""> <b> <blockquote cite=""> <cite> <code> <del datetime=""> <em> <i> <q cite=""> <strike> <strong>