[an error occurred while processing this directive]
en

Cyanobacterial genome transposable element mining and analysis based on 454 deep-sequencing data set

Expand
  • 1. Key Laboratory of Aquatic Biodiversity and Conservation Biology, Institute of Hydrobiology, Chinese Academy of Sciences, Wuhan 430072, China 2. Graduate University of Chinese Academy of Sciences, Beijing 100049, China

Received date: 2010-09-03

  Revised date: 2010-12-20

  Online published: 2011-06-25

Abstract

Researches on the next generation sequencing (NGS) and the comparative genome analysis have recently been concerned. The analyses on transposable element composition and abundance are important parts for genome studies. Generally, the analyses of transposable element system were based on the complete spliced genomes; however, the post-processing and sequence splicing of the huge amount of short sequences from the 454 sequencer always encounter problems. Moreover, the occasion that large amount of repeat elements made up by transposable elements were incorrectly splicing or lost, leading to uncertain results. This study aimed at the construction of a framework to automatically analyze the insert sequence (IS) abundance and their composition based on a stimulated Roche 454 deep-sequencing data set, which was a 33-fold coverage of Microcystis aeruginosa NIES 843 genome. The result from the examination under the setting of three classes of division on the IS element candidates and a separated transposase examination thresholds is the most reliable. It showed that the abundance of IS element in this stimulated dataset was 10.38%, including 14 IS families and 66 IS subfamilies, which demonstrated no significant difference with the two sets of previous analysis results based on the spliced M. aeruginosa NIES 843 genome and a high percentage of IS element sequence overlap, indicating the reliability of this framework.

Cite this article

XIAO Feng, LI Ren-Hui . Cyanobacterial genome transposable element mining and analysis based on 454 deep-sequencing data set[J]. Hereditas(Beijing), 2011 , 33(6) : 654 -660 . DOI: 10.3724/SP.J.1005.2011.00654

References

[1] Rasmussen B, Fletcher IR, Brocks JJ, Kilburn MR. Reas-sessing the first appearance of eukaryotes and cyanobacteria. Nature, 2008, 455(7216): 1101-1104.
[2] Dokulil MT, Teubner K. Cyanobacterial dominance in lakes. Hydrobiologia, 2000, 438(1-3): 1-12.
[3] Lepetit D, Brehm A, Fouillet P, Biémont C. Insertion polymorphism of retrotransposable elements in popula-tions of the insular, endemic species Drosophila madeirensis. Mol Ecol, 2002, 11(3): 347-354.
[4] Nekrutenko A, Li WH. Transposable elements are found in a large number of human protein-coding genes. Trends Genet, 2001, 17(11): 619-621.
[5] Kidwell MG, Lisch DR. Perspective: Transposable elements, parasitic DNA, and genome evolution. Evolution, 2001, 55(1): 1-24.
[6] Gray WD. The nature and processing of errors in interactive behavior. Cognitive Sci, 2000, 24(2): 205-248.
[7] Zhou FF, Olman V, Xu Y. Insertion Sequences show diverse recent activities in Cyanobacteria and Archaea. BMC Genomics, 2008, 9: 36.
[8] Wicker T, Sabot F, Hua-Van A, Bennetzen JL, Capy P, Chalhoub B, Flavell A, Leroy P, Morgante M, Panaud O, Paux E, Sanmiguel P, Schulman AH. A unified classification system for eukaryotic transposable elements. Nat Rev Genet, 2007, 8(12): 973-982.
[9] Yang G, Zhang F, Hancock CN, Wessler SR. Transposition of the rice miniature inverted repeat transposable element mPing in Arabidopsis thaliana. Proc Natl Acad Sci USA, 2007, 104(26): 10962-10967.
[10] Kaneko T, Nakamura Y, Wolk CP, Kuritz T, Sasamoto S, Watanabe A, Iriguchi M, Ishikawa A, Kawashima K, Kimura T, Kishida Y, Kohara M, Matsumoto M, Matsuno A, Muraki A, Nakazaki N, Shimpo S, Sugimoto M, Takazawa M, Yamada M, Yasuda M, Tabata S. Complete genomic sequence of the filamentous nitrogen-fixing cyanobacterium Anabaena sp. strain PCC 7120. DNA Res, 2001, 8(5): 227-253.
[11] Kaneko T, Nakajima N, Okamoto S, Suzuki I, Tanabe Y, Tamaoki M, Nakamura Y, Kasai F, Watanabe A, Kawashima K, Kishida Y, Ono A, Shimizu Y, Takahashi C, Mi-nami C, Fujishiro T, Kohara M, Katoh M, Nakazaki N, Nakayama S, Yamada M, Tabata S, Watanabe MM. Com-plete genomic structure of the bloom-forming toxic cyanobacterium Microcystis aeruginosa NIES-843. DNA Res, 2007, 14(6): 247-256.
[12] Nakamura Y, Kaneko T, Sato S, Mimuro M, Miyashita H, Tsuchiya T, Sasamoto S, Watanabe A, Kawashima K, Kishida Y, Kiyokawa C, Kohara M, Matsumoto M, Matsuno A, Nakazaki N, Shimpo S, Takeuchi C, Yamada M, Tabata S. Complete genome structure of Gloeobacter violaceus PCC 7421, a Cyanobacterium that lacks thylakoids. DNA Res, 2003, 10(4): 137-145.
[13] Zhou Y, Li T, Zhao JD, Luo JC. PGAAS: a prokaryotic genome assembly assistant system. Bioinformatics, 2002, 18(5): 661-665.
[14] Rothberg JM, Leamon JH. The development and impact of 454 sequencing. Nat Biotechnol, 2008, 26(10): 1117-1124.
[15] Lin S, Haas S, Zemojtel T, Xiao P, Vingron M, Li RH. Genome-wide comparison of cyanobacterial transposable elements, potential genetic diversity indicators. Gene, 2011, 473(2): 139-149.
[16] Kurtz S. The Vmatch large scale sequence analysis software. Ref Type: Computer Program, 2010, URL http://www.vmatch.de.
[17] Larkin MA, Blackshields G, Brown NP, Chenna R, McGettigan PA, Mcwilliam H, Valentin F, Wallace IM, Wilm A, Lopez R, Thompson JD, Gibson TJ, Higgins DG. Clustal W and Clustal X version 2.0. Bioinformatics, 2007, 23(21): 2947-2948.
[18] Hall TA. BioEdit: a user-friendly biological sequence alignment editor and analysis program for Windows 95/98/NT. Nucleic Acids Symp Ser, 1999, 41: 95-98.
[19] Huang XQ, Madan A. CAP3: A DNA sequence assembly program. Genome Res, 1999, 9(9): 868-877.
[20] Altschul SF, Gish W, Miller W, Myers EW, Lipman DJ. Basic local alignment search tool.
Outlines

/