DNA Data Bank of Japan DNA Database Release 47, Oct. 2001, including 13,266,610 entries, 14,145,671,645 bases This database may be copied and redistributed without permission on the condition that all the statements in this release note are reproduced in each copy. The present release contains the newest data prepared by the DNA Data Bank of Japan (DDBJ), GenBank, and European Molecular Biology Laboratory/European Bioinformatics Institute (EMBL/EBI) as of Sep. 27, 2001. This unified database was made possible thanks to the international collaboration among the three data banks. All the entries have accordingly been annotated with the feature keys common to them. All the entries designated by the accession numbers with the prefixes "C", "D", "E", "AB", "AG", "AK", "AP", "AT", "AU", "AV", "BA" and "BB" have been collected and processed by DDBJ, and the rest have been prepared by GenBank and EMBL/EBI. There have been a number of genome projects going on worldwide. Among them human genome projects have probably been most productive and yielded a large number of ordinary sequences, huge amounts of ESTs and quantities of genome sequences. Thus, we have the human(HUM) division solely for human sequences and the primate (PRI) division for non-human primate sequences. Note that the EST division also contains human sequences. The present release does not have the ORG division. Thus, if you are interested in human mitochondrial sequences, for example, you are now advised to refer to the HUM division. The HUM division in this release was recorded in 11 files each of which had 300 MB storage capacity. Incidentally, the BCT, INV and PLN divisions were recorded in 3, 3, 4 files, respectively. This release also includes an independent division (PAT) for patent data. The patent data are those which the Japanese Patent Office (JPO), United States Patent and Trademark Office (USPTO), and the European Patent Office (EPO) collected and processed. The accession numbers of the patent data collected by the Japanese Patent Office start with the prefix "E", those collected and supplied by USPTO and GenBank respectively start with "I" and "AR", and those collected and supplied by EPO and EMBL/EBI respectively start with "A" and "AX". The entries with the prefixes "I","AR", "A","AX" and "E" were allocated to a file (ddbjpat.seq) in the DDBJ format. Note also that unauthorized use of the patent data may cause legal issues for which we take no responsibility. In the present release, the SOURCE in the flat file was revisited and revised if necessary in accordance with the unified taxonomy database common to the three data banks. The number of ESTs has been increasing at an enormous rate and is expected to be growing even more rapidly in the future. Therefore, EST data were stored in 97 files each of which had the same storage capacity as the file of the HUM division. The present release includes the GSS division. GSS stands for the Genome Survey Sequence, which is similar to EST, except that GSS is genomic DNA whereas EST is cDNA. This division was recorded in 29 files similarly to the HUM division. This release also includes the High Throughput Genome Sequence (HTGS) which comes mainly from genome project teams which deal with a clone as a sequencing unit. HTGS in this release were recorded in 21 files similarly to the HUM division. The index files are not presented in this release except for ddbjacc.idx, ddbjgen.idx, ddbjjou.idx, and ddbjkey.idx. Instead, we have included a program by which to make the index files not presented in this release. For the use of the program, see the files, seq2indexes.doc, seq2indexes.c, and seq2indexes.h in this release. The present release contains amino acid sequences that were translated from the corresponding nucleotide sequences in our database. In the translation we paid much attention to the fact that some species or organella have a codon different from the universal one, and used the proper codon table. If you find an incorrect codon in a translated sequence, please let us know. The three data banks include the item VERSION in the flat file, which indicates a version of a submitted nucleotide sequence (see Table 1). It is expressed as AB123456.1, in which the digit(s) after the period is a version number. The reason for adding VERSION is that since a released sequence sometimes revised by the submitter, the accession number alone cannot specify the sequence in question causing the user a trouble. The number is increased by one every time when a revised sequence is made public. Accordingly, the translated protein sequence will be accompanied with a /protein_id which is expressed as BAA12345.1, in which the digit(s) after the period is again a version number. The number is increased by one when the corresponding nucleotide sequence is revised and the protein sequence is changed as a result, and when the revised protein sequence is made public. We terminated the RNA division. The RNA data were redistributed according to the category of the organism. Therefore, you will find a human RNA sequence, for example, in the HUM division. The present release includes a division, CON. The CON division is to show the order of related sequences in a genome, and expressed by join and the accession numbers of the sequences. The contents of the CON division are compiled by the three data banks not by the data submitter. The current number of the entries of this division is 9,324. The present release also includes, HTC (High Throughput cDNA). The definition of the HTC division is as follows. This division is to include unfinished high throughput cDNA sequences, each of which has 5'UTR and 3'UTR at both ends and part of a coding region. The sequence may also include introns. When the sequence becomes finished later, it moves to the corresponding taxonomic division. The sequence is accompanied with a keyword, HTC (High Throughput cDNA), which is dropped when the sequence is finished and moved to a taxonomic division. This release is published by the following DDBJ staff. General administration T. Gojobori, Y. Fukuma, Y. Katsube, M. Maruyama, K. Okuda, J. Sugiyama, H. Tsutsui (hold), Y. Ueda, A. Watanabe Database construction Y. Tateno, H. Aono, N. Asakawa, M. Ejima, M. Gojobori, A. Hasegawa, A. Hashizume, M. Hirahata, M. Hirashima, J. Mashima, A. Okada, M. Okaneya, T. Okido, M. Suzuki, H. Tsutsui, T. Umezawa, Y. Yamamoto Database software development and management H. Sugawara, M. Ota, S. Miyazaki, Y. Fujisawa, M. Fumoto, H. Hashimoto, T. Iizuka (hold), N. Ishizaka, K. Kaneda, T. Kato, Y. Kawanishi, T. Kubota, K. Mamiya, S. Misu, N. Nishimiya, T. Okayama, Y. Shigemoto, Y. Sugiyama, K. Suzuki, N. Takahashi, N. Tanaka, T. Takaki System management K. Nishikawa, K. Ikeo, N. Hoshi, T. Iizuka, A. Kusakabe, M. Nagura, F. Sugiyama, Y. Sugisaki, K. Yoshioka Editorial and public relations N. Saitou, K. Fukami-Kobayashi, Y. Daito, H. Ichikawa, K. Ichikawa, T. Kawamoto, J. Kohira, S. Nagira, Y. Sueki Center for Information Biology and DNA Data Bank of Japan National Institute of Genetics Mishima 411-8540, Japan Phone: +81 559 81 6853 FAX: +81 559 81 6849 E-mail: ddbj@ddbj.nig.ac.jp (for general inquiry) ddbjsub@ddbj.nig.ac.jp (for data submission) ddbjupdt@ddbj.nig.ac.jp (for updates and notification of publication) WWW: http://www.ddbj.nig.ac.jp (for DDBJ WWW server) http://sakura.ddbj.nig.ac.jp (for DDBJ sequence data submission system SAKURA) Acknowledgement: We are grateful to NCBI and EMBL/EBI for a firm friendship and an excellent collaboration with us. We also thank the Japanese Patent Office for a steady cooperation with us. The operation of DDBJ is supported by the Ministry of Education, Culture, Sports, Science and Technology, and we would gratefully note this here. DDBJ Database Release History Release Date Entries Bases Comments ------------------------------------------------------------------------------ 47 10/01 13,266,610 14,145,671,645 46 07/01 12,313,759 13,037,646,166 45 04/01 11,434,113 12,207,092,905 HTC division started 44 01/01 10,165,597 11,136,298,841 43 10/00 8,666,551 10,034,532,698 42 07/00 7,554,995 8,880,721,093 41 04/00 5,962,608 6,409,581,885 CON division started 40 01/00 5,388,125 4,762,696,173 RNA division terminated 39 10/99 4,810,773 3,728,000,562 NID and PID discarded 38 07/99 4,294,369 3,098,519,597 37 03/99 3,311,627 2,375,261,951 VERSION, /protein_id started 36 01/99 3,073,166 2,190,425,560 35 10/98 2,759,261 1,957,341,169 34 07/98 2,412,785 1,708,580,623 33 04/98 2,174,769 1,479,303,279 32 01/98 1,956,669 1,300,950,613 31 10/97 1,731,532 1,139,869,464 Adoption of the unified taxonomy database 30 07/97 1,534,115 992,788,339 NID and PID terminated 29 04/97 1,270,194 841,415,232 28 01/97 1,154,120 756,785,219 HTG division started ORG division terminated 27 10/96 936,697 608,103,057 GSS division started 26 07/96 835,552 551,932,448 25 04/96 744,490 499,300,364 /translation started 24 01/96 637,508 431,771,652 23 10/95 569,757 390,694,350 22 07/95 437,588 322,982,425 HUM division started 21 04/95 274,596 250,875,023 20 01/95 239,689 231,299,557 19 10/94 204,332 205,274,131 18 07/94 185,230 192,473,021 17 04/94 169,957 179,942,209 16 01/94 154,626 165,017,628 15 10/93 131,649 147,224,690 14 07/93 120,350 138,686,333 13 04/93 112,067 129,784,445 12 01/93 97,683 120,815,244 EST division started 11 07/92 65,693 84,839,075 10 01/92 59,317 77,805,556 GenBank/EMBL inclusion started 9 07/91 1,130 2,002,124 8 01/91 879 1,573,442 7 07/90 681 1,154,211 6 01/90 496 841,236 5 07/89 395 679,378 4 01/89 302 535,985 3 07/88 230 345,850 2 01/88 142 199,392 1 07/87 66 108,970 Started with DDBJ only ------------------------------------------------------------------------ This release covers 17 categories of organisms and others as follows: ------------------------------------------------------------------------------ ddbjbct.*** Category for bacteria ddbjest.*** Category for EST (expressed sequence tag) ddbjhtg.*** Category for HTG (high throughput genomic sequencing) ddbjhum.*** Category for human ddbjgss.*** Category for GSS (Genome Survey Sequence) ddbjinv.*** Category for invertebrates ddbjmam.*** Category for mammals other than primates and rodents ddbjpat.*** Category for patents ddbjphg.*** Category for phages ddbjpln.*** Category for plants ddbjpri.*** Category for primates other than human ddbjrod.*** Category for rodents ddbjsts.*** Category for STS (sequence tagged site) ddbjsyn.*** Category for synthetic DNAs ddbjuna.*** Category for unannotated sequences ddbjvrl.*** Category for viruses ddbjvrt.*** Category for vertebrates other than mammals ------------------------------------------------------------------------------ Each category then has the following nine files. Note that all the files except for ddbj***.seq are created by the user by use of seq2indexes as mentioned in the release note. ------------------------------------------------------------------------------ ddbj***.seq List of an entry in DDBJ format, see Table 1. ddbj***.acc List of the accession numbers, see Table 2 . ddbj***.aut List of the authors, see Table 3. ddbj***.dir List of the short directory in DDBJ style, see Table 4. ddbj***.idx List of indices, see Table 5. ddbj***.jou List of the journals, see Table 6. ddbj***.key List of the key words, see Table 7. ddbj***.org List of the species names, see Table 8. ddbj***.sdr List of the short directory in DDBJ style, see Table 9. ------------------------------------------------------------------------------ Table 1. Part of the contents in the file 'ddbjbct.seq'. This shows all pieces of information on one entry in DDBJ format. ------------------------------------------------------------------------------ LOCUS D87069 993 bp mRNA BCT 07-FEB-1999 DEFINITION Escherichia coli mRNA for RNA polymerase sigma subunit, truncated form of sigma-38, complete cds. ACCESSION D87069 VERSION D87069.1 KEYWORDS RNA polymerase sigma subunit, truncated form of sigma-38. SOURCE Escherichia coli (strain:W3110) cDNA to mRNA. ORGANISM Escherichia coli Bacteria; Proteobacteria; gamma subdivision; Enterobacteriaceae; Escherichia. REFERENCE 1 (bases 1 to 993) AUTHORS Jishage,M. TITLE Direct Submission JOURNAL Submitted (14-AUG-1996) to the DDBJ/EMBL/GenBank databases. Miki Jishage, National Institute of Genetics, Molecular Genetics; Yata 1111, Mishima, Shizuoka 411, Japan (E-mail:mjishage@lab.nig.ac.jp, Tel:0559-81-6742, Fax:0559-81-6746) REFERENCE 2 (bases 1 to 993) AUTHORS Jishage,M. and Ishihama,A. TITLE Variation in RNA polymerase sigma subunit composition within different stocks of Escherichia coli starin W3110 JOURNAL Unpublished (1996) REFERENCE 3 (sites) AUTHORS Ivanova,A., Renshaw,M., Guntaka,R. and Eisenstark,A. TITLE DNA base sequence variability in katF (putative sigma factor) gene Escherichia coli JOURNAL Nucleic Acids Res. 20, 5479-5480 (1992) REFERENCE 4 (sites) AUTHORS Takayanagi,Y., Tanaka,K. and Takahashi,H. TITLE Structure of the 5' upstream region and the regulation of the rpoS gene of Escherichia coli JOURNAL Mol Gen Genet 243, 525-531 (1994) COMMENT FEATURES Location/Qualifiers source 1..993 /organism="Escherichia coli" /sequenced_mol="cDNA to mRNA" /strain="W3110" CDS 1..810 /note="the gene has four single base changes, resulting in two amino acid substitutions and an amber mutation" /product="RNA polymerase sigma subunit, truncated form of sigma-38" /protein_id="BAA13238.1" /translation="MSQNTLKVHDLNEDAEFDENGVEVFDEKALVEYEPSDNDLAEEE LLSQGATQRVLDATQLYLGEIGYSPLLTAEEEVYFARRALRGDVASRRRMIESNLRLV VKIARRYGNRGLALLDLIEEGNLGLIRAVEKFDPERGFRFSTYATWWIRQTIERAIMN QTRTIRLPIHIVKELNVYLRTARELSHKLDHEPSAEEIAEQLDKPVDDVSRMLRLNER ITSVDTPLGGDSEKALLDILADEKENGPEDTTQDDDMKQSIVKWLFELNAK" /transl_table=11 mutation 75 /citation=[3] /replace="t" mutation 97 /citation=[3] /replace="t" mutation 99 /citation=[3] /replace="t" mutation 808 /citation=[3] /replace="t" BASE COUNT 254 a 223 c 291 g 225 t 0 others ORIGIN 1 atgagtcaga atacgctgaa agttcatgat ttaaatgaag atgcggaatt tgatgagaac 61 ggagttgagg tttttgacga aaaggcctta gtagaatatg aacccagtga taacgatttg 121 gccgaagagg aactgttatc gcagggagcc acacagcgtg tgttggacgc gactcagctt 181 taccttggtg agattggtta ttcaccactg ttaacggccg aagaagaagt ttattttgcg 241 cgtcgcgcac tgcgtggaga tgtcgcctct cgccgccgga tgatcgagag taacttgcgt 301 ctggtggtaa aaattgcccg ccgttatggc aatcgtggtc tggcgttgct ggaccttatc 361 gaagagggca acctggggct gatccgcgcg gtagagaagt ttgacccgga acgtggtttc 421 cgcttctcaa catacgcaac ctggtggatt cgccagacga ttgaacgggc gattatgaac 481 caaacccgta ctattcgttt gccgattcac atcgtaaagg agctgaacgt ttacctgcga 541 accgcacgtg agttgtccca taagctggac catgaaccaa gtgcggaaga gatcgcagag 601 caactggata agccagttga tgacgtcagc cgtatgcttc gtcttaacga gcgcattacc 661 tcggtagaca ccccgctggg tggtgattcc gaaaaagcgt tgctggacat cctggccgat 721 gaaaaagaga acggtccgga agataccacg caagatgacg atatgaagca gagcatcgtc 781 aaatggctgt tcgagctgaa cgccaaatag cgtgaagtgc tggcacgtcg attcggtttg 841 ctggggtacg aagcggcaac actggaagat gtaggtcgtg aaattggcct cacccgtgaa 901 cgtgttcgcc agattcaggt tgaaggcctg cgccgtttgc gcgaaatcct gcaaacgcag 961 gggctgaata tcgaagcgct gttccgcgag taa // ------------------------------------------------------------------------------ Table 2. Part of the contents in the file 'ddbjbct.acc'. The first column refers to the secondary accession number, second column to the locus name, and third to the primary accession number. The primary number may be the same as the secondary number. They are arranged in the ascending order of the secondary accession numbers. ------------------------------------------------------------------------------ D00001 -> ECOPBPAA X04516 D00002 -> ECOPYRH X04469 D00006 -> PNS981TET D00006 D00020 -> COLE2LYS D00020 D00021 -> COLE31YS D00021 D00038 -> BRLAM330 D00038 D00066 -> BAC139AC D00066 D00067 -> ECONANA M20207 D00069 -> ECOUVRD2 D00069 D00087 -> BACXYNAA D00087 ------------------------------------------------------------------------------ Table 3. Part of the contents in the file 'ddbjbct.aut'. For each author name given on the left to the arrow, the corresponding locus name and primary accession number are respectively listed on the right. They are arranged in the alphabetical order of the author names. ------------------------------------------------------------------------------ Aan,F. -> STYCRR X05210 Aan,F. -> STYENZI M76176 Aaronson,W. -> ECOKPSD M64977 Aaronson,W. -> ECONEUA J05023 Abad-Lapuebla,M.A. -> VIBTDHI D90238 Abdel-Mawgood,A.L. -> CYAPSBHA X16394 Abdel-Meguid,S.S. -> TRNGDRECM J01843 Abdelal,A. -> STYCARA M36540 Abdelal,A. -> STYCARAB X13200 Abdelal,A.H. -> PSENOSA M60717 ------------------------------------------------------------------------------ Table 4. Part of the short directory in DDBJ style in the file 'ddbjbct.dir'. For each locus name given in the first column, the corresponding primary accession number, molecular type, number of nucleotide pairs, and description for the locus are respectively listed. They are arranged in the alphabetical order of the locus names. ------------------------------------------------------------------------------ ABCAARAA M34830 ds-DNA 1624 A.aceti acetic acid resistance protein (aarA) gene, complete cds. ABCADHCC D00635 ds-DNA 4230 A. polyoxogenes alcohol dehydrogenase (EC 1.1.99.8) and cytochrome c genes. ABCALDH D00521 ds-DNA 2683 A.polyoxogenes membrane-bound aldehyde dehydrogenase gene, complete cds and flanks. ABCBCSAA M37202 ds-DNA 9540 A.xylinum bcs B, bcs C and bcs D genes, complete cds and bcs A gene, partial cds. ABCCELA M76548 ds-DNA 1165 Acetobacter xylinum UDP pyrophosphorylase (celA) gene, complete cds. ABCCELSYN X54676 ds-DNA 5363 A. xylinum gene for cellulose biosynthesis ABCIS1380 D10043 ds-DNA 1665 A.pasteurianus insertion sequence IS1380. ACAADH1 D90004 ds-DNA 2467 Acetobacter aceti(K6033) alcohol dehydrogenase subunit gene(adh1). ACCAAC2 M62833 ds-DNA 1123 Acinetobacter baumannii aminoglycoside acetyltr ansferase (aac2) gene, complete cds. ACCACEAA M62822 ds-DNA 1874 A.baumannii chloramphenicol acetyltransferase (cat) gene, complete cds. ------------------------------------------------------------------------------ Table 5. Part of the contents in the file 'ddbjbct.idx'. The first column refers to the locus name, second column to the starting site of the locus in byte, and third to its ending site in byte. They are arranged in the alphabetical order of the locus names. ------------------------------------------------------------------------------ %***************************** #ABCAARAA 0 3211 #ABCADHCC 3212 10608 #ABCALDH 10609 15864 #ABCBCSAA 15865 29583 #ABCCELA 29584 32289 #ABCCELSYN 32290 40960 #ABCIS1380 40961 44711 #ACAADH1 44712 49357 #ACCAAC2 49358 52395 ------------------------------------------------------------------------------ Table 6. Part of the contents in the file 'ddbjbct.jou'. This gives information on the journal in which sequence data were published. ------------------------------------------------------------------------------ (in) Chaloupka,J. and Krumphanzl,V. (Eds.); Extracellular Enzymes of Microorganisms: 129-137, Plenum Press, New York (1987) -> BACAMYABS M57457 (in) Ganesan,A.T., Chang,S. and Hoch,J.A. (Eds.); Molecular Cloning and Gene Regulation in Bacilli: 3-10, Academic Press, New York (1982) -> BACRG16S M55011 (in) Ganesan,A.T., Chang,S. and Hoch,J.A. (Eds.); Molecular Cloning and Gene Regulation in Bacilli: 3-10, Academic Press, New York (1982) -> BACRG16SA M55006 (in) Ganesan,A.T., Chang,S. and Hoch,J.A. (Eds.); Molecular Cloning and Gene Regulation in Bacilli: 3-10, Academic Press, New York (1982) -> BACRG16SB M55008 (in) Hoch,J.A. and Setlow,P. (Eds.); Molecular Biology of Microbial Differentiation: 85-94, American Society for Microbiology, Washington, DC (1985) -> BACSPOII M57606 (in) Holmgren,A. (Ed.); Thioredoxin and Glutaredoxin Systems: Structure and Function: 11-19, Unknown name, Unknown city (1986) -> ECOTRXA1 M54881 (in) Kjeldgaard,N.C. and Maaloe,O. (Eds.); Control of ribosome synthesis: 138-143, Academic Press, New York (1976) -> ECOLAC J01636 (in) Losick,R. and Chamberlin,M. (Eds.); RNA polymerase: 455-472, Cold Spring Harbor Laboratory, Cold Spring Harbor, NY (1976) -> ECOTGY1 K01197 (in) Sikes,C.S. and Wheeler,A.P. (Eds.); Surface reactive peptides and polymers. Discovery and commercialization.: 186-200, American Chemical Society, Washington, D.C. (1991) -> ECOTGP J01714 (in) Sund,H. and Blauer,G. (Eds.); Protein-Ligand Interactions: 193-207, Walter de Gruyter, New York (1975) -> ECOLAC J01636 (in) Wu,R. and Grossman,L. (Eds.); Methods in Enzymology, Recombinant DNA, part E: In press, Academic Press, New York, N.Y. (1986) -> PLMCG M11320 Acta Microbiol. Pol. 35, 175-190 (1986) -> ECOTGG1 M54893 Actinomycetologica 5, 14-17 (1991) -> STMARGG D00799 Adv. Biophys. 21, 115-133 (1986) -> R10REP M26840 Adv. Biophys. 21, 175-192 (1986) -> ECONUSAA M26839 Adv. Enzyme Regul. 21, 225-237 (1983) -> ECOPURFA M26893 Adv. Exp. Med. Biol. 195, 239-246 (1986) -> ECOAPT M14040 Agric. Biol. Chem. 50, 2155-2158 (1986) -> ECONANA M20207 Agric. Biol. Chem. 50, 2771-2778 (1986) -> BRLAM330 D00038 Agric. Biol. Chem. 51, 2019-2022 (1987) -> BACCGT D00129 Agric. Biol. Chem. 51, 2641-2648 (1987) -> STRSAGP D00219 Agric. Biol. Chem. 51, 2807-2809 (1987) -> BACPGECR M35503 Agric. Biol. Chem. 51, 3133-3135 (1987) -> BACXYLAP D00312 Agric. Biol. Chem. 51, 455-463 (1987) -> BACHDCRY D00117 Agric. Biol. Chem. 51, 953-955 (1987) -> BACXYNAA D00087 Agric. Biol. Chem. 52, 1565-1573 (1988) -> BACIP135 D00348 Agric. Biol. Chem. 52, 1785-1789 (1988) -> BACTMR D00343 Agric. Biol. Chem. 52, 2243-2246 (1988) -> PSEGI D00342 Agric. Biol. Chem. 52, 399-406 (1988) -> BACAMYEB M35517 Agric. Biol. Chem. 52, 479-487 (1988) -> ECAPALI D00217 ------------------------------------------------------------------------------ Table 7. Part of the contents in the file 'ddbjbct.key'. For the locus and accession number respectively given on the right to the arrow, the corresponding key words are listed on the left. ------------------------------------------------------------------------------ A.aceti acetic acid resistance protein (aarA) gene, complete cds. -> ABCAARAA M34830 acetic acid resistance protein. -> ABCAARAA M34830 Cloning of genes responsible for acetic acid resistance in acetobacter aceti -> ABCAARAA M34830 A. polyoxogenes alcohol dehydrogenase (EC 1.1.99.8) and cytochrome c genes. -> ABCADHCC D00635 alcohol dehydrogenase; cytochrome c. -> ABCADHCC D00635 Cloning and sequencing of the gene cluster encoding two subunits of membrane- bound alcohol dehydrogenase from Acetobacter polyoxogenes -> ABCADHCC D00635 These data kindly submitted in computer readable form by: Toshimi Tamaki Nakano Central Biochemical Institute 2-6 Nakamura-cho Handa-shi, Aichi-ken 475 Japan Phone: 0569-21-3331 Fax: 0569-23-8486 -> ABCADHCC D00635 A.polyoxogenes membrane-bound aldehyde dehydrogenase gene, complete cds and flanks. -> ABCALDH D00521 aldehyde dehydrogenase gene; ethanol oxidation; membrane-bound enzyme. -> ABCALDH D00521 Nucleotide sequence of the membrane-bound aldehyde dehydrogenase gene from Acetobacter polyoxogenes -> ABCALDH D00521 ------------------------------------------------------------------------------ Table 8. Part of the contents in the file 'ddbjbct.org'. For the locus and accession number respectively given on the right to the arrow, the corresponding taxonomic names are listed on the left. They are arranged in the alphabetical order of the species names. ------------------------------------------------------------------------------ A. nidulans 6301 DNA. Anacystis nidulans Prokaryota; Bacteria; Gracilicutes; Oxyphotobacteria; Cyanobacteria. -> ANIRUBPS X00019 A. nidulans DNA, clone pAN4. Anacystis nidulans Prokaryota; Bacteria; Gracilicutes; Oxyphotobacteria; Cyanobacteria. -> ANIRGGX X00343 A. nidulans DNA. Anacystis nidulans Prokaryota; Bacteria; Gracilicutes; Oxyphotobacteria; Cyanobacteria. -> ANIRGG X00512 A. polyoxogenes genomic DNA. Acetobacter polyoxogenes Prokaryota; Bacteria; Gracilicutes; Scotobacteria; Aerobic rods and cocci; Azotobacteraceae. - > ABCADHCC D00635 A. quadruplicatum (strain PR-6) DNA, clone pAQPR1. Agmenellum quadruplicatum Prokaryota; Bacteria; Gracilicutes; Oxyphotobacteria; Cyanobacteria. -> AQUPCAB K02660 A. quadruplicatum (strain PR6) DNA. Agmenellum quadruplicatum Prokaryota; Bacteria; Gracilicutes; Oxyphotobacteria; Cyanobacteria. -> AQUCPCAB K02659 A. vinelandii DNA. Azotobacter vinelandii Prokaryota; Bacteria; Gracilicutes; Scotobacteria; Aerobic rods and cocci; Azotobacteraceae. -> AVINIFUSV M17349 A.aceti (strain 10-8) DNA, clone pAR1611. Acetobacter aceti Prokaryota; Bacteria; Gracilicutes; Scotobacteria; Aerobic rods and cocci; Azotobacteraceae. -> ABCAARAA M34830 A.actinomycetemcomitans (strain JP2) DNA, clone lambda-OP8. Actinobacillus actinomycetemcomitans Prokaryota; Bacteria; Gracilicutes; Scotobacteria; Facultatively anaerobic rods; Pasteurellaceae. -> ACNLKTXN M27399 A.anitratum DNA, clone pLJD1. Acinetobacter anitratum Prokaryota; Bacteria; Gracilicutes; Scotobacteria; Neisseriaceae. -> ACCCITSYN M33037 ------------------------------------------------------------------------------ Table 9. Part of the short directory file in DDBJ style in the file 'ddbjbct.sdr'. The short directory file contains brief descriptions of all of the sequence entries contained in the DDBJ style. ------------------------------------------------------------------------------ ABCAARAA A.aceti acetic acid resistance protein (aarA) gene, complete 1624bp ABCADHCC A. polyoxogenes alcohol dehydrogenase (EC 1.1.99.8) and 4230bp ABCALDH A.polyoxogenes membrane-bound aldehyde dehydrogenase gene, 2683bp ABCBCSABCD A.xylinum bcs A, B, C and D genes, complete cds's. 9540bp ABCCELA Acetobacter xylinum UDP pyrophosphorylase (celA) gene, 1165bp ABCCELSYN A. xylinum gene for cellulose biosynthesis 5363bp ABCIS1380 A.pasteurianus insertion sequence IS1380. 1665bp ACAADH1 Acetobacter aceti(K6033) alcohol dehydrogenase subunit 2467bp ACCAAC2 Acinetobacter baumannii aminoglycoside acetyltransferase 1123bp ACCACEAA A.baumannii chloramphenicol acetyltransferase (cat) gene, 1874bp ACCAPHA6 Acinetobacter baumannii aphA-6 gene. 1170bp ACCBENABCA A.calcoaceticus BenA, BenB, BenC, BenD, and BenE proteins 15922bp ACCCAT Acinetobacter calcoaceticus cat operon. 15922bp ACCCATAM A.calcoaceticus catA and catM genes, encoding catechol 1, 5537bp ACCCHMO Acinetobacter sp. cyclohexanone monooxygenase gene, complete 2128bp ACCCITSYN A.anitratum citrate synthase gene, complete cds. 1895bp ------------------------------------------------------------------------------ In addition to the 9 tables the four following index files are included in this release. These files were prepared irrespective of the 14 categories of taxonomic divisions. Accession number index file Keyword phrase index file Journal citation index file Gene name index file A brief description is given for each file in the following. Table 10. Part of the accession number index file in the 'ddbjacc.idx'. The following excerpt from the accession number index file illustrates the format of the index. ------------------------------------------------------------------------------ D00100 PSEASPAA BCT D00100 D00101 RABNP450R MAM D00101 D00102 HUMLTX HUM D00102 D00103 AFARRN5SA BCT D00103 AFRRN5SA BCT X05517 D00104 AFARRN5SB BCT D00104 AFRRN5SB BCT X05518 D00105 AFARRN5S BCT D00105 ASRRN5S BCT X05524 D00106 ACH5SRR BCT D00106 AXRRN5S BCT X05522 AXRRN5SA BCT X05523 D00107 ACH5SRRX BCT D00107 ACRRN5S BCT X05521 ------------------------------------------------------------------------------ Table 11. Part of the keyword phrase index file in the 'ddbjkey.idx'. Keyword phrases consist of names for gene products and other characteristics of sequence entries. ------------------------------------------------------------------------------ A CHANNEL DROCHA INV M17155 A COMPONENT SQLCVEA VRL M38183 A LOCUS GORGOGOA3 PRI X54375 GORGOGOA4 PRI X54376 A LOCUS ALLELE GORA0101 PRI X60258 GORA0201 PRI X60259 GORA0401 PRI X60257 GORA0501 PRI X60256 A MULTI-GENE FAMILY RICGLUTE PLN D00584 A PROTEIN MS2AAR PHG M25187 ST1APCS PHG M25396 A SEQUENCE HS5TOA30 VRL D00148 HS5TOA31 VRL D00147 ------------------------------------------------------------------------------ Table 12. Part of the author name index file in 'ddbjaut.idx'. The author name index file lists all of the author names that appear in the citations. ------------------------------------------------------------------------------ ABE,A. HUMMHDRBWE PRI M27509 HUMMHDRBWF PRI M27510 HUMMHDRBWG PRI M27511 YSCGAL11A PLN M22481 ABE,C. S85445 BCT S85445 ABE,E. M23442 UNA M23442 ABE,H. CHKADF VRT M55660 CHKCOF VRT M55659 ABE,K. CHPCLAC PRI D11383 CHPIMRF PRI D11384 CUGCUR09 PLN X64110 CUGCUR37 PLN X64111 HPCCEXPA VRL M55970 HPCCPEP1 VRL D10687 HPCCPEP2 VRL D10688 HPCHABC82 VRL X51587 HPCNS2APA VRL M55972 HPCNS2PA VRL M55971 HPCNS2PB VRL M55973 HPCNS5PA VRL M55974 MUSKE2 ROD M65255 MUSKE2A ROD M65256 MZECYS PLN D10622 RICCPI PLN J03469 RICGLUTE PLN D00584 RICLNOCI PLN J05595 RICOCS PLN M29259 RICORYII PLN X57658 RICOZA PLN D90406 RICOZB PLN D90407 RICOZC PLN D90408 S54524 PLN S54524 S54526 PLN S54526 S54530 PLN S54530 S73960 ROD S73960 ------------------------------------------------------------------------------ Table 13. Part of the journal citation index file in 'ddbjjou.idx'. The journal citation index file lists all of the citations that appear in the references. ------------------------------------------------------------------------------ ACTA BIOCHIM. BIOPHYS. SIN. 23, 246-253 (1992) HUMPLASINS HUM M98056 ACTA BIOCHIM. BIOPHYS. SIN. 28, 233-239(1996) TKTII PLN X82230 ACTA BIOCHIM. POL. 24, 301-318 (1977) LUPTRFJ PLN K00345 LUPTRFN PLN K00346 ACTA BIOCHIM. POL. 26, 369-381(1979) HVTRNPHE PLN X02683 ACTA BIOCHIM. POL. 29, 143-149 (1982) EMEMTA PLN M32572 EMEMTB PLN M32573 EMEMTC PLN M32574 EMEMTD PLN M32575 EMEMTE PLN M32576 ACTA BIOCHIM. POL. 34, 21-27 (1987) LUPNOSP PLN M32571 ------------------------------------------------------------------------------ Table 14. Part of the gene name index file in 'ddbjgen.idx'. This file lists all the gene names that appear in the feature table. ------------------------------------------------------------------------------ AACC8 STMAACC8 BCT M55426 AACC9 MPUAACC9 BCT M55427 AACT HUMA1ACM PRI K01500 HUMA1ACMA PRI X00947 HUMA1ACMB PRI M18035 HUMAACT1 PRI M18906 HUMAACT2 PRI M22533 HUMAACTA PRI J05176 AAD INTINTORF BCT L06418 LMOMO229D BCT X17478 AAD A1 ENTAAC3VI BCT M88012 AAD9 ENEAAD9A BCT M69221 AADA LMOMO229A BCT X17479 S52249 BCT S52249 SYNAADA SYN M60473 TRNTAAB BCT M55547 TRNTN21CAS BCT M86913 ------------------------------------------------------------------------------ The files in this release are arranged in the following order with non- labeled format. Release note ddbjrel.txt 1068 records Category for bacteria1, 29025 entries, 121956376 bases ddbjbct1.seq 4832265 records Category for bacteria2, 43866 entries, 118593550 bases ddbjbct2.seq 4883765 records Category for bacteria3, 43799 entries, 91862295 bases ddbjbct3.seq 3997579 records Category for EST1 (expressed sequence tag), 93250 entries, 34643977 bases ddbjest1.seq 5514114 records Category for EST2 (expressed sequence tag), 96152 entries, 39200566 bases ddbjest2.seq 5559781 records Category for EST3 (expressed sequence tag), 97505 entries, 37828724 bases ddbjest3.seq 5555318 records Category for EST4 (expressed sequence tag), 90869 entries, 27873888 bases ddbjest4.seq 5479501 records Category for EST5 (expressed sequence tag), 97158 entries, 38218007 bases ddbjest5.seq 5574042 records Category for EST6 (expressed sequence tag), 101443 entries, 40202921 bases ddbjest6.seq 5622536 records Category for EST7 (expressed sequence tag), 100399 entries, 38835835 bases ddbjest7.seq 5585886 records Category for EST8 (expressed sequence tag), 99295 entries, 38407363 bases ddbjest8.seq 5559184 records Category for EST9 (expressed sequence tag), 100948 entries, 40006175 bases ddbjest9.seq 5614543 records Category for EST10 (expressed sequence tag), 101629 entries, 39760337 bases ddbjest10.seq 5587003 records Category for EST11 (expressed sequence tag), 99066 entries, 41087960 bases ddbjest11.seq 5545339 records Category for EST12 (expressed sequence tag), 101495 entries, 44213727 bases ddbjest12.seq 5591123 records Category for EST13 (expressed sequence tag), 107794 entries, 43197016 bases ddbjest13.seq 5627164 records Category for EST14 (expressed sequence tag), 103258 entries, 41842759 bases ddbjest14.seq 5581832 records Category for EST15 (expressed sequence tag), 99003 entries, 41529235 bases ddbjest15.seq 5538809 records Category for EST16 (expressed sequence tag), 95236 entries, 42111046 bases ddbjest16.seq 5529756 records Category for EST17 (expressed sequence tag), 100881 entries, 41792907 bases ddbjest17.seq 5604864 records Category for EST18 (expressed sequence tag), 99427 entries, 43585353 bases ddbjest18.seq 5585155 records Category for EST19 (expressed sequence tag), 95180 entries, 40108218 bases ddbjest19.seq 5540523 records Category for EST20 (expressed sequence tag), 99301 entries, 41899485 bases ddbjest20.seq 5561150 records Category for EST21 (expressed sequence tag), 102422 entries, 47879727 bases ddbjest21.seq 5354504 records Category for EST22 (expressed sequence tag), 107870 entries, 76084811 bases ddbjest22.seq 5414710 records Category for EST23 (expressed sequence tag), 122163@entries, 60883811 bases ddbjest23.seq 5748756 records Category for EST24 (expressed sequence tag), 108632 entries, 44608346 bases ddbjest24.seq 5673501 records Category for EST25 (expressed sequence tag), 90629 entries, 23647830 bases ddbjest25.seq 5552364 records Category for EST26 (expressed sequence tag), 90128 entries, 25948077 bases ddbjest26.seq 5509789 records Category for EST27 (expressed sequence tag), 59241 entries, 14874068 bases ddbjest27.seq 5212923 records Category for EST28 (expressed sequence tag), 59117 entries, 15356532 bases ddbjest28.seq 5198037 records Category for EST29 (expressed sequence tag), 79594 entries, 28194244 bases ddbjest29.seq 5393525 records Category for EST30 (expressed sequence tag), 122284 entries, 54901272 bases ddbjest30.seq 5784034 records Category for EST31 (expressed sequence tag), 108176 entries, 5516215 bases ddbjest31.seq 5461317 records Category for EST32 (expressed sequence tag), 99661 entries, 46008114 bases ddbjest32.seq 5530853 records Category for EST33 (expressed sequence tag), 90130 entries, 38931486 bases ddbjest33.seq 5420501 records Category for EST34 (expressed sequence tag), 95084 entries, 41607193 bases ddbjest34.seq 5485595 records Category for EST35 (expressed sequence tag), 98220 entries, 40842168 bases ddbjest35.seq 5583241 records Category for EST36 (expressed sequence tag), 101139 entries, 39731548 bases ddbjest36.seq 5643189 records Category for EST37 (expressed sequence tag), 87564 entries, 37661172 bases ddbjest37.seq 5468721 records Category for EST38 (expressed sequence tag), 92570 entries, 41291996 bases ddbjest38.seq 5496262 records Category for EST39 (expressed sequence tag), 103363 entries, 46400512 bases ddbjest39.seq 5634412 records Category for EST40 (expressed sequence tag), 96821 entries, 37370645 bases ddbjest40.seq 5584910 records Category for EST41 (expressed sequence tag), 110612 entries, 46438391 bases ddbjest41.seq 5741655 records Category for EST42 (expressed sequence tag), 71491 entries, 24086523 bases ddbjest42.seq 5269504 records Category for EST43 (expressed sequence tag), 58825 entries, 15873545 bases ddbjest43.seq 5130540 records Category for EST44 (expressed sequence tag), 58893 entries, 17050434 bases ddbjest44.seq 5119921 records Category for EST45 (expressed sequence tag), 59225 entries, 16408208 bases ddbjest45.seq 5111714 records Category for EST46 (expressed sequence tag), 58995 entries, 17134778 bases ddbjest46.seq 5101934 records Category for EST47 (expressed sequence tag), 58896 entries, 16618660 bases ddbjest47.seq 5140862 records Category for EST48 (expressed sequence tag), 59263 entries, 16504244 bases ddbjest48.seq 5112342 records Category for EST49 (expressed sequence tag), 60682 entries, 16387537 bases ddbjest49.seq 5107652 records Category for EST50 (expressed sequence tag), 60421 entries, 16958138 bases ddbjest50.seq 5092673 records Category for EST51 (expressed sequence tag), 60867 entries, 16561733 bases ddbjest51.seq 5162968 records Category for EST52 (expressed sequence tag), 73464 entries, 28493966 bases ddbjest52.seq 5333329 records Category for EST53 (expressed sequence tag), 99160 entries, 40461528 bases ddbjest53.seq 5662732 records Category for EST54 (expressed sequence tag), 100758 entries, 41982653 bases ddbjest54.seq 5633594 records Category for EST55 (expressed sequence tag), 101486 entries, 57363939 bases ddbjest55.seq 5514262 records Category for EST56 (expressed sequence tag), 101573 entries, 55899956 bases ddbjest56.seq 5528979 records Category for EST57 (expressed sequence tag), 104054 entries, 51380344 bases ddbjest57.seq 5635519 records Category for EST58 (expressed sequence tag), 94217 entries, 52993926 bases ddbjest58.seq 5444653 records Category for EST59 (expressed sequence tag), 96399 entries, 47957430 bases ddbjest59.seq 5518645 records Category for EST60 (expressed sequence tag), 94566 entries, 47119291 bases ddbjest60.seq 5509761 records Category for EST61 (expressed sequence tag), 101647 entries, 59771385 bases ddbjest61.seq 5556129 records Category for EST62 (expressed sequence tag), 88180 entries, 47656816 bases ddbjest62.seq 5405834 records Category for EST63 (expressed sequence tag), 97131 entries, 52688576 bases ddbjest63.seq 5525720 records Category for EST64 (expressed sequence tag), 96147 entries, 58685494 bases ddbjest64.seq 5467922 records Category for EST65 (expressed sequence tag), 97630 entries, 60454452 bases ddbjest65.seq 5501209 records Category for EST66 (expressed sequence tag), 92885 entries, 42222631 bases ddbjest66.seq 5555312 records Category for EST67 (expressed sequence tag), 90274 entries, 41781583 bases ddbjest67.seq 5432757 records Category for EST68 (expressed sequence tag), 90552 entries, 51699210 bases ddbjest68.seq 5437070 records Category for EST69 (expressed sequence tag), 95380 entries, 61252757 bases ddbjest69.seq 5494630 records Category for EST70 (expressed sequence tag), 99041 entries, 44056688 bases ddbjest70.seq 5593708 records Category for EST71 (expressed sequence tag), 97195 entries, 38442422 bases ddbjest71.seq 5576780 records Category for EST72 (expressed sequence tag), 97847 entries, 43541607 bases ddbjest72.seq 5565163 records Category for EST73 (expressed sequence tag), 90274 entries, 47377493 bases ddbjest73.seq 5385216 records Category for EST74 (expressed sequence tag), 105060 entries, 62675250 bases ddbjest74.seq 5533355 records Category for EST75 (expressed sequence tag), 101675 entries, 61396867 bases ddbjest75.seq 5497699 records Category for EST76 (expressed sequence tag), 94297 entries, 58552227 bases ddbjest76.seq 5486780 records Category for EST77 (expressed sequence tag), 97664 entries, 65434572 bases ddbjest77.seq 5468766 records Category for EST78 (expressed sequence tag), 91710 entries, 56876254 bases ddbjest78.seq 5352450 records Category for EST79 (expressed sequence tag), 100756 entries, 58767891 bases ddbjest79.seq 5610646 records Category for EST80 (expressed sequence tag), 94744 entries, 64447793 bases ddbjest80.seq 5410715 records Category for EST81 (expressed sequence tag), 94450 entries, 59979256 bases ddbjest81.seq 5422143 records Category for EST82 (expressed sequence tag), 100086 entries, 46259533 bases ddbjest82.seq 5619397 records Category for EST83 (expressed sequence tag), 102138 entries, 50925587 bases ddbjest83.seq 5610038 records Category for EST84 (expressed sequence tag), 98164 entries, 58019605 bases ddbjest84.seq 5516741 records Category for EST85 (expressed sequence tag), 84629 entries, 45010979 bases ddbjest85.seq 5426893 records Category for EST86 (expressed sequence tag), 96213 entries, 52792370 bases ddbjest86.seq 5536339 records Category for EST87 (expressed sequence tag), 87267 entries, 47153352 bases ddbjest87.seq 5394892 records Category for EST88 (expressed sequence tag), 93712 entries, 56964229 bases ddbjest88.seq 5479912 records Category for EST89 (expressed sequence tag), 93554 entries, 58397456 bases ddbjest89.seq 5464631 records Category for EST90 (expressed sequence tag), 99186 entries, 51829141 bases ddbjest90.seq 5458633 records Category for EST91 (expressed sequence tag), 136135 entries, 33542382 bases ddbjest91.seq 5969657 records Category for EST92 (expressed sequence tag), 97555 entries, 48247131 bases ddbjest92.seq 5570014 records Category for EST93 (expressed sequence tag), 95321 entries, 35439485 bases ddbjest93.seq 5530982 records Category for EST94 (expressed sequence tag), 96808 entries, 34599566 bases ddbjest94.seq 5573839 records Category for EST95 (expressed sequence tag), 102837 entries, 35281371 bases ddbjest95.seq 5672828 records Category for EST96 (expressed sequence tag), 93196 entries, 37972270 bases ddbjest96.seq 5502899 records Category for EST97 (expressed sequence tag), 45736 entries, 16291543 bases ddbjest97.seq 2507018 records Category for GSS1 (genome survey sequence), 113664 entries, 76046239 bases ddbjgss1.seq 5397415 records Category for GSS2 (genome survey sequence), 93305 entries, 74250224 bases ddbjgss2.seq 5269542 records Category for GSS3 (genome survey sequence), 84526 entries, 72049195 bases ddbjgss3.seq 5215263 records Category for GSS4 (genome survey sequence), 75857 entries, 72159922 bases ddbjgss4.seq 5115497 records Category for GSS5 (genome survey sequence), 110686 entries, 53644001 bases ddbjgss5.seq 5591828 records Category for GSS6 (genome survey sequence), 118392 entries, 51683027 bases ddbjgss6.seq 6004375 records Category for GSS7 (genome survey sequence), 115857 entries, 55386740 bases ddbjgss7.seq 5945267 records Category for GSS8 (genome survey sequence), 104784 entries, 55330609 bases ddbjgss8.seq 5808709 records Category for GSS9 (genome survey sequence), 103910 entries, 51935407 bases ddbjgss9.seq 5809202 records Category for GSS10 (genome survey sequence), 101896 entries, 52016626 bases ddbjgss10.seq 5764882 records Category for GSS11 (genome survey sequence, 98108 entries, 49464600 bases ddbjgss11.seq 5671745 records Category for GSS12 (genome survey sequence), 99658 entries, 55918266 bases ddbjgss12.seq 5669646 records Category for GSS13 (genome survey sequence), 92238 entries, 47970242 bases ddbjgss13.seq 5547687 records Category for GSS14 (genome survey sequence), 98744 entries, 51564199 bases ddbjgss14.seq 5656280 records Category for GSS15 (genome survey sequence), 94739 entries, 43619385 bases ddbjgss15.seq 5657146 records Category for GSS16 (genome survey sequence), 97284 entries, 48386077 bases ddbjgss16.seq 5672368 records Category for GSS17 (genome survey sequence), 96742 entries, 53806725 bases ddbjgss17.seq 5657340 records Category for GSS18 (genome survey sequence), 82726 entries, 38975604 bases ddbjgss18.seq 5467443 records Category for GSS19 (genome survey sequence), 75198 entries, 38472038 bases ddbjgss19.seq 5340294 records Category for GSS20 (genome survey sequence), 77212 entries, 33269056 bases ddbjgss20.seq 5378133 records Category for GSS21 (genome survey sequence), 86577 entries, 50850954 bases ddbjgss21.seq 5434605 records Category for GSS22 (genome survey sequence), 76572 entries, 36038068 bases ddbjgss22.seq 5363075 records Category for GSS23 (genome survey sequence), 94178 entries, 54524638 bases ddbjgss23.seq 5635452 records Category for GSS24 (genome survey sequence), 78007 entries, 31225336 bases ddbjgss24.seq 5393373 records Category for GSS25 (genome survey sequence), 91809 entries, 42909241 bases ddbjgss25.seq 5593526 records Category for GSS26 (genome survey sequence), 78181 entries, 42709315 bases ddbjgss26.seq 5369603 records Category for GSS27 (genome survey sequence), 99663 entries, 53682686 bases ddbjgss27.seq 5775688 records Category for GSS28 (genome survey sequence), 101985 entries, 64519491 bases ddbjgss28.seq 5729370 records Category for GSS29 (genome survey sequence), 62924 entries, 27707828 bases ddbjgss29.seq 3157657 records Category for HTC (high throughput cDNA), 19658 entries, 24276543 bases ddbjhtc.seq 2294930 records Category for HTG1 (high throughput genome sequence), 1524 entries, 229147997 bases ddbjhtg1.seq 3985356 records Category for HTG2 (high throughput genome sequence), 1525 entries, 227485426 bases ddbjhtg2.seq 3999991 records Category for HTG3 (high throughput genome sequence), 3373 entries, 225068732 bases ddbjhtg3.seq 4025134 records Category for HTG4 (high throughput genome sequence), 2439 entries, 227201818 bases ddbjhtg4.seq 4006727 records Category for HTG5 (high throughput genome sequence), 2430 entries, 227100641 bases ddbjhtg5.seq 4006354 records Category for HTG6 (high throughput genome sequence), 1564 entries, 226029125 bases ddbjhtg6.seq 4010632 records Category for HTG7 (high throughput genome sequence), 1455 entries, 226374174 bases ddbjhtg7.seq 4010760 records Category for HTG8 (high throughput genome sequence), 5663 entries, 219554287 bases ddbjhtg8.seq 4067084 records Category for HTG9 (high throughput genome sequence), 27422 entries, 185373391 bases ddbjhtg9.seq 4378137 records Category for HTG10 (high throughput genome sequence), 8859 entries, 214228961 bases ddbjhtg10.seq 4114407 records Category for HTG11 (high throughput genome sequence), 4710 entries, 222632109 bases ddbjhtg11.seq 4041558 records Category for HTG12 (high throughput genome sequence), 9187 entries, 215481343 bases ddbjhtg12.seq 4109419 records Category for HTG13 (high throughput genome sequence), 5836 entries, 218420358 bases ddbjhtg13.seq 4073954 records Category for HTG14 (high throughput genome sequence), 1662 entries, 227064625 bases ddbjhtg14.seq 3997212 records Category for HTG15 (high throughput genome sequence), 1453 entries, 229924390 bases ddbjhtg15.seq 3980148 records Category for HTG16 (high throughput genome sequence), 2782 entries, 219218423 bases ddbjhtg16seq 4023284 records Category for HTG17 (high throughput genome sequence), 1448 entries, 229171885 bases ddbjhtg17.seq 3990039 records Category for HTG18 (high throughput genome sequence), 1367 entries, 229168472 bases ddbjhtg18.seq 3986506 records Category for HTG19 (high throughput genome sequence), 1235 entries, 231211735 bases ddbjhtg19.seq 3974466 records Category for HTG20 (high throughput genome sequence), 1424 entries, 229952252 bases ddbjhtg20.seq 3989651 records Category for HTG21 (high throughput genome sequence), 871 entries, 114350350 bases ddbjhtg21.seq 1956970 records Category for human1, 7856 entries, 201421931 bases ddbjhum1.seq 4312951 records Category for human2, 1573 entries, 213885483 bases ddbjhum2.seq 4199763 records Category for human3, 1501 entries, 218036101 bases ddbjhum3.seq 4144822 records Category for human4, 1507 entries, 219235166 bases ddbjhum4.seq 4123359 records Category for human5, 22640 entries, 185735264 bases ddbjhum5.seq 4478845 records Category for human6, 34579 entries, 159997213 bases ddbjhum6.seq 4630359 records Category for human7, 4234 entries, 203489825 bases ddbjhum7.seq 4171498 records Category for huma8, 2135 entries, 213017772 bases ddbjhum8.seq 4097033 records Category for human9, 2733 entries, 213574187 bases ddbjhum9.seq 4106797 records Category for human10, 32747 entries, 167961780 bases ddbjhum10.seq 4655143 records Category for human11, 54417 entries, 93053740 bases ddbjhum11.seq 3982437 records Category for invertebrates1, 7643 entries, 211349621 bases ddbjinv1.seq 4109505 records Category for invertebrates2, 45233 entries, 147445392 bases ddbjinv2.seq 4791895 records Category for invertebrates3, 45834 entries, 137405972 bases ddbjinv3.seq 4771338 records Category for mammals, 34676 entries, 33133748 bases ddbjmam.seq 1976145 records Category for patents, 449689 entries, 182691358 bases ddbjpat.seq 12960153 records Category for phages, 1787 entries, 5443220 bases ddbjphg.seq 237678 records Category for plants1, 38511 entries, 152373486 bases ddbjpln1.seq 4680765 records Category for plants2, 70022 entries, 113227591 bases ddbjpln2.seq 5156982 records Category for plants3, 55066 entries, 130767196 bases ddbjpln3.seq 4938320 records Category for plants4, 10645 entries, 26494892 bases ddbjpln4.seq 1021495 records Category for primates, 12889 entries, 16116489 bases ddbjpri.seq 844128 records Category for rodents, 70611 entries, 153898296 bases ddbjrod.seq 5900984 records Category for STS (sequence tagged site), 111710 entries, 44035082 bases ddbjsts.seq 6620426 records Category for synthetic DNAs, 6381 entries, 11773036 bases ddbjsyn.seq 492489 records Category for unannotated sequences, 568 entries, 332899 bases ddbjuna.seq 26256 records Category for viruses, 133264 entries, 117121707 bases ddbjvrl.seq 8193995 records Category for vertebrates, 61140 entries, 57794543 bases ddbjvrt.seq 3535214 records Accession number index file ddbjacc.idx 13297194 records Keyword phrase index file ddbjkey.idx 4737304 records Journal citation index file ddbjjou.idx 7400818 records Gene name index file ddbjgen.idx 740761 records