Processing unspecified species names at NCBI Taxonomy ===================================================== Historically NCBI/GenBank taxonomy database attempted to generate a unique name when the submitted organism name does not have a species name (e.g., just 'Rattus' or 'Rattus sp.' instead of 'Rattus norvegicus'). This was done by adding a unique string so that e.g., 'Rattus sp.' becomes 'Rattus sp. T-612'. To cut down on the number of taxonomic updates required, NCBI/GenBank taxonomy has been adding names without requiring the addition of a strain or another unique identifier since 2017. At the moment, this is restricted to viruses and microbes, including prokaryotes (Bacteria and Archaea) and eukaryotes (Fungi, Stramenopiles & unicellular eukaryotes) and the remaining names in the Metazoa and Viridiplantae will continue to be treated as before. Previously: Bacillus sp. St12345 Now: Bacillus sp. This means that in these cases sequences of several species will be mapping to a single node in the taxonomy database. Exceptions to the stated change, where names will be added as before: a. In cases where it is clear that specific sequences belong to a single provisional species name, or a novel taxon they will still be added with a unique identifier (usually submitter initials and year, or strain). b. Genome project strains continue to have the strain ids. added to the name as sp. . However, metagenome projects do not require a strain/isolate id. in the name unless the submitters are proposing a novel taxon. c. Cyanobacteria continue to have strain ids. added to the name in all instances. d. Specific submitter requests for a unique id will still be considered.