Pathogenwatch
  • Welcome to Pathogenwatch
  • 🎉Announcements
  • ▶️A "Getting Started" Tutorial
  • 🎦Video Tutorials
  • 🧐Useful Links
  • 📖How to use Pathogenwatch
    • Uploading Genomes
    • Genome Reports
    • Browsing Genomes
    • Editing Metadata
    • 🚮Deleting genomes
    • Downloads
    • Creating A Collection
    • Browsing Collections
    • Sharing a collection
    • Genomic Context Search
    • Using The Interactive Collection Views
      • The Map View
      • The Tree Viewer
      • The Filter Bar
      • The Metadata Tables
        • Uploaded Metadata
        • Typing Results
        • Genome Statistics
        • Antimicrobial Resistance
    • Private Metadata
  • 📖Technical Descriptions
    • Species Assignment
      • Speciator
    • Sequence Typing Methods
      • cgMLST
      • Genotyphi
      • Kaptive
      • Kleborate
      • Klebsiella LIN Codes
      • MLST
      • NG-MAST
      • Pangolin
      • PopPUNK
      • SeroBA
      • Vista
      • SISTR
    • Antimicrobial Resistance Prediction
      • SPN-PBP-AMR
      • Kleborate
      • Pathogenwatch AMR
    • Inctyper
    • cgMLST Clustering
    • SARS-CoV-2 Notable Mutations
    • SARS-CoV-2 Genome Tree
    • Core Genome Tree
      • Core Assignment
      • Reference Assignment
      • Core Filter
      • Tree Construction
    • Short Read Assembly
  • ❓FAQ
  • 💾Public data downloads
  • 💊WHO bacterial priority pathogens
  • 📜Release Notes 2025
  • Release Notes 2024
  • Release Notes 2023
  • Release Notes 2022
  • Release Notes 2019-2021
  • ⚠️Privacy and Terms Of Service
  • 📣How to cite
  • 🙏Acknowledgements
  • ❗Report an Issue
Powered by GitBook
On this page
  • About
  • Method
  • Results
  • How to cite
  1. Technical Descriptions
  2. Sequence Typing Methods

MLST

PreviousKlebsiella LIN CodesNextNG-MAST

Last updated 11 months ago

About

MLST schemes are based around a community-agreed set of 7 gene loci present in all strains of the species. A database of validated allele sequences is maintained for each locus and a code assigned to each one. An "ST" code is then generated from the unique combination alleles. The schemes supported by Pathogenwatch are provided by , while an in-house search tool is used to rapidly but accurately assign the correct MLST assignment.

Novel allele and MLST codes are indicated by the "*" character at the start of the code. It may also contain letters instead of numbers. Novel loci codes will be consistent between releases, unless replaced by an assignment from the host scheme.

If your profile includes novel alleles or a novel MLST code, we recommend visiting the source database linked in the results page to submit your genome there. Generated assignments will subsequently be imported in Pathogenwatch at the next update.

Method

The assembly is searched for exact matches to known alleles. A representative set of alleles for each locus are then searched for using Blast. These searches are combined and filtered based on the similarity of the match and length of the match. Novel alleles are hashed using the SHA-1 algorithm, this is then used as their unique identifier. Profiles are assigned based on the combination of alleles detected. Novel profiles are also given a unique identifier using the SHA-1 hash algorithm.

Results

For each assembly the assigned allele codes and combined ST code is provided. If a locus is missing, the allele is represented with a question mark, while if it is a novel allele a four letter code that uniquely represents that allele is shown. In the a novel ST due to a combination of alleles is shown as a unique four letter code, while those due to a new allele also have an asterisk ("*") marking them.

How to cite

Please cite the resource which hosts the MLST scheme. The host of the scheme should linked in individual genome reports. Please contact us if you have any questions.

The software is available under an OSS licence from and .

📖
https://github.com/pathogenwatch-oss/mlst
https://github.com/pathogenwatch-oss/typing-databases
PubMLST
Collection View
Collection view with novel allele and ST