Optimizing a Whole-Genome Sequencing Data Processing Pipeline for Precision Surveillance of Health Care-Associated Infections

DOI

10.3390/microorganisms7100388

Journal Title

Microorganisms

First Page

E388

Document Type

Article

Publication Date

9-1-2019

Department

Pathology, Microbiology and Immunology

Keywords

data processing pipeline, genomic surveillance, health care-associated infection (HAI), k-mer, whole-genome sequencing (WGS)

Disciplines

Medicine and Health Sciences

Abstract

The surveillance of health care-associated infection (HAI) is an essential element of the infection control program. While whole-genome sequencing (WGS) has widely been adopted for genomic surveillance, its data processing remains to be improved. Here, we propose a three-level data processing pipeline for the precision genomic surveillance of microorganisms without prior knowledge: species identification, multi-locus sequence typing (MLST), and sub-MLST clustering. The former two are closely connected to what have widely been used in current clinical microbiology laboratories, whereas the latter one provides significantly improved resolution and accuracy in genomic surveillance. Comparing to a broadly used reference-dependent alignment/mapping method and an annotation-dependent pan-/core-genome analysis, we implemented our reference- and annotation-independent, k-mer-based, simplified workflow to a collection of Acinetobacter and Enterococcus clinical isolates for tests. By taking both single nucleotide variants and genomic structural changes into account, the optimized k-mer-based pipeline demonstrated a global view of bacterial population structure in a rapid manner and discriminated the relatedness between bacterial isolates in more detail and precision. The newly developed WGS data processing pipeline would facilitate WGS application to the precision genomic surveillance of HAI. In addition, the results from such a WGS-based analysis would be useful for the precision laboratory diagnosis of infectious microorganisms.

Recommended Citation

Huang, W., Wang, G., Yin, C., Chen, D., Dhand, A., Chanza, M., Dimitrova, N., & Fallon, J. (2019). Optimizing a Whole-Genome Sequencing Data Processing Pipeline for Precision Surveillance of Health Care-Associated Infections. Microorganisms, 7 (10), E388. https://doi.org/10.3390/microorganisms7100388

NYMC Faculty Publications

Optimizing a Whole-Genome Sequencing Data Processing Pipeline for Precision Surveillance of Health Care-Associated Infections

DOI

Journal Title

First Page

Document Type

Publication Date

Department

Keywords

Disciplines

Abstract

Recommended Citation

Search

Browse

Author Corner

NYMC Faculty Publications

Optimizing a Whole-Genome Sequencing Data Processing Pipeline for Precision Surveillance of Health Care-Associated Infections

Authors

DOI

Journal Title

First Page

Document Type

Publication Date

Department

Keywords

Disciplines

Abstract

Recommended Citation

Share

Search

Browse

Author Corner