De novo identification of microbial contaminants in low microbial biomass microbiomes with Squeegee

Abstract

Computational analysis of host-associated microbiomes has opened the door to numerous discoveries relevant to human health and disease. However, contaminant sequences in metagenomic samples can potentially impact the interpretation of findings reported in microbiome studies, especially in low-biomass environments. Contamination from DNA extraction kits or sampling lab environments leaves taxonomic "bread crumbs" across multiple distinct sample types. Here we describe Squeegee, a de novo contamination detection tool that is based upon this principle, allowing the detection of microbial contaminants when negative controls are unavailable. On the low-biomass samples, we compare Squeegee predictions to experimental negative control data and show that Squeegee accurately recovers putative contaminants. We analyze samples of varying biomass from the Human Microbiome Project and identify likely, previously unreported kit contamination. Collectively, our results highlight that Squeegee can identify microbial contaminants with high precision and thus represents a computational approach for contaminant detection when negative controls are unavailable.

Document Details

Document Type
Pub Defense Publication
Publication Date
Nov 10, 2022
Source ID
10.1038/s41467-022-34409-z

Entities

People

  • Kjersti M Aagaard
  • Michael D. Jochum
  • R A Leo Elworth
  • Todd J Treangen
  • Yunxi Liu

Organizations

  • Intelligence Advanced Research Projects Activity
  • National Institute of Allergy and Infectious Diseases

Tags

Fields of Study

  • Biology
  • Environmental science

Readers

  • Computer Vision.
  • Gulf War Illness and Chronic Multisymptom Illness in Veterans.
  • Microbial Pathology

Technology Areas

  • Biotechnology
  • Biotechnology - Bioremediation