Cargando…
Recovery of metagenomic data from the Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR
Background: Ongoing research of the mosquito microbiome aims to uncover novel strategies to reduce pathogen transmission. Sequencing costs, especially for metagenomics, are however still significant. A resource that is increasingly used to gain insights into host-associated microbiomes is the large...
Autores principales: | , , , |
---|---|
Formato: | Online Artículo Texto |
Lenguaje: | English |
Publicado: |
F1000 Research Limited
2023
|
Materias: | |
Acceso en línea: | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC10412942/ https://www.ncbi.nlm.nih.gov/pubmed/37577055 http://dx.doi.org/10.12688/wellcomeopenres.19155.2 |
_version_ | 1785087027547471872 |
---|---|
author | Foo, Aidan Cerdeira, Louise Hughes, Grant L. Heinz, Eva |
author_facet | Foo, Aidan Cerdeira, Louise Hughes, Grant L. Heinz, Eva |
author_sort | Foo, Aidan |
collection | PubMed |
description | Background: Ongoing research of the mosquito microbiome aims to uncover novel strategies to reduce pathogen transmission. Sequencing costs, especially for metagenomics, are however still significant. A resource that is increasingly used to gain insights into host-associated microbiomes is the large amount of publicly available genomic data based on whole organisms like mosquitoes, which includes sequencing reads of the host-associated microbes and provides the opportunity to gain additional value from these initially host-focused sequencing projects. Methods: To analyse non-host reads from existing genomic data, we developed a snakemake workflow called MINUUR (Microbial INsights Using Unmapped Reads). Within MINUUR, reads derived from the host-associated microbiome were extracted and characterised using taxonomic classifications and metagenome assembly followed by binning and quality assessment. We applied this pipeline to five publicly available Aedes aegypti genomic datasets, consisting of 62 samples with a broad range of sequencing depths. Results: We demonstrate that MINUUR recovers previously identified phyla and genera and is able to extract bacterial metagenome assembled genomes (MAGs) associated to the microbiome. Of these MAGS, 42 are high-quality representatives with >90% completeness and <5% contamination. These MAGs improve the genomic representation of the mosquito microbiome and can be used to facilitate genomic investigation of key genes of interest. Furthermore, we show that samples with a high number of KRAKEN2 assigned reads produce more MAGs. Conclusions: Our metagenomics workflow, MINUUR, was applied to a range of Aedes aegypti genomic samples to characterise microbiome-associated reads. We confirm the presence of key mosquito-associated symbionts that have previously been identified in other studies and recovered high-quality bacterial MAGs. In addition, MINUUR and its associated documentation are freely available on GitHub and provide researchers with a convenient workflow to investigate microbiome data included in the sequencing data for any applicable host genome of interest. |
format | Online Article Text |
id | pubmed-10412942 |
institution | National Center for Biotechnology Information |
language | English |
publishDate | 2023 |
publisher | F1000 Research Limited |
record_format | MEDLINE/PubMed |
spelling | pubmed-104129422023-08-11 Recovery of metagenomic data from the Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR Foo, Aidan Cerdeira, Louise Hughes, Grant L. Heinz, Eva Wellcome Open Res Method Article Background: Ongoing research of the mosquito microbiome aims to uncover novel strategies to reduce pathogen transmission. Sequencing costs, especially for metagenomics, are however still significant. A resource that is increasingly used to gain insights into host-associated microbiomes is the large amount of publicly available genomic data based on whole organisms like mosquitoes, which includes sequencing reads of the host-associated microbes and provides the opportunity to gain additional value from these initially host-focused sequencing projects. Methods: To analyse non-host reads from existing genomic data, we developed a snakemake workflow called MINUUR (Microbial INsights Using Unmapped Reads). Within MINUUR, reads derived from the host-associated microbiome were extracted and characterised using taxonomic classifications and metagenome assembly followed by binning and quality assessment. We applied this pipeline to five publicly available Aedes aegypti genomic datasets, consisting of 62 samples with a broad range of sequencing depths. Results: We demonstrate that MINUUR recovers previously identified phyla and genera and is able to extract bacterial metagenome assembled genomes (MAGs) associated to the microbiome. Of these MAGS, 42 are high-quality representatives with >90% completeness and <5% contamination. These MAGs improve the genomic representation of the mosquito microbiome and can be used to facilitate genomic investigation of key genes of interest. Furthermore, we show that samples with a high number of KRAKEN2 assigned reads produce more MAGs. Conclusions: Our metagenomics workflow, MINUUR, was applied to a range of Aedes aegypti genomic samples to characterise microbiome-associated reads. We confirm the presence of key mosquito-associated symbionts that have previously been identified in other studies and recovered high-quality bacterial MAGs. In addition, MINUUR and its associated documentation are freely available on GitHub and provide researchers with a convenient workflow to investigate microbiome data included in the sequencing data for any applicable host genome of interest. F1000 Research Limited 2023-05-26 /pmc/articles/PMC10412942/ /pubmed/37577055 http://dx.doi.org/10.12688/wellcomeopenres.19155.2 Text en Copyright: © 2023 Foo A et al. https://creativecommons.org/licenses/by/4.0/This is an open access article distributed under the terms of the Creative Commons Attribution Licence, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited. |
spellingShingle | Method Article Foo, Aidan Cerdeira, Louise Hughes, Grant L. Heinz, Eva Recovery of metagenomic data from the Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR |
title | Recovery of metagenomic data from the
Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR |
title_full | Recovery of metagenomic data from the
Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR |
title_fullStr | Recovery of metagenomic data from the
Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR |
title_full_unstemmed | Recovery of metagenomic data from the
Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR |
title_short | Recovery of metagenomic data from the
Aedes aegypti microbiome using a reproducible snakemake pipeline: MINUUR |
title_sort | recovery of metagenomic data from the
aedes aegypti microbiome using a reproducible snakemake pipeline: minuur |
topic | Method Article |
url | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC10412942/ https://www.ncbi.nlm.nih.gov/pubmed/37577055 http://dx.doi.org/10.12688/wellcomeopenres.19155.2 |
work_keys_str_mv | AT fooaidan recoveryofmetagenomicdatafromtheaedesaegyptimicrobiomeusingareproduciblesnakemakepipelineminuur AT cerdeiralouise recoveryofmetagenomicdatafromtheaedesaegyptimicrobiomeusingareproduciblesnakemakepipelineminuur AT hughesgrantl recoveryofmetagenomicdatafromtheaedesaegyptimicrobiomeusingareproduciblesnakemakepipelineminuur AT heinzeva recoveryofmetagenomicdatafromtheaedesaegyptimicrobiomeusingareproduciblesnakemakepipelineminuur |