Editorial: Marine microbiomes: towards standard methods and Best Practices
Bibliographic record
Abstract
The microbiome is key to understanding and sustaining the services that ocean ecosystems provide (AORA 2020). The marine microbiome -an ensemble of microscopic organisms that inhabit water columns, sediments, and aquatic organisms -contains members spanning in size from viruses of a few tens of nanometers to metazoans of several centimeters. Together, the microbiome forms the base of the food web, maintains animal health, and regulates most fluxes of energy and matter. Marine microbiome discovery is part of a great campaign to explore the earth's oceans, and rapid advances in high throughput sequencing are allowing a glimpse into this hidden world. Furthermore, these techniques have been adapted to detect DNA in the environment (eDNA) from organisms of all trophic levels.Biomolecular observations can provide important insights into ecosystem structure and function, development of new indicators of ecosystem health, and warnings of potential hazards to living resources and humans. Endorsements by the United Nations Ocean Decade 1 reflect the growing demand for affordable, largescale biological observations provided by biomolecule detection. Examples include the Ocean Biomolecular Observing Network (OBON) Program (Leinen et al. 2022), which aims to transform how we sense, harvest, protect and manage ocean life. OBON actions supporting these aims include the Observing and Promoting Atlantic Microbiomes 2 project hosted by the Atlantic Ocean Research Alliance (AORA) Marine Microbiome Working Group 3 that called for this special issue. This Frontiers Research Topic was motivated by the recognition that a number of crosscutting challenges need to be addressed to fully unlock the marine microbiome for environmental and societal benefit. Such challenges include the development and adoption of standards, common methods, Best Practices, and FAIR (Findable, Accessible, Interoperable, and 1 https://oceandecade.org/ 2 https://oceandecade.org/actions/ocean-biomolecular-observing-network-obon/ 3 https://www.marinemicrobiome.org/The marine microbiome is a largely unexplored treasure for society. Illustration credit: Rán Flygenring.Re-usable) data principles (AORA 2020). For some authors, the emphasis was on cyberinfrastructure to ensure that both sequence and environmental data are FAIR (Blumberg et al. 2021). Others focused on developing a Minimum Information for an Omic Protocol (MIOP) and a public repository of protocols that can be both searched and prioritized for use (Samuel et al. 2021). Both of these manuscripts highlighted the importance of machine readable data and products to the achievement of FAIR principles. The need to implement and sustain a global and publicly supported platform to share, discover, and compare practices and protocols was emphasized.Papers in this special issue highlighted that harmonization across the full workflow -from methods through data reporting -is needed to achieve global scale biodiversity observations that can be integrated over space and time. Some manuscripts offered general overviews and "tricks of the trade" to guide microbiome sample collection and processing for coral tissues (Silva et al. 2023) or pelagic waters for a variety of molecular targets and size fractions (Patin and Goodwin 2023). These papers reviewed methods for multiple sections of the overall workflow with detailed guidance provided for sample collection, preservation, and processing. Other manuscripts focused on specific details, such as DNA isolation. For example, Wietz et al. ( 2022) described extraction of DNA from samples preserved in formalin or HgCl2, preservatives commonly used in sediment trap studies. Korlević et al. ( 2021) described a procedure to specifically isolate DNA and protein from macrophyte epiphytic communities to avoid overwhelming microbiome samples with host DNA. Gu et al. ( 2022) described a new analytical protocol to determine Protoporphyrin IX (PPIX) in microbial cells and provided results with coastal aquatic samples to demonstrate the potential to use PPIX as an indicator of microbial productivity. This diversity of topics underscores the large range of microbiome applications.The growth of publicly available sequence data has increased the ability to perform metaanalysis to investigate broad scale environmental change. However, the rapid expansion of molecular techniques has created disparate protocols and workflows. A number of authors thus addressed the question of whether datasets can be combined across studies by exploring the sensitivity of taxonomic annotation to variations in sequencing methods. For example, taxonomic assignments were compared for the 16S rRNA gene V3-V4 and V4-V5 primer sets as applied to a variety of sample types collected from Arctic Ocean marine systems. In this case, V4-V5 was recommended due to superior inclusion of archaeal taxa (Fadeev et al. 2021). In another case, a single primer set was applied to coral tissues that were processed separately (DNA extraction through library preparation) and then sequenced on different platforms (MiSeq and HiSeq). Despite past studies suggesting that MiSeq and HiSeq data could be combined to provide microbiome taxonomic analysis, the study here cautioned that significant differences in compositional assignments could arise from protocol variations (Epstein et al. 2021). This work suggested that projects that seek to understand and overcome sources of technical variation remain needed. Multiple studies also highlighted the continued need to build out reference databases to improve annotation of sequence data. This Research Topic fostered cross-community exchange of standards and Best Practices. It provided an opportunity for different communities working on marine microbiomes to communicate the advantages and limitations of various sampling, laboratory, and data processing methods and to open community discussion on how to move towards large scale operationalization. Although the need for harmonization was recognized, workflows must be fit for purpose; meaning that they must meet the objectives and logistical constraints of a study, application, management objective, or time series (including legacy sampling). Even with Best Practices in hand, pilot studies must be conducted to validate and optimize workflows against different sample types, geographies, or molecular targets. To support international efforts to develop guidance on Best Practices, a centralized platform could compile protocols, metadata, and data produced by specific methods. Such could include successful, unsuccessful, anecdotal, or unpublished information to provide real-world feedback. Machine Learning approaches could potentially help define optimal workflows from the growing observations. Overall, the pursuit of cross-community standards and Best Practices will foster data integration across heterogeneous methods, improve future ocean observations, and expand the trusted use of microbiome and eDNA science.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.016 | 0.061 |
| Meta-epidemiology (narrow) | 0.005 | 0.002 |
| Meta-epidemiology (broad) | 0.004 | 0.004 |
| Bibliometrics | 0.004 | 0.002 |
| Science and technology studies | 0.004 | 0.005 |
| Scholarly communication | 0.011 | 0.008 |
| Open science | 0.005 | 0.003 |
| Research integrity | 0.018 | 0.024 |
| Insufficient payload (model declined to judge) | 0.021 | 0.022 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".