eDNAqua-Plan Data Management Plan: DMP template, list of datasets, metadata standard vocabularies, repositories and file names for eDNA Studies
Authors/Creators
Description
We here provide a template for a data management plan specifically tailored to projects that generate eDNA datasets from aquatic environments. It may be challenging at the very start of a project to determine which metadata standard vocabulary and repository will be used, and this template aims to inform the researcher of the available options. The DMP template follows the digital landscape structure of eDNAqua-Plan (https://doi.org/10.5281/zenodo.16367374) and outlines the workflow steps for eDNA studies, including sampling, laboratory procedures, bioinformatics analyses, and taxonomic assignment, while specifying the applicable methodologies (metabarcoding, metagenomics, targeted assays, and morphology). The DMP aims to guide users on which datasets can be generated, which file naming should be used, where and how the data should be handled, including information on checklists to be used for documentation and for improved FAIRness of the datasets.
The DMP template is available in DMPonline (for Belgian users). For research organizations outside Belgium, the template is provided below as a Word file and a PDF file. By answering the questions and saving the Word file, the DMP can be uploaded to the management system of the research institute. The “Table_datasets_eDNA_studies_V3” file lists the most common options for repositories and standard vocabularies.
Key features include:
- Standardized file naming conventions (aligned with FAIR guidelines).
- Recommended metadata standards (e.g., MIxS, DwC, ENVO) and reproducibility management systems (e.g., NCBI BioProject, Zenodo, GitHub).
- File extension types for data files (e.g., CSV, FASTQ, FASTA).
- Guidance on data types, from project and sample metadata to raw/processed sequencing data, reference libraries, and bioinformatics code.
The document serves as a practical reference for researchers, ensuring consistency, interoperability, and reproducibility in eDNA data management and sharing.