Aller au contenu principal
Accès ouvert déclaré 2026 article

SCIENCE BFSP CoS | WP6 | D6.2 Biodiv FAIR metrics implementation

0Citations signalées, ce qui n’est pas une note de qualité
6Institutions déclarées
3Pays d’affiliation déclarés

Rattachement africain : ch, fr, gb. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

The biodiversity community has already done impressive work registering vast numbers of sample records in large reference data repositories such as the European Nucleotide Archive (ENA) and BioSamples. Yet the metadata attached to these records is too often incomplete: essential fields are left blank, filled with placeholders like ‘missing’, or the information is simply never captured in the first place. These gaps are exactly what make otherwise valuable records hard to find, trust, and reuse. As part of the BioDiv-FC project which extends the open-source FAIR-Checker tool with domain-aware, community-defined FAIR metrics for biodiversity data this deliverable reports a profile-based method to quantify the metadata completeness of ENA/BioSamples records against a community checklist. We take the ENA Tree of Life checklist (ERC000053, "ENA53"), express it as a machine-actionable profile partitioned into mandatory, recommended and optional elements, and compile that profile into a generic Shapes Constraint Language (SHACL) shape. Each sample is harvested as schema.org Resource Description Framework (RDF) and validated against the shape, yielding results at two complementary levels: (i) a single weighted completeness score (0–100 %) that summarises the record at a glance, and (ii) a detailed report listing precisely which mandatory, recommended, and optional properties are missing. The score supports monitoring metadata quality across large sample collections, while the detailed report turns each evaluation into an actionable recommendation for progressively improving a record's metadata completeness. The approach is checklist-agnostic: it is by no means limited to the ENA checklists but applies to any other community profile, since retargeting the evaluation only requires swapping the three element lists.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

La source scientifique ouverte est momentanément indisponible.

Où se fait cette recherche

  • SIB Swiss Institute of Bioinformatics pays non établi dans la notice
    Organisation à but non lucratif
  • Centre National de la Recherche Scientifique pays non établi dans la notice
    Organisme public
  • Institut National de Recherche pour l'Agriculture pays non établi dans la notice
    Organisme public
  • Services déconcentrés d'appui à la recherche Occitanie-Montpellier pays non établi dans la notice
    Structure de recherche
  • University of Manchester pays non établi dans la notice
    Université ou école supérieure
  • Inserm pays non établi dans la notice
    Organisme public
  • INRAe pays non établi dans la notice
    Institution
  • INRA Centre de Montpellier pays non établi dans la notice
    Institution
  • French National Center for Scientific Research (head office) pays non établi dans la notice
    Institution

SIB Swiss Institute of Bioinformatics, Centre National de la Recherche Scientifique et Institut National de Recherche pour l'Agriculture, avec 6 autres affiliations.

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Research Data Management PracticesSpecies Distribution and Climate ChangeScientific Computing and Data Management

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.