talk · Friday 25 September · SAL C

Approaches to the Automated Tracking of Taxonomic Research Activities

David Fichtmueller · Publishing and Communications Promoting Biodiversity, Natural History Collections, People, Data and Data Standards

Recording time 51:36–1:06:29Open on Vimeo ↗

The short versionTETTRIX aims to build a machine-run, regularly updated picture of who is doing taxonomy on what and where, and is asking the community which sources it is missing.

Overview

What this was about

David Fichtmueller introduced work package 1 of the newly started EU project TETTRIX, which will automatically gather ongoing taxonomic research activities from many sources into a central database (TaxoTrack), align them with taxonomic backbones and identifiers, and produce outputs such as maps, gap analyses, dashboards and policy briefs. He described how it builds on TETTRIs expertise assessments from publications (OpenAlex, Meise) and GBIF specimen data (Berlin), laid out the partner-led source streams, and asked the audience how their research could be detected. Feedback stressed ORCIDs on publications, registering nomenclatural acts, and institutional evaluation reports as sources.

Why it matters. A systematic, repeatable overview of taxonomic research activity could expose understudied taxa and regions and inform funders and policy makers, but it depends on the research being machine-detectable through identifiers such as ORCID.

Key ideas

In the room

  • TETTRIX (successor of TETTRIs) started in September 2026; work package 1 will automatically gather ongoing taxonomic research activities, store them centrally and align them with taxonomic backbones.
  • Earlier TETTRIs work assessed expertise from publications filtered in OpenAlex (Meise) and from GBIF specimen data (Berlin); the specimen approach was harder because person data in GBIF are ambiguous and rarely carry identifiers.
  • Planned source streams: taxonomic publication databases incl. sub-article data (Pensoft: BLR, Biodiversity PMC, OpenBiodiv), funding records and project registries (UFZ Leipzig), taxonomic data networks and name registries (Berlin, Meise, Prague with a Central/Eastern European focus), 'dark' pre-naming activity via DNA barcodes, BINs and species hypotheses (NTNU), and expertise/service marketplaces.
  • All sources feed a central database, TaxoTrack, aligned with identifiers such as Wikidata, ROR and ORCID, to generate maps, taxonomic coverage evaluations, policy briefs and dashboards.
  • The team does not expect researchers to change workflows; it wants machine-accessible sources so the analysis can be re-run every few weeks or months.
  • A survey was launched targeting funding agencies and project registries, with an open section for other source ideas.
Jump in

Notable moments

In their words

Transcript

Automatically generated captions can contain mistakes, especially in names and technical terms. Times are relative to the room recording.

Read the transcript ↓
Loading transcript…