A conference theme

Data rescue and historical data

Rescuing at-risk and historical biodiversity data and using it to establish baselines.

7 talks & discussions
Across rooms and sessions

Explore the conversation.

talk · Tuesday 22 September · SAL A

A closer look at steps needed to extract research-ready biodiversity data from legacy literature

Legacy literature data can only be unlocked by combining better OCR/NLP with taxonomic rigour that accounts for synonymy, shifting concepts, coreference and mixed languages.

Gerwin Kasperek ↗
talk · Tuesday 22 September · ODIN

From Pixels to Policy: Closing the Gap between In Situ Monitoring, Earth Observation, and Policy Actionability

Packaging data with its semantics, provenance and machine-readable policy as FAIR Digital Objects makes it rescue-ready and redeployable elsewhere.

Claus Weiland ↗
talk · Tuesday 22 September · ODIN

Supporting resilient data infrastructure in a changing environment

Technical copies are not enough: durable data infrastructure depends on preserved human expertise, messy maintenance work and data standardized close to collection.

Rachael Blake ↗
talk · Thursday 24 September · FORUM

Digitising Plant and Fungal Collections at Royal Botanic Gardens, Kew: Applying International Standards to Support Data Sharing and Geospatial Analysis

WGSRPD enables cross-collection geospatial analysis at Kew, but historical, heterogeneous data leave half of records without fine-level assignment.

Udayangani Liu ↗
talk · Thursday 24 September · FORUM

From preservation to perseverance: bridging disciplines for data resilience

Data resilience in a crisis is about prioritisation and mutual aid rather than perfection, and infrastructures need to be kept appropriately visible to be funded.

Andrea Thomer ↗
talk · Thursday 24 September · SAL C

Mobilising and integrating extracted information from the literature into a biodiversity data ecosystem - the BIOfid approach

BIOfid mobilises Central European biodiversity literature with NLP, ontologies and standards, because raw LLM extraction still hallucinates identifiers and needs verification.

Gerwin Kasperek ↗
talk · Friday 25 September · SAL A

Extracting Structured Biodiversity Knowledge from Literature with AI-Assisted Workflows

A hybrid text-layer/OCR tool can free tables from grey-literature PDFs into CSV with about 83% precision, but it misses a third of tables, especially narrow ones.

Yağmur Güleç ↗