PIDfest 26

Boost Metadata Quality with the DataCite Metadata Generator
2026-10-28 , LUMC04

Do you struggle with creating complete, high-quality DataCite metadata or reusing existing records efficiently? In this session, you will discover the updated DataCite Metadata Generator developed by LMU Munich, designed to simplify and enhance your metadata workflows.

You will learn how the tool enables you to create, upload, validate, and improve metadata in both XML and JSON formats. It also enriches metadata quality by automatically integrating PIDs via APIs such as ORCID and ROR.

You will also see how the generator is designed to be easily updated when new DataCite schema versions are released, ensuring long-term usability. You will explore how this open-source solution can be integrated into your own institutional workflows and adapted to specific needs. By the end of the session, you will better understand how to streamline metadata creation while improving consistency and reuse.


Do you want to create high-quality DataCite metadata without struggling through complex forms or incomplete records? Are you looking for ways to reuse and improve existing metadata efficiently?

In this session, you will explore the updated version of the DataCite Metadata Generator developed at LMU Munich (University Library and IT-Gruppe Geisteswissenschaften). The tool is designed to simplify metadata creation while improving quality and interoperability across systems.

You will discover how the web-based form helps you generate structured DataCite metadata that can be directly reused in publication workflows or used to enhance existing records. The new version goes beyond simple form-based input: you can upload existing metadata files in XML or JSON format, receive feedback on completeness, edit them within the tool, and download improved versions.

You will also learn how the generator automatically enriches metadata by connecting to external PID systems. By integrating APIs such as ORCID and ROR, the tool retrieves additional information linked to PIDs, helping you improve consistency and reduce manual input.

Another key feature is its future-proof design: the generator is built so that it can be easily updated when new DataCite schema versions are released. This ensures that your workflows remain compatible with evolving standards without requiring major redevelopment.

In addition, you will hear how the tool supports flexible adoption in different institutional contexts. The source code is openly available on GitHub, allowing you to integrate the generator into your own workflows and adapt it to your needs, e.g., by extending the form with institution-specific fields.

After this session, you will:
• understand how to create, validate, and reuse DataCite metadata more efficiently
• know how to improve metadata quality through PID integration and automated enrichment
• gain insights into how you can adapt and integrate the tool into your own infrastructure

This session is aimed at practitioners, repository managers, and anyone working with research metadata who wants practical solutions to improve metadata workflows and quality.

I am part of the RDM information team at the University Library of LMU Munich, where I support researchers across all disciplines with services and infrastructure for managing their research data. My background combines library and information management with computer science, allowing me to work at the intersection of metadata, infrastructure, and user-oriented services.