It’s difficult. I think real change will come with the next generation. If working with data in this way becomes standard practice for students, it will become a natural part of research. But even today, there are many people in the MATECH community, which focuses on materials science, who want to get involved.
We don’t want to dictate to researchers how they should work with their data, so we are finding out what they actually need. At the same time, we are trying to make storing data as simple as possible, for example by connecting the repository to electronic lab notebooks or enabling data to be uploaded directly from instruments via an API.
What are the main steps involved in building the repository?
A lot of it is about communicating with researchers. We need to know how they will want to search for data, what information matters to them, and which features they can do without.
A good metadata model is fundamental. Michal Med, the author of the Czech Core Metadata Model for Research Data (CCMM), compares it to a bottle of beer: the data are the beer, while the metadata are the label telling you what’s inside. Without the label, you don’t really know what’s in the bottle.
Then there are the technical and legal aspects. We are building the repository on the Invenio platform and developing the rules for how it will operate and connect to other services. We are trying to set everything up so that storing and finding data is as easy as possible for researchers.
What is your role in all of this?