Describing your dataset with metadata
The information (metadata) you provide when depositing your dataset is essential for making your data FAIR (Findable, Accessible, Interoperable, and Reusable). Well-described datasets are easier to find, understand, cite and reuse. Metadata is also harvested by search engines and data catalogues, increasing the visibility and impact of your research.
Most data repositories, such as DataverseNL, provide you with a metadata form (template) to fill out to describe your dataset. The template in DataverseNL is based on the OpenAire and DataCite metadata schemas, which you may also refer to in your data management plan when asked which metadata standard or /schema you will use.
Essential metadata in DataverseNL
The list below contains the essential elements, also known as rich metadata, that ensure that your data can be found, understood, reused and cited. Each element is accompanied by a short description or practical guidance. Although DataverseNL offers additional optional metadata fields, completing the essential elements as best as possible will significantly improve the quality and discoverability of your dataset.
Essential metadata in DataverseNL (Table)
Essential metadata in DataverseNL
|
Metadata field
|
Subfield
|
Description of the metadata field
|
Additional information
|
|
Title
|
Title with which the dataset will be published.
|
This title does not have to be the same as the title of your publication; after all, this one is about the data.
|
|
|
Author(s)
| |||
|
Name
|
The most common format is Family Name, Initials (or first name).
|
There is no specific convention enforced in DataverseNL. Consistency with other publications is advised.
|
|
|
Affiliation
|
University of Groningen, or University Medical Center Groningen, Name of external organisation.
|
Use of the ROR identifier of organisations is essential to uniquely identify, e.g. the UG or UMCG. This will be checked by the DCC curators.
If the affiliation of external authors is unknown, this field can be left blank.
|
|
|
Identifier Type
|
Identifier of the author(s). ORCID, ISNI or SCOPUS-ID are most often used.
|
ORCID of at least one UG/UMCG author will be required.
|
|
|
Identifier
|
Identifier number.
| ||
|
Point of contact
|
The default is the Groningen Digital Competence Centre (DCC). Other points of contact can be added if needed.
|
Researchers may leave the UG or the UMCG in future, and are therefore not sustainable points of contact. Also note that if visitors want to get in touch, all points of contact receive the messages.
|
|
|
Description
|
Free text that should describe your dataset, not necessarily the research project or the publication
|
The abstract of your publication can be a source of inspiration.
The layout of the text can be improved by using HTML tags. This can be done by the curators, too.
|
|
|
Subject
|
Select from the list all that apply
|
If the tag Social Sciences is included, Odissei will harvest the metadata for its Portal.
|
|
|
Keyword
|
Term
|
Keywords enhance the findability of your dataset. You can add as many as you need.
|
Think of important variables, methods used, and instruments used. Use terms that are well-defined in your field. Keywords defined in controlled vocabularies often have an identifier and a link. If so, please add them.
|
|
Related publication
|
Reference to one or more publications that have this dataset to underpin the findings.
|
If you do not have the full reference at the moment of publication of your dataset, this information can be added later on.
|
|
|
Notes
|
Free text, relevant information that you cannot share in any other metadata field.
| ||
|
Language
|
Select one or more from the list.
|
Tagging the used languages can enhance the reuse of your dataset
|
|
|
Producer
|
Since the affiliation mentioned above is on an organisational level, you can add the name of the research institute or lab of the principal investigator here.
|
||
|
Contributor
|
Give credit to anyone who contributed to your dataset but is not an author.
|
Think of data collectors, data scientists, organisations that supported you (not with a grant), etc.
|
|
|
Funding information
|
Organisations that funded your research, including the ID/number of the grant(s).
| ||
|
Distributor
|
By default: DataverseNL Network
|
This information will be added by the curators of the DCC
|
|
|
Depositor
|
By default: the person who creates the dataset. This is often the curator of the DCC who supports you.
|
For reasons of data governance and integrity, this is automatically generated.
|
|
|
Deposit date
|
By default: the date when the dataset was created as a draft
|
From a user’s point of view not that important, but relevant for safeguarding the integrity and the management of the dataset.
|
|
|
Software
|
The name and version of the software you used to collect or analyse your data.
|
This could also be the software you created. Please check out the information on Research Software Management.
|
|
|
Data source
|
If (part of) your data originates from particular sources, please describe them here.
|
This information can also be part of your Readme file.
|
