In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures.

FAIR Training Materials for Disciplinary Research Infrastructures: Metadata, Vocabularies, and Selective Harvesting in the H2IOSC Ecosystem

Daniele Melaccio
;
Giulia Pedonese;Francesca Frontini;
2026

Abstract

In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures.
2026
Istituto di linguistica computazionale "Antonio Zampolli" - ILC
FAIR learning resources, metadata interoperability, CLARIN, H2IOSC, SSHOC, selective harvesting
File in questo prodotto:
Non ci sono file associati a questo prodotto.

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/20.500.14243/596303
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ente

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus ND
  • ???jsp.display-item.citation.isi??? ND
social impact