In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures.

FAIR Training Materials for Disciplinary Research Infrastructures: Metadata, Vocabularies, and Selective Harvesting in the H2IOSC Ecosystem

Daniele Melaccio
;
Giulia Pedonese;Francesca Frontini;
2026

Abstract

In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures.
Campo DC Valore Lingua
dc.authority.orgunit Istituto di linguistica computazionale "Antonio Zampolli" - ILC en
dc.authority.people Daniele Melaccio en
dc.authority.people Giulia Pedonese en
dc.authority.people Francesca Frontini en
dc.authority.people Iulianna van der Lek en
dc.authority.people Thalassia Kontino en
dc.collection.id.s 71c7200a-7c5f-4e83-8d57-d3d2ba88f40d *
dc.collection.name 04.01 Contributo in Atti di convegno *
dc.contributor.appartenenza ASR - Unità Internal Audit *
dc.contributor.appartenenza Istituto di linguistica computazionale "Antonio Zampolli" - ILC *
dc.contributor.appartenenza.mi 918 *
dc.contributor.appartenenza.mi 1177 *
dc.contributor.area Non assegn *
dc.contributor.area Non assegn *
dc.contributor.area Non assegn *
dc.date.firstsubmission 2026/08/28 17:39:46 *
dc.date.issued 2026 -
dc.date.submission 2026/08/28 17:39:46 *
dc.description.abstracteng In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures. -
dc.description.allpeople Melaccio, Daniele; Pedonese, Giulia; Frontini, Francesca; Van Der Lek, Iulianna; Kontino, Thalassia -
dc.description.allpeopleoriginal Daniele Melaccio, Giulia Pedonese, Francesca Frontini, Iulianna van der Lek, Thalassia Kontino en
dc.description.fulltext none en
dc.description.international si en
dc.description.numberofauthors 5 -
dc.identifier.source manual *
dc.identifier.uri https://hdl.handle.net/20.500.14243/596303 -
dc.language.iso eng en
dc.relation.allauthors Daniele Melaccio, Giulia Pedonese, Francesca Frontini, Iulianna van der Lek, Thalassia Kontino en
dc.relation.conferencedate 19-20 febbraio 2026 en
dc.relation.conferencename Information and Research Science Connecting to Digital and Library Science 2026 en
dc.relation.conferenceplace Modena, Italia en
dc.relation.ispartofbook Proceedings of the 22nd Conference on Information and Research Science Connecting to Digital and Library Science en
dc.subject.keywordseng FAIR learning resources, metadata interoperability, CLARIN, H2IOSC, SSHOC, selective harvesting -
dc.subject.singlekeyword FAIR learning resources *
dc.subject.singlekeyword metadata interoperability *
dc.subject.singlekeyword CLARIN *
dc.subject.singlekeyword H2IOSC *
dc.subject.singlekeyword SSHOC *
dc.subject.singlekeyword selective harvesting *
dc.title FAIR Training Materials for Disciplinary Research Infrastructures: Metadata, Vocabularies, and Selective Harvesting in the H2IOSC Ecosystem en
dc.type.circulation Internazionale en
dc.type.driver info:eu-repo/semantics/conferenceObject -
dc.type.full 04 Contributo in convegno::04.01 Contributo in Atti di convegno it
dc.type.invited contributo en
dc.type.miur 273 -
iris.orcid.lastModifiedDate 2026/08/28 17:39:46 *
iris.orcid.lastModifiedMillisecond 1787931586459 *
iris.sitodocente.maxattempts 1 -
File in questo prodotto:
Non ci sono file associati a questo prodotto.

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/20.500.14243/596303
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ente

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus ND
  • ???jsp.display-item.citation.isi??? ND
social impact