In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures.
FAIR Training Materials for Disciplinary Research Infrastructures: Metadata, Vocabularies, and Selective Harvesting in the H2IOSC Ecosystem
Daniele Melaccio
;Giulia Pedonese;Francesca Frontini;
2026
Abstract
In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures.| Campo DC | Valore | Lingua |
|---|---|---|
| dc.authority.orgunit | Istituto di linguistica computazionale "Antonio Zampolli" - ILC | en |
| dc.authority.people | Daniele Melaccio | en |
| dc.authority.people | Giulia Pedonese | en |
| dc.authority.people | Francesca Frontini | en |
| dc.authority.people | Iulianna van der Lek | en |
| dc.authority.people | Thalassia Kontino | en |
| dc.collection.id.s | 71c7200a-7c5f-4e83-8d57-d3d2ba88f40d | * |
| dc.collection.name | 04.01 Contributo in Atti di convegno | * |
| dc.contributor.appartenenza | ASR - Unità Internal Audit | * |
| dc.contributor.appartenenza | Istituto di linguistica computazionale "Antonio Zampolli" - ILC | * |
| dc.contributor.appartenenza.mi | 918 | * |
| dc.contributor.appartenenza.mi | 1177 | * |
| dc.contributor.area | Non assegn | * |
| dc.contributor.area | Non assegn | * |
| dc.contributor.area | Non assegn | * |
| dc.date.firstsubmission | 2026/08/28 17:39:46 | * |
| dc.date.issued | 2026 | - |
| dc.date.submission | 2026/08/28 17:39:46 | * |
| dc.description.abstracteng | In a lifelong learning society and especially now in the era of Artificial Intelligence, high-quality, flexible learning resources are essential for acquiring the new skills required by the ever-evolving job market. Research infrastructures play a central role in providing users with specialised training and advanced digital tools, fostering innovation in their fields. CLARIN ERIC, the Common Language Resources and Technology Infrastructure, has consistently promoted training initiatives in linguistic and language technologies, especially through the CLARIN Learning Hub. The Italian national consortium CLARIN-IT has developed this approach within the NRRP project H2IOSC, aiming to create a cluster of research infrastructures in the field of Social Sciences and Humanities. In this context, CLARIN-IT led the work package dedicated to training and developed two platforms: an e-learning portal and a digital library for long-term deposit and curation of training materials as FAIR digital objects. The Skills4EOSC FAIR-by-Design methodology and the SSHOC vocabularies adopted in the H2IOSC project ensured metadata compatibility and interoperability of learning resources across domains. Moreover, the adopted standards will enable harvesting resources in the H2IOSC Marketplace and selective harvesting for domain-specific platforms such as the CLARIN Learning Resource Catalogue. This paper proposes an architecture for metadata mapping and harvesting of training resources to maximise their reusability across infrastructures. | - |
| dc.description.allpeople | Melaccio, Daniele; Pedonese, Giulia; Frontini, Francesca; Van Der Lek, Iulianna; Kontino, Thalassia | - |
| dc.description.allpeopleoriginal | Daniele Melaccio, Giulia Pedonese, Francesca Frontini, Iulianna van der Lek, Thalassia Kontino | en |
| dc.description.fulltext | none | en |
| dc.description.international | si | en |
| dc.description.numberofauthors | 5 | - |
| dc.identifier.source | manual | * |
| dc.identifier.uri | https://hdl.handle.net/20.500.14243/596303 | - |
| dc.language.iso | eng | en |
| dc.relation.allauthors | Daniele Melaccio, Giulia Pedonese, Francesca Frontini, Iulianna van der Lek, Thalassia Kontino | en |
| dc.relation.conferencedate | 19-20 febbraio 2026 | en |
| dc.relation.conferencename | Information and Research Science Connecting to Digital and Library Science 2026 | en |
| dc.relation.conferenceplace | Modena, Italia | en |
| dc.relation.ispartofbook | Proceedings of the 22nd Conference on Information and Research Science Connecting to Digital and Library Science | en |
| dc.subject.keywordseng | FAIR learning resources, metadata interoperability, CLARIN, H2IOSC, SSHOC, selective harvesting | - |
| dc.subject.singlekeyword | FAIR learning resources | * |
| dc.subject.singlekeyword | metadata interoperability | * |
| dc.subject.singlekeyword | CLARIN | * |
| dc.subject.singlekeyword | H2IOSC | * |
| dc.subject.singlekeyword | SSHOC | * |
| dc.subject.singlekeyword | selective harvesting | * |
| dc.title | FAIR Training Materials for Disciplinary Research Infrastructures: Metadata, Vocabularies, and Selective Harvesting in the H2IOSC Ecosystem | en |
| dc.type.circulation | Internazionale | en |
| dc.type.driver | info:eu-repo/semantics/conferenceObject | - |
| dc.type.full | 04 Contributo in convegno::04.01 Contributo in Atti di convegno | it |
| dc.type.invited | contributo | en |
| dc.type.miur | 273 | - |
| iris.orcid.lastModifiedDate | 2026/08/28 17:39:46 | * |
| iris.orcid.lastModifiedMillisecond | 1787931586459 | * |
| iris.sitodocente.maxattempts | 1 | - |
I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.


