This paper presents an updated quantitative and graph-based profile of CompL-it, an open computational lexicon of contemporary Italian represented according to the Linguistic Linked Open Data paradigm and the OntoLex-Lemon model. Rather than focusing on the construction or enrichment of the resource, we examine its current structure as a lexical-semantic infrastructure. The analysis covers lexical entries, inflected forms, lexical senses, semantic relations, semantic types, semantic traits, definitions and examples, with particular attention to the sense-based semantic layer. We combine distributional statistics and coverage indicators with a preliminary graph-structural analysis based on relation classes, reciprocity, connectivity patterns and fragmentation under relation-class ablation. The paper also presents the main access modes to the resource, including data download, SPARQL querying and graphical browsing. By making the internal organisation of CompL-it explicit and measurable, the contribution provides a basis for future work on Italian NLP scenarios that require structured lexical-semantic knowledge.

Exploring CompL-it: Quantitative and Graph-based Profiling of an Open Italian Lexical-Semantic Resource

Emiliano Giovannetti
Primo
;
Andrea Bellandi;Mafalda Papini;Simone Marchi
2026

Abstract

This paper presents an updated quantitative and graph-based profile of CompL-it, an open computational lexicon of contemporary Italian represented according to the Linguistic Linked Open Data paradigm and the OntoLex-Lemon model. Rather than focusing on the construction or enrichment of the resource, we examine its current structure as a lexical-semantic infrastructure. The analysis covers lexical entries, inflected forms, lexical senses, semantic relations, semantic types, semantic traits, definitions and examples, with particular attention to the sense-based semantic layer. We combine distributional statistics and coverage indicators with a preliminary graph-structural analysis based on relation classes, reciprocity, connectivity patterns and fragmentation under relation-class ablation. The paper also presents the main access modes to the resource, including data download, SPARQL querying and graphical browsing. By making the internal organisation of CompL-it explicit and measurable, the contribution provides a basis for future work on Italian NLP scenarios that require structured lexical-semantic knowledge.
2026
Istituto di linguistica computazionale "Antonio Zampolli" - ILC
Computational lexicon, Italian NLP, Linguistic Linked Open Data, OntoLex-Lemon
File in questo prodotto:
Non ci sono file associati a questo prodotto.

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/20.500.14243/594301
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ente

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus ND
  • ???jsp.display-item.citation.isi??? ND
social impact