This paper presents an updated quantitative and graph-based profile of CompL-it, an open computational lexicon of contemporary Italian represented according to the Linguistic Linked Open Data paradigm and the OntoLex-Lemon model. Rather than focusing on the construction or enrichment of the resource, we examine its current structure as a lexical-semantic infrastructure. The analysis covers lexical entries, inflected forms, lexical senses, semantic relations, semantic types, semantic traits, definitions and examples, with particular attention to the sense-based semantic layer. We combine distributional statistics and coverage indicators with a preliminary graph-structural analysis based on relation classes, reciprocity, connectivity patterns and fragmentation under relation-class ablation. The paper also presents the main access modes to the resource, including data download, SPARQL querying and graphical browsing. By making the internal organisation of CompL-it explicit and measurable, the contribution provides a basis for future work on Italian NLP scenarios that require structured lexical-semantic knowledge.
Exploring CompL-it: Quantitative and Graph-based Profiling of an Open Italian Lexical-Semantic Resource
Emiliano Giovannetti
Primo
;Andrea Bellandi;Mafalda Papini;Simone Marchi
2026
Abstract
This paper presents an updated quantitative and graph-based profile of CompL-it, an open computational lexicon of contemporary Italian represented according to the Linguistic Linked Open Data paradigm and the OntoLex-Lemon model. Rather than focusing on the construction or enrichment of the resource, we examine its current structure as a lexical-semantic infrastructure. The analysis covers lexical entries, inflected forms, lexical senses, semantic relations, semantic types, semantic traits, definitions and examples, with particular attention to the sense-based semantic layer. We combine distributional statistics and coverage indicators with a preliminary graph-structural analysis based on relation classes, reciprocity, connectivity patterns and fragmentation under relation-class ablation. The paper also presents the main access modes to the resource, including data download, SPARQL querying and graphical browsing. By making the internal organisation of CompL-it explicit and measurable, the contribution provides a basis for future work on Italian NLP scenarios that require structured lexical-semantic knowledge.I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.


