<?xml version="1.0" encoding="UTF-8"?><?xml-stylesheet type="text/xsl" href="static/CINECAstyle.xsl"?><OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd"><responseDate>2026-08-17T17:03:29Z</responseDate><request verb="GetRecord" identifier="oai:iris.cnr.it:20.500.14243/570801" metadataPrefix="oai_dc">https://iris.cnr.it/oai/request</request><GetRecord><record><header><identifier>oai:iris.cnr.it:20.500.14243/570801</identifier><datestamp>2026-03-03T18:40:13Z</datestamp><setSpec>ou_ou239</setSpec></header><metadata><oai_dc:dc xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xmlns:doc="http://www.lyncode.com/xoai" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
<dc:title>Generating and Evaluating Multi-Level Text Simplification: A Case Study on Italian</dc:title>
<dc:creator>Michele Papucci</dc:creator>
<dc:creator>Giulia Venturi</dc:creator>
<dc:creator>Felice Dell'Orletta</dc:creator>
<dc:contributor>Michele Papucci, Giulia Venturi, Felice Dell’Orletta</dc:contributor>
<dc:contributor>Papucci, Michele</dc:contributor>
<dc:contributor> Venturi, Giulia</dc:contributor>
<dc:contributor> Dell'Orletta, Felice</dc:contributor>
<dc:subject>Automatic Text Simplification, Large Language Models, Synthetic Data, Linguistic Complexity, Sentence Readability</dc:subject>
<dc:description>Recent advances in Generative AI and Large Language Models (LLMs) have enabled the creation of highly realistic synthetic content, yet controlling model outputs remains a challenge. In this study, we explore the use of LLMs to generate high-quality synthetic data for Automatic Text Simplification (ATS), evaluating the ability of models fine-tuned on Italian to produce multiple simplified versions of the same original sentence that vary in readability and in their lexical and (morpho-)syntactic characteristics. The approach is tested across two domains, Wikipedia and Public Administration, allowing us to explore domain sensitivity. Additionally, we compare the linguistic phenomena observed in the generated data with those found in ATS resources previously created through manual or semi-automatic methods. Our results suggest that the best-performing LLM can generate linguistically diverse simplifications that align with known simplification patterns, offering a promising direction for building reliable ATS resources, including simplifications suited to varying levels of reader proficiency.</dc:description>
<dc:date>2025</dc:date>
<dc:type>info:eu-repo/semantics/conferenceObject</dc:type>
<dc:identifier>https://hdl.handle.net/20.500.14243/570801</dc:identifier>
<dc:relation>info:eu-repo/semantics/altIdentifier/isbn/979-12-243-0587-3</dc:relation>
<dc:identifier>https://aclanthology.org/2025.clicit-1.82/</dc:identifier>
<dc:language>eng</dc:language>
<dc:relation>ispartofbook:Proceedings of the Eleventh Italian Conference on Computational Linguistics (CLiC-it 2025)</dc:relation>
<dc:relation>Eleventh Italian Conference on Computational Linguistics (CLiC-it 2025)</dc:relation>
<dc:relation>firstpage:870</dc:relation>
<dc:relation>lastpage:885</dc:relation>
<dc:publisher>CEUR Workshop Proceedings</dc:publisher>
</oai_dc:dc></metadata></record></GetRecord></OAI-PMH>