Papers › A User-Centered Evaluation of Spanish Text Simplification

A User-Centered Evaluation of Spanish Text Simplification

15 Aug 2023arXiv:2308.07556archive 2025-07-28

Adrian de Wynter, Anthony Hevia, Si-Qing Chen

We present an evaluation of text simplification (TS) in Spanish for a production system, by means of two corpora focused in both complex-sentence and complex-word identification. We compare the most prevalent Spanish-specific readability scores with neural networks, and show that the latter are consistently better at predicting user preferences regarding TS. As part of our analysis, we find that multilingual models underperform against equivalent Spanish-only models on the same task, yet all models focus too often on spurious statistical features, such as sentence length. We release the corpora in our evaluation to the broader community with the hopes of pushing forward the state-of-the-art in Spanish natural language processing.

PaperPDFCode

Code

microsoft/breve-claro officialmentioned in papermentioned on GitHub report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Complex Word IdentificationSentenceText Simplification

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

FocusTS

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections