Policy-based Sentence Simplification: Replacing Parallel Corpora with LLM-as-a-Judge

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wu, Xuanxin, Arase, Yuki, Nagata, Masaaki
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915658096181248
author Wu, Xuanxin
Arase, Yuki
Nagata, Masaaki
author_facet Wu, Xuanxin
Arase, Yuki
Nagata, Masaaki
contents Sentence simplification aims to modify a sentence to make it easier to read and understand while preserving the meaning. Different applications require distinct simplification policies, such as replacing only complex words at the lexical level or rewriting the entire sentence while trading off details for simplicity. However, achieving such policy-driven control remains an open challenge. In this work, we introduce a simple yet powerful approach that leverages Large Language Model-as-a-Judge (LLM-as-a-Judge) to automatically construct policy-aligned training data, completely removing the need for costly human annotation or parallel corpora. Our method enables building simplification systems that adapt to diverse simplification policies. Remarkably, even small-scale open-source LLMs such as Phi-3-mini-3.8B surpass GPT-4o on lexical-oriented simplification, while achieving comparable performance on overall rewriting, as verified by both automatic metrics and human evaluations. The consistent improvements across model families and sizes demonstrate the robustness of our approach.
format Preprint
id arxiv_https___arxiv_org_abs_2512_06228
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Policy-based Sentence Simplification: Replacing Parallel Corpora with LLM-as-a-Judge
Wu, Xuanxin
Arase, Yuki
Nagata, Masaaki
Computation and Language
Sentence simplification aims to modify a sentence to make it easier to read and understand while preserving the meaning. Different applications require distinct simplification policies, such as replacing only complex words at the lexical level or rewriting the entire sentence while trading off details for simplicity. However, achieving such policy-driven control remains an open challenge. In this work, we introduce a simple yet powerful approach that leverages Large Language Model-as-a-Judge (LLM-as-a-Judge) to automatically construct policy-aligned training data, completely removing the need for costly human annotation or parallel corpora. Our method enables building simplification systems that adapt to diverse simplification policies. Remarkably, even small-scale open-source LLMs such as Phi-3-mini-3.8B surpass GPT-4o on lexical-oriented simplification, while achieving comparable performance on overall rewriting, as verified by both automatic metrics and human evaluations. The consistent improvements across model families and sizes demonstrate the robustness of our approach.
title Policy-based Sentence Simplification: Replacing Parallel Corpora with LLM-as-a-Judge
topic Computation and Language
url https://arxiv.org/abs/2512.06228