arXiv · 1810.07931
Unsupervised Neural Text Simplification
Abstract
The paper presents a first attempt towards unsupervised neural text simplification that relies only on unlabeled text corpora. The core framework is composed of a shared encoder and a pair of attentional-decoders and gains knowledge of simplification through discrimination based-losses and denoising. The framework is trained using unlabeled text collected from en-Wikipedia dump. Our analysis (both quantitative and qualitative involving human evaluators) on a public test data shows that the proposed model can perform text-simplification at both lexical and syntactic levels, competitive to existing supervised methods. Addition of a few labelled pairs also improves the performance further.
Explore related subjects
Keep this discovery
Sai Surya, Abhijit Mishra, Anirban Laha, Parag Jain, Karthik Sankaranarayanan. 2018-10-18. Unsupervised Neural Text Simplification. https://arxiv.org/abs/1810.07931
Cite the original work for its findings. Save a collection to share your selection of sources.