arXiv · 2305.12463
Teaching the Pre-trained Model to Generate Simple Texts for Text Simplification
Abstract
Randomly masking text spans in ordinary texts in the pre-training stage hardly allows models to acquire the ability to generate simple texts. It can hurt the performance of pre-trained models on text simplification tasks. In this paper, we propose a new continued pre-training strategy to teach the pre-trained model to generate simple texts. We continue pre-training BART, a representative model, to obtain SimpleBART. It consistently and significantly improves the results on lexical simplification, sentence simplification, and document-level simplification tasks over BART. At the end, we compare SimpleBART with several representative large language models (LLMs).
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Renliang Sun, Wei Xu, Xiaojun Wan. 2023-05-21. Teaching the Pre-trained Model to Generate Simple Texts for Text Simplification. https://arxiv.org/abs/2305.12463
Cite the original work for its findings. Save a collection to share your selection of sources.