Optimizing Statistical Machine Translation for Text Simplification
Abstract
Most recent sentence simplification systems use basic machine translation models to learn lexical and syntactic paraphrases from a manually simplified parallel corpus. These methods are limited by the quality and quantity of manually simplified corpora, which are expensive to build. In this paper, we conduct an in-depth adaptation of statistical machine translation to perform text simplification, taking advantage of large-scale paraphrases learned from bilingual texts and a small amount of manual simplifications with multiple references. Our work is the first to design automatic metrics that are effective for tuning and evaluating simplification systems, which will facilitate iterative development for this task.
Full Text:
PDF (presented at ACL 2016)Refbacks
- There are currently no refbacks.
Copyright (c) 2016 Association for Computational Linguistics

This work is licensed under a Creative Commons Attribution 4.0 International License.