From Paraphrase Database to Compositional Paraphrase Model and Back | Zendy

John Wieting | Zendy; Mohit Bansal | Zendy; Kevin Gimpel | Zendy; Karen Livescu | Zendy

AI Assistant Blog Pricing

Home ZAIA Blog

Open Access

From Paraphrase Database to Compositional Paraphrase Model and Back

Author(s) -

John Wieting,

Mohit Bansal,

Kevin Gimpel,

Karen Livescu

Publication year - 2015

Publication title -

transactions of the association for computational linguistics

Language(s) - English

Resource type - Journals

ISSN - 2307-387X

DOI - 10.1162/tacl_a_00143

Subject(s) - paraphrase , bigram , computer science , phrase , leverage (statistics) , natural language processing , artificial intelligence , information retrieval , trigram

The Paraphrase Database (PPDB; Ganitkevitch et al., 2013) is an extensive semantic resource, consisting of a list of phrase pairs with (heuristic) confidence estimates. However, it is still unclear how it can best be used, due to the heuristic nature of the confidences and its necessarily incomplete coverage. We propose models to leverage the phrase pairs from the PPDB to build parametric paraphrase models that score paraphrase pairs more accurately than the PPDB’s internal scores while simultaneously improving its coverage. They allow for learning phrase embeddings as well as improved word embeddings. Moreover, we introduce two new, manually annotated datasets to evaluate short-phrase paraphrasing models. Using our paraphrase model trained using PPDB, we achieve state-of-the-art results on standard word and bigram similarity tasks and beat strong baselines on our new short phrase paraphrase tasks.

The content you want is available to Zendy users.

Already have an account? Click here to sign in.

Having issues? You can contact us here

Accelerating Research