Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models

Bogdan Nicula; Mihai Dascalu; Natalie Newton; Ellen Orcutt; Danielle S. McNamara

doi:10.1007/978-3-030-80421-3_36

Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models

Bogdan Nicula, Mihai Dascalu, Natalie Newton, Ellen Orcutt, Danielle S. McNamara

Psychology

Research output: Chapter in Book/Report/Conference proceeding › Conference contribution

6 Scopus citations

Abstract

The ability to automatically assess the quality of paraphrases can be very useful for facilitating literacy skills and providing timely feedback to learners. Our aim is twofold: a) to automatically evaluate the quality of paraphrases across four dimensions: lexical similarity, syntactic similarity, semantic similarity and paraphrase quality, and b) to assess how well models trained for this task generalize. The task is modeled as a classification problem and three different methods are explored: a) manual feature extraction combined with an Extra Trees model, b) GloVe embeddings and a Siamese neural network, and c) using a pretrained BERT model fine-tuned on our task. Starting from a dataset of 1998 paraphrases from the User Language Paraphrase Corpus (ULPC), we explore how the three models trained on the ULPC dataset generalize when applied on a separate, small paraphrase corpus based on children inputs. The best out-of-the-box generalization performance is obtained by the Extra Trees model with at least 75% average F1-scores for the three similarity dimensions. We also show that the Siamese neural network and BERT models can obtain an improvement of at least 5% after fine-tuning across all dimensions.

Original language	English (US)
Title of host publication	Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings
Editors	Alexandra I. Cristea, Christos Troussas
Publisher	Springer Science and Business Media Deutschland GmbH
Pages	333-340
Number of pages	8
ISBN (Print)	9783030804206
DOIs	https://doi.org/10.1007/978-3-030-80421-3_36
State	Published - 2021
Event	17th International Conference on Intelligent Tutoring Systems, ITS 2021 - Virtual, Online Duration: Jun 7 2021 → Jun 11 2021

Publication series

Name	Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
Volume	12677 LNCS
ISSN (Print)	0302-9743
ISSN (Electronic)	1611-3349

Conference

Conference	17th International Conference on Intelligent Tutoring Systems, ITS 2021
City	Virtual, Online
Period	6/7/21 → 6/11/21

Keywords

Language models
Natural language processing
Paraphrase quality assessment
Recurrent neural networks

ASJC Scopus subject areas

Theoretical Computer Science
General Computer Science

Access to Document

10.1007/978-3-030-80421-3_36

Cite this

Nicula, B., Dascalu, M., Newton, N., Orcutt, E., & McNamara, D. S. (2021). Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models. In A. I. Cristea, & C. Troussas (Eds.), Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings (pp. 333-340). (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics); Vol. 12677 LNCS). Springer Science and Business Media Deutschland GmbH. https://doi.org/10.1007/978-3-030-80421-3_36

Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models. / Nicula, Bogdan; Dascalu, Mihai; Newton, Natalie et al.
Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings. ed. / Alexandra I. Cristea; Christos Troussas. Springer Science and Business Media Deutschland GmbH, 2021. p. 333-340 (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics); Vol. 12677 LNCS).

Research output: Chapter in Book/Report/Conference proceeding › Conference contribution

Nicula, B, Dascalu, M, Newton, N, Orcutt, E & McNamara, DS 2021, Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models. in AI Cristea & C Troussas (eds), Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings. Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics), vol. 12677 LNCS, Springer Science and Business Media Deutschland GmbH, pp. 333-340, 17th International Conference on Intelligent Tutoring Systems, ITS 2021, Virtual, Online, 6/7/21. https://doi.org/10.1007/978-3-030-80421-3_36

Nicula B, Dascalu M, Newton N, Orcutt E, McNamara DS. Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models. In Cristea AI, Troussas C, editors, Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings. Springer Science and Business Media Deutschland GmbH. 2021. p. 333-340. (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)). doi: 10.1007/978-3-030-80421-3_36

Nicula, Bogdan ; Dascalu, Mihai ; Newton, Natalie et al. / Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models. Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings. editor / Alexandra I. Cristea ; Christos Troussas. Springer Science and Business Media Deutschland GmbH, 2021. pp. 333-340 (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)).

@inproceedings{5bfdf5a3d467437c83e1d379bac46ade,

title = "Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models",

abstract = "The ability to automatically assess the quality of paraphrases can be very useful for facilitating literacy skills and providing timely feedback to learners. Our aim is twofold: a) to automatically evaluate the quality of paraphrases across four dimensions: lexical similarity, syntactic similarity, semantic similarity and paraphrase quality, and b) to assess how well models trained for this task generalize. The task is modeled as a classification problem and three different methods are explored: a) manual feature extraction combined with an Extra Trees model, b) GloVe embeddings and a Siamese neural network, and c) using a pretrained BERT model fine-tuned on our task. Starting from a dataset of 1998 paraphrases from the User Language Paraphrase Corpus (ULPC), we explore how the three models trained on the ULPC dataset generalize when applied on a separate, small paraphrase corpus based on children inputs. The best out-of-the-box generalization performance is obtained by the Extra Trees model with at least 75% average F1-scores for the three similarity dimensions. We also show that the Siamese neural network and BERT models can obtain an improvement of at least 5% after fine-tuning across all dimensions.",

keywords = "Language models, Natural language processing, Paraphrase quality assessment, Recurrent neural networks",

author = "Bogdan Nicula and Mihai Dascalu and Natalie Newton and Ellen Orcutt and McNamara, {Danielle S.}",

note = "Funding Information: Acknowledgments. The work was funded by a grant of the Romanian National Authority for Scientific Research and Innovation, CNCS – UEFISCDI, project number TE 70 PN-III-P1-1.1-TE-2019-2209, ATES – “Automated Text Evaluation and Simplification”. This research was also supported in part by the Institute of Education Sciences (R305A190063 and R305A190050) and the Office of Naval Research (N00014-17-1-2300 and N00014-19-1-2424). The opinions expressed are those of the authors and do not represent views of the IES or ONR. Publisher Copyright: {\textcopyright} 2021, Springer Nature Switzerland AG.; 17th International Conference on Intelligent Tutoring Systems, ITS 2021 ; Conference date: 07-06-2021 Through 11-06-2021",

year = "2021",

doi = "10.1007/978-3-030-80421-3_36",

language = "English (US)",

isbn = "9783030804206",

series = "Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)",

publisher = "Springer Science and Business Media Deutschland GmbH",

pages = "333--340",

editor = "Cristea, {Alexandra I.} and Christos Troussas",

booktitle = "Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings",

address = "Germany",

}

TY - GEN

T1 - Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models

AU - Nicula, Bogdan

AU - Dascalu, Mihai

AU - Newton, Natalie

AU - Orcutt, Ellen

AU - McNamara, Danielle S.

N1 - Funding Information: Acknowledgments. The work was funded by a grant of the Romanian National Authority for Scientific Research and Innovation, CNCS – UEFISCDI, project number TE 70 PN-III-P1-1.1-TE-2019-2209, ATES – “Automated Text Evaluation and Simplification”. This research was also supported in part by the Institute of Education Sciences (R305A190063 and R305A190050) and the Office of Naval Research (N00014-17-1-2300 and N00014-19-1-2424). The opinions expressed are those of the authors and do not represent views of the IES or ONR. Publisher Copyright: © 2021, Springer Nature Switzerland AG.

PY - 2021

Y1 - 2021

N2 - The ability to automatically assess the quality of paraphrases can be very useful for facilitating literacy skills and providing timely feedback to learners. Our aim is twofold: a) to automatically evaluate the quality of paraphrases across four dimensions: lexical similarity, syntactic similarity, semantic similarity and paraphrase quality, and b) to assess how well models trained for this task generalize. The task is modeled as a classification problem and three different methods are explored: a) manual feature extraction combined with an Extra Trees model, b) GloVe embeddings and a Siamese neural network, and c) using a pretrained BERT model fine-tuned on our task. Starting from a dataset of 1998 paraphrases from the User Language Paraphrase Corpus (ULPC), we explore how the three models trained on the ULPC dataset generalize when applied on a separate, small paraphrase corpus based on children inputs. The best out-of-the-box generalization performance is obtained by the Extra Trees model with at least 75% average F1-scores for the three similarity dimensions. We also show that the Siamese neural network and BERT models can obtain an improvement of at least 5% after fine-tuning across all dimensions.

AB - The ability to automatically assess the quality of paraphrases can be very useful for facilitating literacy skills and providing timely feedback to learners. Our aim is twofold: a) to automatically evaluate the quality of paraphrases across four dimensions: lexical similarity, syntactic similarity, semantic similarity and paraphrase quality, and b) to assess how well models trained for this task generalize. The task is modeled as a classification problem and three different methods are explored: a) manual feature extraction combined with an Extra Trees model, b) GloVe embeddings and a Siamese neural network, and c) using a pretrained BERT model fine-tuned on our task. Starting from a dataset of 1998 paraphrases from the User Language Paraphrase Corpus (ULPC), we explore how the three models trained on the ULPC dataset generalize when applied on a separate, small paraphrase corpus based on children inputs. The best out-of-the-box generalization performance is obtained by the Extra Trees model with at least 75% average F1-scores for the three similarity dimensions. We also show that the Siamese neural network and BERT models can obtain an improvement of at least 5% after fine-tuning across all dimensions.

KW - Language models

KW - Natural language processing

KW - Paraphrase quality assessment

KW - Recurrent neural networks

UR - http://www.scopus.com/inward/record.url?scp=85112250426&partnerID=8YFLogxK

UR - http://www.scopus.com/inward/citedby.url?scp=85112250426&partnerID=8YFLogxK

U2 - 10.1007/978-3-030-80421-3_36

DO - 10.1007/978-3-030-80421-3_36

M3 - Conference contribution

AN - SCOPUS:85112250426

SN - 9783030804206

T3 - Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)

SP - 333

EP - 340

BT - Intelligent Tutoring Systems - 17th International Conference, ITS 2021, Proceedings

A2 - Cristea, Alexandra I.

A2 - Troussas, Christos

PB - Springer Science and Business Media Deutschland GmbH

T2 - 17th International Conference on Intelligent Tutoring Systems, ITS 2021

Y2 - 7 June 2021 through 11 June 2021

ER -

Automated Paraphrase Quality Assessment Using Recurrent Neural Networks and Language Models

Abstract

Publication series

Conference

Keywords

ASJC Scopus subject areas

Access to Document

Other files and links

Fingerprint

Cite this