Deep neural self-training for scientific keyphrase extraction

Research Area: Machine Learning

Abstract:

Scientific information extraction is a crucial step for understanding scientific publications. In this paper, we focus on scientific keyphrase extraction, which aims to identify keyphrases from scientific articles and classify them into predefined categories. We present a neural network based approach for this task, which employs the bidirectional long short-memory (LSTM) to represent the sentences in the article. On top of the bidirectional LSTM layer in our neural model, conditional random field (CRF) is used to predict the label sequence for the whole sentence. Considering the expensive annotated data for supervised learning methods, we introduce self-training method into our neural model to leverage the unlabeled articles. Experimental results on the ScienceIE corpus and ACL keyphrase corpus show that our neural model achieves promising performance without any hand-designed features and external knowledge resources. Furthermore, it efficiently incorporates the unlabeled data and achieve competitive performance compared with previous state-of-the-art systems.

Keywords:

Author(s) Name: Xun Zhu,Chen Lyu ,Donghong Ji ,Han Liao,Fei Li

Journal name: PLoS ONE

Conferrence name:

Publisher name: PLOS

DOI: 10.1371/journal.pone.0232547

Volume Information: Volume 15, Issue No (5)

Paper Link: https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0232547

Office Address

Social List

Deep neural model with self-training for scientific keyphrase extraction - 2020

Abstract:

S-Logix (OPC) Private Limited

Office Address

Deep neural model with self-training for scientific keyphrase extraction - 2020

Abstract:

Related Papers