Neural Correction Model for Open-Domain Named Entity Recognition

Zhu, Mengdi; Deng, Zheye; Xiong, Wenhan; Yu, Mo; Zhang, Ming; Wang, William Yang

Computer Science > Computation and Language

arXiv:1909.06058 (cs)

[Submitted on 13 Sep 2019 (v1), last revised 1 Nov 2020 (this version, v2)]

Title:Neural Correction Model for Open-Domain Named Entity Recognition

Authors:Mengdi Zhu, Zheye Deng, Wenhan Xiong, Mo Yu, Ming Zhang, William Yang Wang

View PDF

Abstract:Named Entity Recognition (NER) plays an important role in a wide range of natural language processing tasks, such as relation extraction, question answering, etc. However, previous studies on NER are limited to particular genres, using small manually-annotated or large but low-quality datasets. Meanwhile, previous datasets for open-domain NER, built using distant supervision, suffer from low precision, recall and ratio of annotated tokens (RAT). In this work, to address the low precision and recall problems, we first utilize DBpedia as the source of distant supervision to annotate abstracts from Wikipedia and design a neural correction model trained with a human-annotated NER dataset, DocRED, to correct the false entity labels. In this way, we build a large and high-quality dataset called AnchorNER and then train various models with it. To address the low RAT problem of previous datasets, we introduce a multi-task learning method to exploit the context information. We evaluate our methods on five NER datasets and our experimental results show that models trained with AnchorNER and our multi-task learning method obtain state-of-the-art performances in the open-domain setting.

Subjects:	Computation and Language (cs.CL); Information Retrieval (cs.IR); Machine Learning (cs.LG)
Cite as:	arXiv:1909.06058 [cs.CL]
	(or arXiv:1909.06058v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1909.06058

Submission history

From: Mengdi Zhu [view email]
[v1] Fri, 13 Sep 2019 06:44:30 UTC (308 KB)
[v2] Sun, 1 Nov 2020 10:14:01 UTC (505 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2019-09

Change to browse by:

cs
cs.IR
cs.LG

References & Citations

DBLP - CS Bibliography

listing | bibtex

Wenhan Xiong
Mo Yu
Ming Zhang
William Yang Wang

export BibTeX citation

Computer Science > Computation and Language

Title:Neural Correction Model for Open-Domain Named Entity Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Neural Correction Model for Open-Domain Named Entity Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators