Continual Named Entity Recognition without Catastrophic Forgetting

Zhang, Duzhen; Cong, Wei; Dong, Jiahua; Yu, Yahan; Chen, Xiuyi; Zhang, Yonggang; Fang, Zhen

Computer Science > Computation and Language

arXiv:2310.14541 (cs)

[Submitted on 23 Oct 2023]

Title:Continual Named Entity Recognition without Catastrophic Forgetting

Authors:Duzhen Zhang, Wei Cong, Jiahua Dong, Yahan Yu, Xiuyi Chen, Yonggang Zhang, Zhen Fang

View PDF

Abstract:Continual Named Entity Recognition (CNER) is a burgeoning area, which involves updating an existing model by incorporating new entity types sequentially. Nevertheless, continual learning approaches are often severely afflicted by catastrophic forgetting. This issue is intensified in CNER due to the consolidation of old entity types from previous steps into the non-entity type at each step, leading to what is known as the semantic shift problem of the non-entity type. In this paper, we introduce a pooled feature distillation loss that skillfully navigates the trade-off between retaining knowledge of old entity types and acquiring new ones, thereby more effectively mitigating the problem of catastrophic forgetting. Additionally, we develop a confidence-based pseudo-labeling for the non-entity type, \emph{i.e.,} predicting entity types using the old model to handle the semantic shift of the non-entity type. Following the pseudo-labeling process, we suggest an adaptive re-weighting type-balanced learning strategy to handle the issue of biased type distribution. We carried out comprehensive experiments on ten CNER settings using three different datasets. The results illustrate that our method significantly outperforms prior state-of-the-art approaches, registering an average improvement of $6.3$\% and $8.0$\% in Micro and Macro F1 scores, respectively.

Comments:	Accepted by EMNLP2023 main conference as a long paper
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2310.14541 [cs.CL]
	(or arXiv:2310.14541v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2310.14541

Submission history

From: Duzhen Zhang [view email]
[v1] Mon, 23 Oct 2023 03:45:30 UTC (1,926 KB)

Computer Science > Computation and Language

Title:Continual Named Entity Recognition without Catastrophic Forgetting

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Continual Named Entity Recognition without Catastrophic Forgetting

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators