MasakhaNER : Named Entity Recognition for African Languages

Adelani, David Ifeoluwa and Chukwuneke, Chiamaka (2021) MasakhaNER : Named Entity Recognition for African Languages. arXiv. ISSN 2331-8422

Full text not available from this repository.

Abstract

We take a step towards addressing the under-representation of the African continent in NLP research by creating the first large publicly available high-quality dataset for named entity recognition (NER) in ten African languages, bringing together a variety of stakeholders. We detail characteristics of the languages to help researchers understand the challenges that these languages pose for NER. We analyze our datasets and conduct an extensive empirical evaluation of state-of-the-art methods across both supervised and transfer learning settings. We release the data, code, and models in order to inspire future research on African NLP.

Item Type:
Journal Article
Journal or Publication Title:
arXiv
Additional Information:
Accepted at the AfricaNLP Workshop @EACL 2021
ID Code:
154462
Deposited By:
Deposited On:
23 Jun 2021 04:28
Refereed?:
Yes
Published?:
Published
Last Modified:
15 Jul 2024 21:45