ResearchTrend.AI
  • Papers
  • Communities
  • Events
  • Blog
  • Pricing
Papers
Communities
Social Events
Terms and Conditions
Pricing
Parameter LabParameter LabTwitterGitHubLinkedInBlueskyYoutube

© 2025 ResearchTrend.AI, All rights reserved.

  1. Home
  2. Papers
  3. 2304.03952
59
0

MphayaNER: Named Entity Recognition for Tshivenda

8 April 2023
R. Mbuvha
David Ifeoluwa Adelani
Tendani Mutavhatsindi
Tshimangadzo Rakhuhu
A. Mauda
Tshifhiwa Joshua Maumela
Andisani Masindi
Seani Rananga
Vukosi Marivate
T. Marwala
ArXivPDFHTML
Abstract

Named Entity Recognition (NER) plays a vital role in various Natural Language Processing tasks such as information retrieval, text classification, and question answering. However, NER can be challenging, especially in low-resource languages with limited annotated datasets and tools. This paper adds to the effort of addressing these challenges by introducing MphayaNER, the first Tshivenda NER corpus in the news domain. We establish NER baselines by \textit{fine-tuning} state-of-the-art models on MphayaNER. The study also explores zero-shot transfer between Tshivenda and other related Bantu languages, with chiShona and Kiswahili showing the best results. Augmenting MphayaNER with chiShona data was also found to improve model performance significantly. Both MphayaNER and the baseline models are made publicly available.

View on arXiv
Comments on this paper