DOAJ Open Access 2021

MasakhaNER: Named Entity Recognition for African Languages

David Ifeoluwa Adelani Jade Abbott Graham Neubig Daniel D’souza Julia Kreutzer +56 lainnya

Abstrak

AbstractWe take a step towards addressing the under- representation of the African continent in NLP research by bringing together different stakeholders to create the first large, publicly available, high-quality dataset for named entity recognition (NER) in ten African languages. We detail the characteristics of these languages to help researchers and practitioners better understand the challenges they pose for NER tasks. We analyze our datasets and conduct an extensive empirical evaluation of state- of-the-art methods across both supervised and transfer learning settings. Finally, we release the data, code, and models to inspire future research on African NLP.1

Penulis (61)

D

David Ifeoluwa Adelani

J

Jade Abbott

G

Graham Neubig

D

Daniel D’souza

J

Julia Kreutzer

C

Constantine Lignos

C

Chester Palen-Michel

H

Happy Buzaaba

S

Shruti Rijhwani

S

Sebastian Ruder

S

Stephen Mayhew

I

Israel Abebe Azime

S

Shamsuddeen H. Muhammad

C

Chris Chinenye Emezue

J

Joyce Nakatumba-Nabende

P

Perez Ogayo

A

Aremu Anuoluwapo

C

Catherine Gitau

D

Derguene Mbaye

J

Jesujoba Alabi

S

Seid Muhie Yimam

T

Tajuddeen Rabiu Gwadabe

I

Ignatius Ezeani

R

Rubungo Andre Niyongabo

J

Jonathan Mukiibi

V

Verrah Otiende

I

Iroro Orife

D

Davis David

S

Samba Ngom

T

Tosin Adewumi

P

Paul Rayson

M

Mofetoluwa Adeyemi

G

Gerald Muriuki

E

Emmanuel Anebi

C

Chiamaka Chukwuneke

N

Nkiruka Odu

E

Eric Peter Wairagala

S

Samuel Oyerinde

C

Clemencia Siro

T

Tobius Saul Bateesa

T

Temilola Oloyede

Y

Yvonne Wambui

V

Victor Akinode

D

Deborah Nabagereka

M

Maurice Katusiime

A

Ayodele Awokoya

M

Mouhamadane MBOUP

D

Dibora Gebreyohannes

H

Henok Tilaye

K

Kelechi Nwaike

D

Degaga Wolde

A

Abdoulaye Faye

B

Blessing Sibanda

O

Orevaoghene Ahia

B

Bonaventure F. P. Dossou

K

Kelechi Ogueji

T

Thierno Ibrahima DIOP

A

Abdoulaye Diallo

A

Adewale Akinfaderin

T

Tendai Marengereke

S

Salomey Osei

Format Sitasi

Adelani, D.I., Abbott, J., Neubig, G., D’souza, D., Kreutzer, J., Lignos, C. et al. (2021). MasakhaNER: Named Entity Recognition for African Languages. https://doi.org/10.1162/tacl_a_00416

Akses Cepat

PDF tidak tersedia langsung

Cek di sumber asli →
Lihat di Sumber doi.org/10.1162/tacl_a_00416
Informasi Jurnal
Tahun Terbit
2021
Sumber Database
DOAJ
DOI
10.1162/tacl_a_00416
Akses
Open Access ✓