Library Faculty Publications

SIGMORPHON 2021 Shared Task on Morphological Reinflection: Generalization Across Languages

Tiago Pimentel, University of Cambridge
Maria Ryskina, Carnegie Mellon University
Christopher Straughn, Northeastern Illinois UniversityFollow

Document Type

Conference Proceeding

Publication Date

2021

Abstract

This year’s iteration of the SIGMORPHON Shared Task on morphological reinflection focuses on typological diversity and cross-lingual variation of morphosyntactic features. In terms of the task, we enrich UniMorph with new data for 32 languages from 13 language families, with most of them being under-resourced: Kunwinjku, Classical Syriac, Arabic (Modern Standard, Egyptian, Gulf), Hebrew, Amharic, Aymara, Magahi, Braj, Kurdish (Central, Northern, Southern), Polish, Karelian, Livvi, Ludic, Veps, Võro, Evenki, Xibe, Tuvan, Sakha, Turkish, Indonesian, Kodi, Seneca, Asháninka, Yanesha, Chukchi, Itelmen, Eibela. We evaluate six systems on the new data and conduct an extensive error analysis of the systems’ predictions. Transformer-based models generally demonstrate superior performance on the majority of languages, achieving >90% accuracy on 65% of them. The languages on which systems yielded low accuracy are mainly under-resourced, with a limited amount of data. Most errors made by the systems are due to allomorphy, honorificity, and form variation. In addition, we observe that systems especially struggle to inflect multiword lemmas. The systems also produce misspelled forms or end up in repetitive loops (e.g., RNN-based models). Finally, we report a large drop in systems’ performance on previously unseen lemmas.

Creative Commons License

This work is licensed under a Creative Commons Attribution 4.0 International License.

Publication Title

Proceedings of the 18th SIGMORPHON Workshop on Computational Research in Phonetics, Phonology, and Morphology

First Page

229

Last Page

259

Recommended Citation

Pimentel, Tiago; Ryskina, Maria; et al., "SIGMORPHON 2021 Shared Task on Morphological Reinflection: Generalization Across Languages" (2021). Library Faculty Publications. 10. https://neiudc.neiu.edu/lib-pub/10

Download

Find in your library

Included in

Computational Linguistics Commons, Morphology Commons

COinS

NEIU Digital Commons

Library Faculty Publications

SIGMORPHON 2021 Shared Task on Morphological Reinflection: Generalization Across Languages

Document Type

Publication Date

Abstract

Creative Commons License

Publication Title

First Page

Last Page

Recommended Citation

Included in

Search

Browse

Links

NEIU Digital Commons

Library Faculty Publications

SIGMORPHON 2021 Shared Task on Morphological Reinflection: Generalization Across Languages

Authors

Document Type

Publication Date

Abstract

Creative Commons License

Publication Title

First Page

Last Page

Recommended Citation

Included in

Share

Search

Browse

Links