Abstract
Primary data from small, low-resource languages of Oceania have only recently become available through language documentation. In our study, we explore corpus data of five Oceanic languages of Melanesia which are known to be mood-prominent (in the sense of Bhat, 1999). In order to find out more about tense, aspect, modality, and polarity, we tagged these categories in a subset of our corpora. For the category of modality, we developed a novel tag set (MelaTAMP, 2017), which categorizes clauses into factual, possible, and counterfactual. Based on an analysis of the inter-annotator consistency, we argue that our tag set for the modal domain is efficient for our subject languages and might be useful for other languages and purposes.- Anthology ID:
- W19-4008
- Volume:
- Proceedings of the 13th Linguistic Annotation Workshop
- Month:
- August
- Year:
- 2019
- Address:
- Florence, Italy
- Editors:
- Annemarie Friedrich, Deniz Zeyrek, Jet Hoek
- Venue:
- LAW
- SIG:
- SIGANN
- Publisher:
- Association for Computational Linguistics
- Note:
- Pages:
- 65–70
- Language:
- URL:
- https://aclanthology.org/W19-4008
- DOI:
- 10.18653/v1/W19-4008
- Cite (ACL):
- Annika Tjuka, Lena Weißmann, and Kilu von Prince. 2019. Tagging modality in Oceanic languages of Melanesia. In Proceedings of the 13th Linguistic Annotation Workshop, pages 65–70, Florence, Italy. Association for Computational Linguistics.
- Cite (Informal):
- Tagging modality in Oceanic languages of Melanesia (Tjuka et al., LAW 2019)
- PDF:
- https://preview.aclanthology.org/revert-3132-ingestion-checklist/W19-4008.pdf