Bermet Chontaeva
2026
Negation of Turkic non-verbal clauses: Analysis and Universal Dependencies Implementation
Nikolett Mus | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Sardana Ivanova | Çağrı Çöltekin | Jonathan N. Washington | Gulnura Dzhumalieva | Aida Kasieva
Proceedings of the Ninth Workshop on Universal Dependencies (UDW 2026)
Nikolett Mus | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Sardana Ivanova | Çağrı Çöltekin | Jonathan N. Washington | Gulnura Dzhumalieva | Aida Kasieva
Proceedings of the Ninth Workshop on Universal Dependencies (UDW 2026)
The paper examines the grammatical behavior of the negative element used to negate predicates in non-verbal clauses in three Turkic languages: Azerbaijani, Kyrgyz, and Turkish. We focus on its interaction with verbal copulas, subject agreement, and the distribution of agreement suffixes, as well as its position within the predicate phrase. The study draws on both previously described corpus data and newly collected examples. Across all three languages, agreement features are realised on the negative element only in the absence of an overt copula. The agreement morphology involved is identical to that found with nominal, adjectival, and adverbial predicates. In all the languages examined, the negative element remains within the predicate phrase; thus, its position is syntactically constrained. At the same time, we observe differences among the languages in the degree to which the position of the negator is fixed within the predicate. In Turkish and Azerbaijani, regardless of which element of the nominal predicate it negates, the negator invariably follows the predicate. In Kyrgyz, by contrast, it consistently appears immediately after the element it negates within the predicate. These patterns suggest that the negative element behaves syntactically as a phrasal operator associated with non-verbal predicates. For annotation purposes, we therefore propose analysing the negative element as a negation modifier, assigning it the POS tag ADV and the dependency relation advmod:neg.
Tokenisation of Turkic Copula Constructions in Universal Dependencies
Cagri Coltekin | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Sardana Ivanova | Gulnura Dzhumalieva | Aida Kasieva | Nikolett Mus | Jonathan Washington
Proceedings of the Second Workshop Natural Language Processing for Turkic Languages (SIGTURK 2026)
Cagri Coltekin | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Sardana Ivanova | Gulnura Dzhumalieva | Aida Kasieva | Nikolett Mus | Jonathan Washington
Proceedings of the Second Workshop Natural Language Processing for Turkic Languages (SIGTURK 2026)
Identifying units, ’syntactic words’, for morphosyntactic analysis is important yet challenging for morphologically rich languages. In this paper we propose a set of guiding principles to determine units of morphosyntactic analysis, and apply them to the case of copular constructions in Turkic languages, in the context of Universal Dependencies (UD) framework. We also provide a survey of the practice in the Turkic UD treebanks published to date, and discuss the advantages and disadvantages of the proposed tokenisation for a selection of Turkic languages.
2025
Parallel Universal Dependencies Treebanks for Turkic Languages
Arofat Akhundjanova | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Cagri Coltekin
Proceedings of the Eighth Workshop on Universal Dependencies (UDW, SyntaxFest 2025)
Arofat Akhundjanova | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Cagri Coltekin
Proceedings of the Eighth Workshop on Universal Dependencies (UDW, SyntaxFest 2025)
We introduce the first fully aligned and manually annotated parallel Universal Dependencies (UD) treebanks for four Turkic languages: Azerbaijani, Kyrgyz, Turkish, and Uzbek. These resources currently consist of 148 strategically selected sentences that illustrate typologically significant morphosyntactic phenomena across these related yet distinct languages. These parallel treebanks enable systematic comparative studies of Turkic syntax and may be instrumental in cross-lingual NLP applications. All treebanks are available as part of UD v2.16.
2024
Strategies for the Annotation of Pronominalised Locatives in Turkic Universal Dependency Treebanks
Jonathan Washington | Çağrı Çöltekin | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Gulnura Jumalieva | Aida Kasieva | Aslı Kuzgun | Büşra Marşan | Chihiro Taguchi
Proceedings of the Joint Workshop on Multiword Expressions and Universal Dependencies (MWE-UD) @ LREC-COLING 2024
Jonathan Washington | Çağrı Çöltekin | Furkan Akkurt | Bermet Chontaeva | Soudabeh Eslami | Gulnura Jumalieva | Aida Kasieva | Aslı Kuzgun | Büşra Marşan | Chihiro Taguchi
Proceedings of the Joint Workshop on Multiword Expressions and Universal Dependencies (MWE-UD) @ LREC-COLING 2024
As part of our efforts to develop unified Universal Dependencies (UD) guidelines for Turkic languages, we evaluate multiple approaches to a difficult morphosyntactic phenomenon, pronominal locative expressions formed by a suffix -ki. These forms result in multiple syntactic words, with potentially conflicting morphological features, and participating in different dependency relations. We describe multiple approaches to the problem in current (and upcoming) Turkic UD treebanks, and show that none of them offers a solution that satisfies a number of constraints we consider (including constraints imposed by UD guidelines). This calls for a compromise with the ‘least damage’ that should be adopted by most, if not all, Turkic treebanks. Our discussion of the phenomenon and various annotation approaches may also help treebanking efforts for other languages or language families with similar constructions.