Introducing the Digital Language Equality Metric: Contextual Factors

Annika Grützner-Zahn, Georg Rehm


Abstract
In our digital age, digital language equality is an important goal to enable participation in society for all citizens, independent of the language they speak. To assess the current state of play with regard to Europe’s languages, we developed, in the project European Language Equality, a metric for digital language equality that consists of two parts, technological and contextual (i.e., non-technological) factors. We present a metric for calculating the contextual factors for over 80 European languages. For each language, a score is calculated that reflects the broader context or socio-economic ecosystem of a language, which has, for a given language, a direct impact for technology and resource development; it is important to note, though, that Language Technologies and Resources related aspects are reflected by the technological factors. To reduce the vast number of potential contextual factors to an adequate number, five different configurations were calculated and evaluated with a panel of experts. The best results were achieved by a configuration in which 12 manually curated factors were included. In the factor selection process, attention was paid to data quality, automatic updatability, inclusion of data from different domains, and a balance between different data types. The evaluation shows that this specific configuration is stable for the official EU languages; while for regional and minority languages, as well as national non-official EU languages, there is room for improvement.
Anthology ID:
2022.tdle-1.2
Volume:
Proceedings of the Workshop Towards Digital Language Equality within the 13th Language Resources and Evaluation Conference
Month:
June
Year:
2022
Address:
Marseille, France
Venue:
TDLE
SIG:
Publisher:
European Language Resources Association
Note:
Pages:
13–26
Language:
URL:
https://aclanthology.org/2022.tdle-1.2
DOI:
Bibkey:
Cite (ACL):
Annika Grützner-Zahn and Georg Rehm. 2022. Introducing the Digital Language Equality Metric: Contextual Factors. In Proceedings of the Workshop Towards Digital Language Equality within the 13th Language Resources and Evaluation Conference, pages 13–26, Marseille, France. European Language Resources Association.
Cite (Informal):
Introducing the Digital Language Equality Metric: Contextual Factors (Grützner-Zahn & Rehm, TDLE 2022)
Copy Citation:
PDF:
https://preview.aclanthology.org/ingestion-script-update/2022.tdle-1.2.pdf