Marco González
Also published as: Marco Gonzalez
2026
DGS-BIGEKO: A Dataset for Hypothetical Emergency Scenarios in German Sign Language
Cristina Luna Jimenez | Lennart Eing | Daksitha Senel Withanage Don | Marco González | Fabrizio Nunnari | Pamela Perniss | Patrick Gebhard | Elisabeth Andre
Proceedings of the LREC 2026 12th Workshop on the Representation and Processing of Sign Languages: Language in Motion
Cristina Luna Jimenez | Lennart Eing | Daksitha Senel Withanage Don | Marco González | Fabrizio Nunnari | Pamela Perniss | Patrick Gebhard | Elisabeth Andre
Proceedings of the LREC 2026 12th Workshop on the Representation and Processing of Sign Languages: Language in Motion
In this article, we describe DGS-BIGEKO, a sign language dataset containing a conversation in a crisis scenario signed by a professional interpreter in German Sign Language (DGS). The dataset comprises 14 sentences with common questions and answers from protocols occurring in emergency call scenarios translated into DGS. Additionally, the dataset contains signs for an additional 108 concepts that are relevant to emergency call scenarios. The dataset is intended to support research in sign language linguistics and sign language machine translation by providing resources in a very specific domain, where no previous resources are available in DGS. The dataset is freely available for research purposes at the following address: https://doi.org/10.5281/zenodo.18458557
2024
DGS-Fabeln-1: A Multi-Angle Parallel Corpus of Fairy Tales between German Sign Language and German Text
Fabrizio Nunnari | Eleftherios Avramidis | Cristina España-Bonet | Marco González | Anna Hennes | Patrick Gebhard
Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024)
Fabrizio Nunnari | Eleftherios Avramidis | Cristina España-Bonet | Marco González | Anna Hennes | Patrick Gebhard
Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024)
We present the acquisition process and the data of DGS-Fabeln-1, a parallel corpus of German text and videos containing German fairy tales interpreted into the German Sign Language (DGS) by a native DGS signer. The corpus contains 573 segments of videos with a total duration of 1 hour and 32 minutes, corresponding with 1428 written sentences. It is the first corpus of semi-naturally expressed DGS that has been filmed from 7 angles, and one of the few sign language (SL) corpora globally which have been filmed from more than 3 angles and where the listener has been simultaneously filmed. The corpus aims at aiding research at SL linguistics, SL machine translation and affective computing, and is freely available for research purposes at the following address: https://doi.org/10.5281/zenodo.10822097.
2017
Processo de construção de um corpus anotado com Entidades Geológicas visando REN (Building an annotated corpus with geological entities for NER)[In Portuguese]
Daniela Amaral | Sandra Collovini | Anny Figueira | Renata Vieira | Marco Gonzalez
Proceedings of the 11th Brazilian Symposium in Information and Human Language Technology
Daniela Amaral | Sandra Collovini | Anny Figueira | Renata Vieira | Marco Gonzalez
Proceedings of the 11th Brazilian Symposium in Information and Human Language Technology