Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish

Antonio Pareja Lora; Guadalupe Aguado de Cea

Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish

Antonio Pareja-Lora, Guadalupe Aguado de Cea

Abstract

In this paper, we present an ontology-based methodology and architecture for the comparison, assessment, combination (and, to some extent, also contrastive evaluation) of the results of different linguistic tools. More specifically, we describe an experiment aiming at the improvement of the correctness of lemma tagging for Spanish. This improvement was achieved by means of the standardisation and combination of the results of three different linguistic annotation tools (Bitexts DataLexica, Connexors FDG Parser and LACELLs POS tagger), using (1) ontologies, (2) a set of lemma tagging correction rules, determined empirically during the experiment, and (3) W3C standard languages, such as XML, RDF(S) and OWL. As we show in the results of the experiment, the interoperation of these tools by means of ontologies and the correction rules applied in the experiment improved significantly the quality of the resulting lemma tagging (when compared to the separate lemma tagging performed by each of the tools that we made interoperate).

Anthology ID:: L10-1054
Volume:: Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)
Month:: May
Year:: 2010
Address:: Valletta, Malta
Editors:: Nicoletta Calzolari, Khalid Choukri, Bente Maegaard, Joseph Mariani, Jan Odijk, Stelios Piperidis, Mike Rosner, Daniel Tapias
Venue:: LREC
SIG:
Publisher:: European Language Resources Association (ELRA)
Note:
Pages:
Language:
URL:: http://www.lrec-conf.org/proceedings/lrec2010/pdf/92_Paper.pdf
DOI:
Bibkey:
Cite (ACL):: Antonio Pareja-Lora and Guadalupe Aguado de Cea. 2010. Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish. In Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10), Valletta, Malta. European Language Resources Association (ELRA).
Cite (Informal):: Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish (Pareja-Lora & de Cea, LREC 2010)
Copy Citation:
PDF:: http://www.lrec-conf.org/proceedings/lrec2010/pdf/92_Paper.pdf

PDF Cite Search Fix data