A Morphologically-Analyzed CHILDES Corpus of Hebrew

Bracha Nir; Brian MacWhinney; Shuly Wintner

A Morphologically-Analyzed CHILDES Corpus of Hebrew

Bracha Nir, Brian MacWhinney, Shuly Wintner

Abstract

We present a corpus of transcribed spoken Hebrew that forms an integral part of a comprehensive data system that has been developed to suit the specific needs and interests of child language researchers: CHILDES (Child Language Data Exchange System). We introduce a dedicated transcription scheme for the spoken Hebrew data that is aware both of the phonology and of the standard orthography of the language. We also introduce a morphological analyzer that was specifically developed for this corpus.

Anthology ID:: L10-1112
Volume:: Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)
Month:: May
Year:: 2010
Address:: Valletta, Malta
Editors:: Nicoletta Calzolari, Khalid Choukri, Bente Maegaard, Joseph Mariani, Jan Odijk, Stelios Piperidis, Mike Rosner, Daniel Tapias
Venue:: LREC
SIG:
Publisher:: European Language Resources Association (ELRA)
Note:
Pages:
Language:
URL:: http://www.lrec-conf.org/proceedings/lrec2010/pdf/172_Paper.pdf
DOI:
Bibkey:
Cite (ACL):: Bracha Nir, Brian MacWhinney, and Shuly Wintner. 2010. A Morphologically-Analyzed CHILDES Corpus of Hebrew. In Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10), Valletta, Malta. European Language Resources Association (ELRA).
Cite (Informal):: A Morphologically-Analyzed CHILDES Corpus of Hebrew (Nir et al., LREC 2010)
Copy Citation:
PDF:: http://www.lrec-conf.org/proceedings/lrec2010/pdf/172_Paper.pdf

PDF Cite Search Fix data