Type

Conference Proceedings

Authors

Josef van Genabith
Mark Dras
Elaine Ui Dhonnchadha
Jennifer Foster
Ozlem Cetinoglu
Teresa Lynn

Subjects

Linguistics

Topics
language resources dependency computational linguistics irish language irish low density linguistics treebank

Irish treebanking and parsing: a preliminary evaluation (2012)

Abstract Language resources are essential for linguistic research and the development of NLP applications. Low- density languages, such as Irish, therefore lack significant research in this area. This paper describes the early stages in the development of new language resources for Irish – namely the first Irish dependency treebank and the first Irish statistical dependency parser. We present the methodology behind building our new treebank and the steps we take to leverage upon the few existing resources. We discuss language specific choices made when defining our dependency labelling scheme, and describe interesting Irish language characteristics such as prepositional attachment, copula and clefting. We manually develop a small treebank of 300 sentences based on an existing POS-tagged corpus and report an inter-annotator agreement of 0.7902. We train MaltParser to achieve preliminary parsing results for Irish and describe a bootstrapping approach for further stages of development.
Collections Ireland -> Dublin City University -> Publication Type = Conference or Workshop Item
Ireland -> Dublin City University -> Subject = Humanities: Linguistics
Ireland -> Dublin City University -> Subject = Computer Science
Ireland -> Dublin City University -> DCU Faculties and Centres = Research Initiatives and Centres: Centre for Next Generation Localisation (CNGL)
Ireland -> Dublin City University -> Subject = Humanities: Irish language
Ireland -> Dublin City University -> Status = Published
Ireland -> Dublin City University -> DCU Faculties and Centres = Research Initiatives and Centres: National Centre for Language Technology (NCLT)
Ireland -> Dublin City University -> Subject = Computer Science: Computational linguistics
Ireland -> Dublin City University -> Subject = Humanities
Ireland -> Dublin City University -> DCU Faculties and Centres = Research Initiatives and Centres

Full list of authors on original publication

Josef van Genabith, Mark Dras, Elaine Ui Dhonnchadha, Jennifer Foster, Ozlem Cetinoglu, Teresa Lynn

Experts in our system

1
Jennifer Foster
Dublin City University
Total Publications: 53
 
2
Ozlem Cetinoglu
Dublin City University
Total Publications: 10
 
3
Teresa Lynn
Dublin City University
Total Publications: 20