Abstract
This paper deals with the uses of the annotations of third person singular neuter pronouns in the DAD parallel and comparable corpora of Danish and Italian texts and spoken data. The annotations contain information about the functions of these pronouns and their uses as abstract anaphora. Abstract anaphora have constructions such as verbal phrases, clauses and discourse segments as antecedents and refer to abstract objects comprising events, situations and propositions. The analysis of the annotated data shows the language specific characteristics of abstract anaphora in the two languages compared with the uses of abstract anaphora in English. Finally, the paper presents machine learning experiments run on the annotated data in order to identify the functions of third person singular neuter personal pronouns and neuter demonstrative pronouns. The results of these experiments vary from corpus to corpus. However, they are all comparable with the results obtained in similar tasks in other languages. This is very promising because the experiments have been run on both written and spoken data using a classification of the pronominal functions which is much more fine-grained than the classifications used in other studies.
Original language | English |
---|---|
Title of host publication | Proceedings of the Seventh International Conference on Language Resources and Evaluation : LREC 2010 |
Number of pages | 8 |
Place of Publication | Valletta, Malta |
Publisher | European Language Resources Association |
Publication date | 2010 |
Pages | 705-712 |
Publication status | Published - 2010 |
Event | Language Resources and Evaluation Conference - ELRA, Malta Duration: 19 May 2010 → 21 May 2010 Conference number: 7 |
Conference
Conference | Language Resources and Evaluation Conference |
---|---|
Number | 7 |
Country/Territory | Malta |
City | ELRA |
Period | 19/05/2010 → 21/05/2010 |