Skip to main navigation Skip to search Skip to main content

An Annotated Corpus of Direct Speech

Research output: Chapters, Conference Papers, Creative and Literary WorksRGC 32 - Refereed conference paper (with host publication)peer-review

Abstract

We propose a scheme for annotating direct speech in literary texts, based on the Text Encoding Initiative (TEI) and the coreference annotation guidelines from the Message Understanding Conference (MUC). The scheme encodes the speakers and listeners of utterances in a text, as well as the quotative verbs that reports the utterances. We measure inter-annotator agreement on this annotation task. We then present statistics on a manually annotated corpus that consists of books from the New Testament. Finally, we visualize the corpus as a conversational network.
Original languageEnglish
Title of host publicationProceedings of the Tenth International Conference on Language Resources and Evaluation (LREC 2016)
EditorsNicoletta Calzolari
Place of PublicationParis
PublisherEuropean Language Resources Association (ELRA)
Pages1059-1063
ISBN (Print)9782951740891
Publication statusPublished - May 2016
Event 10th International Conference on Language Resources and Evaluation (LREC 2016) - Portorož, Slovenia
Duration: 23 May 201628 May 2016
http://lrec2016.lrec-conf.org/en/
https://aclanthology.org/volumes/L16-1/

Conference

Conference 10th International Conference on Language Resources and Evaluation (LREC 2016)
PlaceSlovenia
CityPortorož
Period23/05/1628/05/16
Internet address

Research Keywords

  • Coreference
  • Corpus annotation
  • Direct speech

Fingerprint

Dive into the research topics of 'An Annotated Corpus of Direct Speech'. Together they form a unique fingerprint.

Cite this