Title Eliciting and Annotating Uncertainty in Spoken Language
Authors Heather Pon-Barry, Stuart Shieber and Nicholas Longenbaugh
Abstract A major challenge in the field of automatic recognition of emotion and affect in speech is the subjective nature of affect labels. The most common approach to acquiring affect labels is to ask a panel of listeners to rate a corpus of spoken utterances along one or more dimensions of interest. For applications ranging from educational technology to voice search to dictation, a speaker's level of certainty is a primary dimension of interest. In such applications, we would like to know the speaker's actual level of certainty, but past research has only revealed listeners' perception of the speaker's level of certainty. In this paper, we present a method for eliciting spoken utterances using stimuli that we design such that they have a quantitative, crowdsourced legibility score. While we cannot control a speaker's actual internal level of certainty, the use of these stimuli provides a better estimate of internal certainty compared to existing speech corpora. The Harvard Uncertainty Speech Corpus, containing speech data, certainty annotations, and prosodic features, is made available to the research community.
Topics Emotion Recognition/Generation, Prosody
Full paper Eliciting and Annotating Uncertainty in Spoken Language
Bibtex @InProceedings{PONBARRY14.1167,
  author = {Heather Pon-Barry and Stuart Shieber and Nicholas Longenbaugh},
  title = {Eliciting and Annotating Uncertainty in Spoken Language},
  booktitle = {Proceedings of the Ninth International Conference on Language Resources and Evaluation (LREC'14)},
  year = {2014},
  month = {may},
  date = {26-31},
  address = {Reykjavik, Iceland},
  editor = {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Thierry Declerck and Hrafn Loftsson and Bente Maegaard and Joseph Mariani and Asuncion Moreno and Jan Odijk and Stelios Piperidis},
  publisher = {European Language Resources Association (ELRA)},
  isbn = {978-2-9517408-8-4},
  language = {english}
