Summary of the paper

Title Extraction of Informative Expressions from Domain-specific Documents
Authors Eiko Yamamoto, Hitoshi Isahara, Akira Terada and Yasunori Abe
Abstract What kinds of lexical resources are helpful for extracting useful information from domain-specific documents? Although domain-specific documents contain much useful knowledge, it is not obvious how to extract such knowledge efficiently from the documents. We need to develop techniques for extracting hidden information from such domain-specific documents. These techniques do not necessarily use state-of-the-art technologies and achieve deep and accurate language understanding, but are based on huge amounts of linguistic resources, such as domain-specific lexical databases. In this paper, we introduce two techniques for extracting informative expressions from documents: the extraction of related words that are not only taxonomically related but also thematically related, and the acquisition of salient terms and phrases. With these techniques we then attempt to automatically and statistically extract domain-specific informative expressions in aviation documents as an example and evaluate the results.
Language Multiple languages
Topics Information Extraction, Information Retrieval, Statistical methods, Text mining
Full paper Extraction of Informative Expressions from Domain-specific Documents
Slides -
Bibtex @InProceedings{YAMAMOTO08.410,
  author = {Eiko Yamamoto, Hitoshi Isahara, Akira Terada and Yasunori Abe},
  title = {Extraction of Informative Expressions from Domain-specific Documents},
  booktitle = {Proceedings of the Sixth International Conference on Language Resources and Evaluation (LREC'08)},
  year = {2008},
  month = {may},
  date = {28-30},
  address = {Marrakech, Morocco},
  editor = {Nicoletta Calzolari (Conference Chair), Khalid Choukri, Bente Maegaard, Joseph Mariani, Jan Odijk, Stelios Piperidis, Daniel Tapias},
  publisher = {European Language Resources Association (ELRA)},
  isbn = {2-9517408-4-0},
  note = {http://www.lrec-conf.org/proceedings/lrec2008/},
  language = {english}
  }

Powered by ELDA © 2008 ELDA/ELRA