Skip to main navigation Skip to search Skip to main content

Constructing large proposition databases

Peter Exner, Pierre Nugues

Research output: Chapter in Book/Report/Conference proceedingPaper in conference proceedingpeer-review

120 Downloads (Pure)

Abstract

With the advent of massive online encyclopedic corpora such as Wikipedia, it has become possible to apply a systematic analysis to a wide range of documents covering a significant part of human knowledge. Using semantic parsers, it has become possible to extract such knowledge in the form of propositions (predicate―argument structures) and build large proposition databases from these documents. This paper describes the creation of multilingual proposition databases using generic semantic dependency parsing. Using Wikipedia, we extracted, processed, clustered, and evaluated a large number of propositions. We built an architecture to provide a complete pipeline dealing with the input of text, extraction of knowledge, storage, and presentation of the resulting propositions
Original languageEnglish
Title of host publicationProceedings of the Eight International Conference on Language Resources and Evaluation (LREC'12)
PublisherEuropean Language Resources Association
Pages3836-3839
ISBN (Electronic) 978-2-9517408-7-7
Publication statusPublished - 2012
EventThe eighth international conference on Language Resources and Evaluation (LREC 2012) - Istanbul, Turkey
Duration: 2012 May 212012 May 27

Conference

ConferenceThe eighth international conference on Language Resources and Evaluation (LREC 2012)
Country/TerritoryTurkey
CityIstanbul
Period2012/05/212012/05/27

Subject classification (UKÄ)

  • Computer Sciences

Free keywords

  • Knowledge Discovery/Representation
  • Information Extraction
  • Information Retrieval
  • Semantics

Fingerprint

Dive into the research topics of 'Constructing large proposition databases'. Together they form a unique fingerprint.

Cite this