Artikel in Tagungsband INPROC-2008-150

Bibliograph.
Daten
Kassner, Laura; Nastase, Vivi; Strube, Michael: Acquiring a Taxonomy from the German Wikipedia.
In: Nicoletta Calzolari (Hrsg): Proceedings of the Sixth International Conference on Language Resources and Evaluation (LREC'08).
Universität Stuttgart, Fakultät Informatik, Elektrotechnik und Informationstechnik.
S. 1-4, englisch.
European Language Resources Association (ELRA), Mai 2008.
ISBN: 2-9517408-4-0.
Artikel in Tagungsband (Konferenz-Beitrag).
KörperschaftConference on Language Resources and Evaluation (LREC)
CR-Klassif.I.2.4 (Knowledge Representation Formalisms and Methods)
I.2.7 (Natural Language Processing)
Keywordstaxonomy; ontology; taxonomy generation; ontology generation; semantic network; Wikipedia; WordNet; GermaNet; multilinguality
Kurzfassung

This paper presents the process of acquiring a large, domain independent, taxonomy from the German Wikipedia. We build upon a previously implemented platform that extracts a semantic network and taxonomy from the English version of theWikipedia. We describe two accomplishments of our work: the semantic network for the German language in which isa links are identifed and annotated, and an expansion of the platform for easy adaptation for a new language. We identify the platform's strengths and shortcomings, which stem from the scarcity of free processing resources for languages other than English. We show that the taxonomy induction process is highly reliable - evaluated against the German version of WordNet, GermaNet, the resource obtained shows an accuracy of 83.34%.

Volltext und
andere Links
PDF (355462 Bytes)
LREC-Proceedings
Kontaktlaura.kassner@gsame.uni-stuttgart.de
Abteilung(en)Universität Stuttgart, Institut für Parallele und Verteilte Systeme, Anwendersoftware
Eingabedatum8. April 2013
   Publ. Institut   Publ. Informatik