Titel: PoliMorf: a (not so) new open morphological dictionary for Polish
Personen:Woliński, Marcin/Miłkowski, Marcin/Ogrodniczuk, Maciej/Przepiórkowski, Adam/Szałkiewicz, Łukasz
Jahr: 2012
Typ: Aufsatz
Verlag: European Language Resources Association (ELRA)
Ortsangabe: Istanbul
In: Calzolari, Nicoletta/Choukri, Khalid/Declerck, Thierry/Doğan, Mehmet U./Maegaard, Bente/Mariani, Joseph/Odijk, Jan/Piperidis, Stelios (Hgg.): Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC 2012), Istanbul, 23 - 25 May 2012
Seiten: 860-864
Untersuchte Sprachen: Polnisch*Polish
Schlagwörter: Datenbank*data base
Datenmodellierung*data modelling
Grammatik im Wörterbuch*grammar in dictionaries
Internet-Lexikografie/Online-Lexikografie*internet lexicography/online lexicography
Wortbildung im Wörterbuch*word formation in dictionaries
Medium: Online
URI: http://www.lrec-conf.org/proceedings/lrec2012/pdf/263_Paper.pdf
Zuletzt besucht: 17.09.2018
Abstract: This paper presents preliminary results of an effort aiming at the creation of a morphological dictionary of Polish, PoliMorf, available under a very liberal BSD-style license. The dictionary is a result of a merger of two existing resources, SGJP and Morfologik and was prepared within the CESAR/META-NET initiative. The work completed so far includes re-licensing of the two dictionaries and filling the new resource with the morphological data semi-automatically unified from both sources. The merging process is controlled by the collaborative dictionary development web application Kuźnia, also implemented within the project. The tool involves several advanced features such as using SGJP inflectional patterns for form generation, possibility of attaching dictionary labels and classification schemes to lexemes, dictionary source record and change tracking. Since SGJP and Morfologik are already used in a significant number of Natural Language Processing projects in Poland, we expect PoliMorf to become the Polish morphological dictionary of choice for many years to come.