CLARIN-PL
Not a member yet
504 research outputs found
Sort by
Polish Spatial Texts (PST) 1.0
Texts derived from polish travel blogs manually annotated with spatial expressions, A spatial expression is a text fragment which describes a relative location of two or more physical objects to each other
Polish language sentiment dependency treebank (Treebank Wydzwieku)
The dataset is a dependency treebank with sentiment annotations
Extended dictionary of named entities NELexicon connected with Linked Open Data
This resource contains Polish named entities connected with terminology from available resources within Linked Open Data (e.g. WordNet, DBPedia, Wikipedia, etc.)
Emotional Annotations Dictionary
List of lexical units with emotional annotation extracted from Polish Wordnet (Słowosieć 4.0
Walenty (2018-06-29)
Walenty is a valence dictionary of Polish developed at the Institute of Computer Science, Polish Academy of Sciences (IPI PAN).
The original formalism of Walenty was established by Filip Skwarski, Elżbieta Hajnicz, Agnieszka Patejuk, Adam Przepiórkowski, Marcin Woliński, Marek Świdziński, and Magdalena Zawisławska. It has been further developed by Elżbieta Hajnicz, Agnieszka Patejuk, Adam Przepiórkowski, and Marcin Woliński. The semantic layer has been developed by Elżbieta Hajnicz and Anna Andrzejczuk.
The original seed of Walenty was provided by the automatic conversion, manually reviewed by Filip Skwarski, of the verbal valence dictionary used by the Świgra2 parser (6396 schemata for 1462 lemmata), which was in turn based on SDPV, the Syntactic Dictionary of Polish Verbs by Marek Świdziński (4148 schemata for 1064 lemmata). Afterwards, Walenty has been developed independently by adding new entries, syntactic schemata, in particular phraseological ones, and semantic frames.
Walenty has been edited and compiled using the Slowal tool (http://zil.ipipan.waw.pl/Slowal) created by Bartłomiej Nitoń and Tomasz Bartosiak.
The version of Walenty from 2018.06.29 contains 101 047 syntactic schemata and 28 321 semantic frames of 13022 verbs 4070 nouns, 950 adjectives and 200 nouns
Lexical Perspective on Wordnet to Wordnet Mapping
The paper presents a feature-based model of equivalence targeted at (manual) sense linking between Princeton WordNet and plWordNet. The model incorporates insights from lexicographic and translation theories on bilingual equivalence and draws on the results of earlier synsetlevel mapping of nouns between Princeton WordNet and plWordNet. It takes into account all basic aspects of language such as form, meaning and function and supplements them with (parallel) corpus frequency and translatability. Three types of equivalence are distinguished, namely strong, regular and weak depending on the conformity with the proposed features. The presented solutions are language neutral and they can be easily applied to language pairs other than Polish and English. Sense-level mapping is a more finegrained mapping than the existing synset mappings and is thus of great potential to human and machine translation