504 research outputs found

    Polish Spatial Texts (PST) 1.0

    No full text
    Texts derived from polish travel blogs manually annotated with spatial expressions, A spatial expression is a text fragment which describes a relative location of two or more physical objects to each other

    Polish language sentiment dependency treebank (Treebank Wydzwieku)

    No full text
    The dataset is a dependency treebank with sentiment annotations

    SpakowanesermonyEN

    No full text
    Sermon

    POE: Microcorpus of 20th century Polish poetry

    No full text
    Microcorpus of 20th century Polish poetr

    Korpus - specyfikacje

    No full text
    n/

    Extended dictionary of named entities NELexicon connected with Linked Open Data

    No full text
    This resource contains Polish named entities connected with terminology from available resources within Linked Open Data (e.g. WordNet, DBPedia, Wikipedia, etc.)

    Polish-Russian Parallel Corpus

    No full text
    Polish-Russian Parallel Corpu

    Emotional Annotations Dictionary

    No full text
    List of lexical units with emotional annotation extracted from Polish Wordnet (Słowosieć 4.0

    Walenty (2018-06-29)

    No full text
    Walenty is a valence dictionary of Polish developed at the Institute of Computer Science, Polish Academy of Sciences (IPI PAN). The original formalism of Walenty was established by Filip Skwarski, Elżbieta Hajnicz, Agnieszka Patejuk, Adam Przepiórkowski, Marcin Woliński, Marek Świdziński, and Magdalena Zawisławska. It has been further developed by Elżbieta Hajnicz, Agnieszka Patejuk, Adam Przepiórkowski, and Marcin Woliński. The semantic layer has been developed by Elżbieta Hajnicz and Anna Andrzejczuk. The original seed of Walenty was provided by the automatic conversion, manually reviewed by Filip Skwarski, of the verbal valence dictionary used by the Świgra2 parser (6396 schemata for 1462 lemmata), which was in turn based on SDPV, the Syntactic Dictionary of Polish Verbs by Marek Świdziński (4148 schemata for 1064 lemmata). Afterwards, Walenty has been developed independently by adding new entries, syntactic schemata, in particular phraseological ones, and semantic frames. Walenty has been edited and compiled using the Slowal tool (http://zil.ipipan.waw.pl/Slowal) created by Bartłomiej Nitoń and Tomasz Bartosiak. The version of Walenty from 2018.06.29 contains 101 047 syntactic schemata and 28 321 semantic frames of 13022 verbs 4070 nouns, 950 adjectives and 200 nouns

    Lexical Perspective on Wordnet to Wordnet Mapping

    Get PDF
    The paper presents a feature-based model of equivalence targeted at (manual) sense linking between Princeton WordNet and plWordNet. The model incorporates insights from lexicographic and translation theories on bilingual equivalence and draws on the results of earlier synsetlevel mapping of nouns between Princeton WordNet and plWordNet. It takes into account all basic aspects of language such as form, meaning and function and supplements them with (parallel) corpus frequency and translatability. Three types of equivalence are distinguished, namely strong, regular and weak depending on the conformity with the proposed features. The presented solutions are language neutral and they can be easily applied to language pairs other than Polish and English. Sense-level mapping is a more finegrained mapping than the existing synset mappings and is thus of great potential to human and machine translation

    40

    full texts

    504

    metadata records
    Updated in last 30 days.
    CLARIN-PL
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇