What's New
toolService

Description:
STARK is a highly customizable tool designed for extracting different types of syntactic structures (trees) from parsed corpora (treebanks), aimed at corpus-driven linguistic investigations of syntactic and lexical phenomena ...
Ta vnos vsebuje 1 datoteko (3.17
MB).
Publicly Available
corpus

Description:
The Frenk-MRW dataset contains French and Slovene socially unacceptable Facebook comments that are manually annotated for metaphor and metonymy based on the observed incongruity between the basic and contextual meaning. ...
Ta vnos vsebuje 1 datoteko (1.82
MB).
Academic Use



lexicalConceptualResource

Description:
ILS is a dataset containing Slovene word forms containing a single lC bigram, i.e. an "l" grapheme preceding a consonant grapheme (a bigram of "l"+C(onsonant) = lC bigram). This combination is one of the less predictable ...
Ta vnos vsebuje 1 datoteko (1.05
MB).
Publicly Available



Največ ogledov
V preteklem tednu
lexicalConceptualResource

Description:
A lexicon of 751 emoji characters with automatically assigned sentiment.
The sentiment is computed from 70,000 tweets, labeled by 83 human annotators
in 13 European languages.
The process and analysis of emoji sentiment ...
Ta vnos vsebuje 3 datotek(e) (93.95
KB).
Publicly Available



toolService

Description:
This Conformer CTC BPE E2E Automated Speech Recognition model was trained following the NVIDIA NeMo Conformer-CTC fine-tuning recipe (for details see the official NVIDIA NeMo NMT documentation, https://docs.nvidia.com/de ...
Ta vnos vsebuje 1 datoteko (430.87
MB).
Publicly Available
lexicalConceptualResource

Description:
A list of headwords from the collection "Besede slovenskega jezika" (Words of Slovenian Language).
Ta vnos vsebuje 1 datoteko (997.48
KB).
Publicly Available


