Najnovejše

 corpus 
corpus
Opis:
This corpus contains the Q&A archive of the ŠUSS language consultancy service. The ŠUSS internet forum was active 1998-2010. Questions posted by users were answered by a group of students of linguistics and related ...
 Ta vnos vsebuje 2 datotek(e) (5.46 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required
 corpus 
corpus
Opis:
Janes-Tag is a manually annotated corpus of Slovene Computer-Mediated Communication (CMC). It is meant as a gold-standard training and testing dataset for tokenisation, sentence segmentation, word normalisation, morphosyntactic ...
 Ta vnos vsebuje 7 datotek(e) (5.68 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required Share Alike
 lexicalConceptualResource 
lexicalConceptualResource
Opis:
Aleks is a lexical database with 463 entries typical for general Slovene academic discourse. The entries include typical context examples (collocations and examples of use) taken from KAS, a corpus of Slovene academic ...
 Ta vnos vsebuje 1 datoteko (16.39 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required

Največ ogledov

V preteklem tednu
 lexicalConceptualResource 
lexicalConceptualResource
Opis:
srLex is a large inflectional lexicon of Serbian language where each entry consists of a (wordform, lemma, MSD, frequency, per-million frequency) 5-tuple. The (wordform, lemma, MSD) triple frequencies are calculated on the ...
 Ta vnos vsebuje 1 datoteko (29.54 MB).
 
Publicly Available
 lexicalConceptualResource 
lexicalConceptualResource
Opis:
The MULTEXT-East morphosyntactic lexicons have a simple structure, where each line is a lexical entry with three tab-separated fields: (1) the word-form, the inflected form of the word; (2) the lemma, the base-form of the ...
 Ta vnos vsebuje 6 datotek(e) (12.05 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required Noncommercial
 lexicalConceptualResource 
lexicalConceptualResource
Opis:
A lexicon of 751 emoji characters with automatically assigned sentiment. The sentiment is computed from 70,000 tweets, labeled by 83 human annotators in 13 European languages. The process and analysis of emoji sentiment ...
 Ta vnos vsebuje 3 datotek(e) (93.95 KB).
 
Publicly Available Distributed under Creative Commons Attribution Required Share Alike