Najnovejše

 corpus 
corpus
Opis:
The GORDAN 1.0 corpus contains authentic data of spoken communication, annotated for dialogue acts. This entry contains the complete audio files of the corpus (seven wav files, 1 hour of recording), and video files (four ...
 Ta vnos vsebuje 1 datoteko (1.87 GB).
 
Publicly Available Distributed under Creative Commons Attribution Required Noncommercial Share Alike
 corpus 
corpus
Avtor(ji):
Opis:
The GORDAN 1.0 corpus contains authentic data of spoken communication, annotated for dialogue acts according to the GORDAN 1.0 dialogue act annotation scheme, included in the data. The corpus data were selected from existing ...
 Ta vnos vsebuje 1 datoteko (479.09 KB).
 
Publicly Available Distributed under Creative Commons Attribution Required
 toolService 
toolService
Avtor(ji):
Opis:
The model for lemmatisation of standard Slovenian was built with the CLASSLA-StanfordNLP tool (https://github.com/clarinsi/classla-stanfordnlp) by training on the ssj500k training corpus (http://hdl.handle.net/11356/1210) ...
 Ta vnos vsebuje 1 datoteko (317.83 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required Share Alike

Največ ogledov

V preteklem tednu
 corpus 
corpus
Avtor(ji):
Opis:
The corpus contains 256,567 documents from the Slovenian news portals 24ur, Dnevnik, Finance, Rtvslo, and Žurnal24. These portals contain political, business, economic and financial content. The submission contains 7 files: ...
 Ta vnos vsebuje 8 datotek(e) (616.88 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required Share Alike
 lexicalConceptualResource 
lexicalConceptualResource
Opis:
srLex is a large inflectional lexicon of Serbian language where each entry consists of a (wordform, lemma, MSD, frequency, per-million frequency) 5-tuple. The (wordform, lemma, MSD) triple frequencies are calculated on the ...
 Ta vnos vsebuje 1 datoteko (29.54 MB).
 
Publicly Available
 corpus 
corpus
Avtor(ji):
Opis:
SentiCoref 1.0 corpus consists of 837 documents selected from SentiNews 1.0 corpus (http://hdl.handle.net/11356/1110). The documents were selected based on the number of automatically detected named entities (using Polyglot, ...
 Ta vnos vsebuje 2 datotek(e) (6.64 MB).
 
Publicly Available Distributed under Creative Commons Attribution Required