What's New
toolService
Description:
This resource provides automatic speech recognition (ASR) models for the Slovenian Gail Valley dialect, developed using the open-source WeNet 3.1 framework. The models were trained on 84.25 hours of speech from the Corpus ...
Ta vnos vsebuje 1 datoteko (523.03
MB).
Publicly Available
corpus
Description:
This entry contains the first part of the audiobook "Beli Rom" (The white Roma) by author Sandi Horvat (COBISS ID: 283921923, ISBN: 978-961-7198-98-0).
We are all children of the stars. Every single one of us. Some say ...
Ta vnos vsebuje 9 datotek(e) (152.72
MB).
Publicly Available
corpus
Description:
This entry contains the first part of the audiobook "Skrivnost glasbene sole" (The secret of music school) by author Marija Rupnik (COBISS ID: 282019843, ISBN: 978-961-7198-96-6).
A story about love, empathy and ...
Ta vnos vsebuje 4 datotek(e) (169.76
MB).
Publicly Available
Največ ogledov
V preteklem tednu
corpus
Description:
ParlaMint 5.0 is a set of comparable corpora containing transcriptions of parliamentary debates of 29 European countries and autonomous regions, mostly starting in 2015 and extending to mid-2022. The individual corpora ...
Ta vnos vsebuje 31 datotek(e) (5.94
GB).
Publicly Available
lexicalConceptualResource
Description:
Frazeološki rječnik hrvatskoga jezika is an open-access dictionary of Croatian idioms based on data from a large electronic corpus. The resulting dictionary will serve as a gateway for a large number of users and researchers ...
Ta vnos ne vsebuje datotek.
corpus
Description:
goo300k is a manually annotated reference corpus of historical Slovene. It contains 1,100 pages (about 300,000 tokens) sampled from 89 texts from the period 1584-1899.
Each text contains extensive meta-data and per-page ...
Ta vnos vsebuje 2 datotek(e) (8.9
MB).
Publicly Available