• Repository
  • About
  • Contact
  • CLARIN
  •  Login
  • English Slovenščina
  • CLARIN.SI repository
  • Search
  • CLARIN logo
  •   Browse  
    •    All of the Repository  
      •   Issue Date
      •   Authors
      •   Titles
      •   Subjects
      •   Publisher
      •   Language
      •   Type
      •   Rights Label
  •   My Account  
    •    Login
  •   General Information  
    •    Deposit
    •    Cite
    •    Submission Lifecycle
    •    FAQ
    •    About
    •    Help Desk
 

 
Selected Filters
 Language : Slovenian     Clear All
Advanced Search

Filters

Use filters to refine the search results.

Current Filters:
New Filters:

Limit your search

Author  
    • Erjavec, Tomaž (91)
    • Ljubešić, Nikola (78)
    • Dobrovoljc, Kaja (67)
    • Krek, Simon (65)
    • Arhar Holdt, Špela (64)
    • Čibej, Jaka (53)
    • Kosem, Iztok (45)
    • Fišer, Darja (39)
    • Robnik-Šikonja, Marko (35)
    • Gantar, Polona (27)
    • Laskowski, Cyprian (23)
    • Terčon, Luka (22)
    • Pori, Eva (21)
    • Krsnik, Luka (20)
    • Klemenc, Bojan (17)
    • Kuzman, Taja (17)
    • Rupnik, Peter (16)
    • Pollak, Senja (15)
    • Ledinek, Nina (14)
    • Ferme, Marko (13)
    • ... View More
Subject  
    • TEI (62)
    • manual annotation (30)
    • lemmatisation (23)
    • language model (21)
    • part-of-speech tagging (21)
    • terminology (21)
    • computer-mediated communication (19)
    • dictionary (19)
    • spoken corpus (19)
    • multilingual (18)
    • lexicography (17)
    • tokenisation (16)
    • historical language (14)
    • parliamentary debates (14)
    • n-grams (13)
    • parsing (13)
    • collocations (12)
    • morphology (12)
    • named entities (12)
    • parallel corpus (12)
    • ... View More
Rights  
    • PUB (313)
    • ACA (19)
    • RES (1)
Language (ISO)  
    • English (60)
    • Croatian (33)
    • Hungarian (23)
    • Bulgarian (21)
    • German (21)
    • Serbian (21)
    • Spanish (20)
    • Dutch (17)
    • Estonian (17)
    • Portuguese (17)
    • French (16)
    • Russian (16)
    • Czech (15)
    • Danish (15)
    • Italian (15)
    • Polish (15)
    • Bosnian (14)
    • Latvian (14)
    • Swedish (14)
    • ... View More
Type  
    • text (297)
    • corpus (167)
    • lexicalConceptualResource (139)
    • toolService (47)
    • audio (10)
    • image (1)
    • languageDescription (1)
    • video (1)
Contain Files  
    • yes (333)
    • no (21)

Showing 1 through 20 out of 354 results

  • 1
  • 2
  • 3
  •  
  • 18
  •    
    • Sort items by
    •  Relevance
    • Title Asc
    • Title Desc
    • Issue Date Asc
    • Issue Date Desc
    •  
    • Results/page
    • 5
    • 10
    •  20
    • 40
    • 60
    • 80
    • 100

  • toolService
    CLARIN.SI data & tools
    toolService
    CroSloEngual BERT 1.1
    (Faculty of Computer and Information Science, University of Ljubljana / 2020-07-09)
    
    Author(s):
    Ulčar, Matej and Robnik-Šikonja, Marko
     This item contains 3 files (476.35 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required

  • toolService
    CLARIN.SI data & tools
    toolService
    Slovenian RoBERTa contextual embeddings model: SloBERTa 2.0
    (Faculty of Computer and Information Science, University of Ljubljana / 2021-01-17)
    
    Author(s):
    Ulčar, Matej and Robnik-Šikonja, Marko
     This item contains 2 files (1.29 GB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Morphological lexicon Sloleks 3.0
    (Centre for Language Resources and Technologies, University of Ljubljana / 2022-12-05)
    
    Author(s):
    Čibej, Jaka ; et al.show everyone Čibej, Jaka ; Gantar, Kaja ; Dobrovoljc, Kaja ; Krek, Simon ; Holozan, Peter ; Erjavec, Tomaž ; Romih, Miro ; Arhar Holdt, Špela ; Krsnik, Luka ; Robnik-Šikonja, Marko
     This item contains 1 file (239.75 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Parallel corpus EN-SL RSDO4 2.0
    (Centre for Language Resources and Technologies, University of Ljubljana / 2021-10-28)
    
    Author(s):
    Repar, Andraž and Lebar Bajec, Iztok
     This item contains 1 file (189.06 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Frequency lists of word parts from the GOS 1.0 corpus 1.1
    (Centre for Language Resources and Technologies, University of Ljubljana; Jožef Stefan Institute / 2020-10-28)
    
    Author(s):
    Čibej, Jaka ; Arhar Holdt, Špela ; Dobrovoljc, Kaja and Krek, Simon
     This item contains 1 file (33.41 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Frequency lists of words from the GOS 1.0 corpus 1.1
    (Centre for Language Resources and Technologies, University of Ljubljana; Jožef Stefan Institute / 2020-10-28)
    
    Author(s):
    Čibej, Jaka ; Arhar Holdt, Špela ; Dobrovoljc, Kaja and Krek, Simon
     This item contains 1 file (4.5 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Automatically stress labelled morphological lexicon Sloleks 1.2, version 1.1
    (Faculty of Computer and Information Science, University of Ljubljana; Centre for Language Resources and Technologies, University of Ljubljana / 2018-05-08)
    
    Author(s):
    Krsnik, Luka ; Robnik-Šikonja, Marko ; Šef, Tomaž and Krek, Simon
     This item contains 2 files (55.91 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Noncommercial Share Alike

  • toolService
    CLARIN.SI data & tools
    toolService
    The CLASSLA-Stanza model for lemmatisation of standard Slovenian 2.0
    (Jožef Stefan Institute / 2023-01-31)
    
    Author(s):
    Terčon, Luka ; Čibej, Jaka and Ljubešić, Nikola
     This item contains 1 file (2.09 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    ASR database ARTUR 1.0 (transcriptions)
    (Faculty of Electrical Engineering and Computer Science, University of Maribor; Faculty of Electrical Engineering, University of Ljubljana; Faculty of Computer and Information Science, University of Ljubljana; ZRC SAZU / 2023-02-22)
    
    Author(s):
    Verdonik, Darinka ; et al.show everyone Verdonik, Darinka ; Bizjak, Andreja ; Sepesy Maučec, Mirjam ; Gril, Lucija ; Dobrišek, Simon ; Križaj, Janez ; Strle, Gregor ; Bajec, Marko ; Lebar Bajec, Iztok ; Jelovšek, Tjaša ; Lokovšek, Jure ; Trojar, Mitja ; Erjavec, Tomaž ; Bernjak, Mitja ; Žganec Gros, Jerneja ; Čakš, Peter ; Pucer, Matevž ; Cvetko, Mitja ; Pavlič, Jani ; Zelenik, Marijana ; Ivanovska, Marija ; Grm, Klemen ; Longyka, Jure ; Mihelič, Aleš ; Vesnicer, Boštjan ; Dretnik, Naum
     This item contains 1 file (48.63 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Frequency lists of character-level n-grams from the GOS 1.0 corpus 1.1
    (Centre for Language Resources and Technologies, University of Ljubljana; Jožef Stefan Institute / 2020-10-28)
    
    Author(s):
    Čibej, Jaka ; Arhar Holdt, Špela ; Dobrovoljc, Kaja and Krek, Simon
     This item contains 1 file (2.56 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • toolService
    CLARIN.SI data & tools
    toolService
    The CLASSLA-Stanza model for morphosyntactic annotation of standard Slovenian 2.0
    (Jožef Stefan Institute / 2023-01-31)
    
    Author(s):
    Ljubešić, Nikola ; Terčon, Luka and Čibej, Jaka
     This item contains 2 files (509.87 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Kres corpus n-grams 2.0
    (Centre for Language Resources and Technologies, University of Ljubljana / 2018-08-03)
    
    Author(s):
    Dobrovoljc, Kaja
     This item contains 3 files (2.34 GB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Frequency lists of word-level n-grams from the GOS 1.0 corpus 1.1
    (Centre for Language Resources and Technologies, University of Ljubljana; Jožef Stefan Institute / 2020-10-28)
    
    Author(s):
    Čibej, Jaka ; Arhar Holdt, Špela ; Dobrovoljc, Kaja and Krek, Simon
     This item contains 3 files (287.52 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Frequency list of words from the Trendi corpus 2021
    (Jožef Stefan Institute / 2022-10-28)
    
    Author(s):
    Čibej, Jaka and Kosem, Iztok
     This item contains 1 file (25.06 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Gos corpus n-grams 2.0
    (Centre for Language Resources and Technologies, University of Ljubljana / 2018-08-03)
    
    Author(s):
    Dobrovoljc, Kaja
     This item contains 3 files (21.02 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • toolService
    CLARIN.SI data & tools
    toolService
    The CLASSLA-Stanza model for JOS dependency parsing of standard Slovenian 2.0
    (Jožef Stefan Institute / 2023-01-31)
    
    Author(s):
    Terčon, Luka and Ljubešić, Nikola
     This item contains 2 files (176.5 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • toolService
    CLARIN.SI data & tools
    toolService
    The CLASSLA-Stanza model for semantic role labeling of standard Slovenian 2.0
    (Jožef Stefan Institute / 2023-01-31)
    
    Author(s):
    Terčon, Luka and Ljubešić, Nikola
     This item contains 1 file (58.69 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Slovene corpus for general relation extraction SloREL 1.1
    (Faculty of Computer and Information Science, University of Ljubljana / 2022-09-15)
    
    Author(s):
    Štravs, Miha ; Knez, Timotej and Žitnik, Slavko
     This item contains 1 file (38.71 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required

  • lexicalConceptualResource
    CLARIN.SI data & tools
    lexicalConceptualResource
    Frequency lists of word-level n-grams from the Trendi corpus 2021
    (Jožef Stefan Institute / 2022-10-28)
    
    Author(s):
    Čibej, Jaka and Kosem, Iztok
     This item contains 1 file (1.03 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • toolService
    CLARIN.SI data & tools
    toolService
    The CLASSLA-Stanza model for morphosyntactic annotation of non-standard Slovenian 2.1
    (Jožef Stefan Institute / 2023-03-30)
    
    Author(s):
    Terčon, Luka and Ljubešić, Nikola
     This item contains 2 files (504.03 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • 1
  • 2
  • 3
  •  
  • 18
  •    
    • Sort items by
    •  Relevance
    • Title Asc
    • Title Desc
    • Issue Date Asc
    • Issue Date Desc
    •  
    • Results/page
    • 5
    • 10
    •  20
    • 40
    • 60
    • 80
    • 100
 

Partners

  • Alpineon, d.o.o.
  • Amebis, d.o.o.
  • Institute of Contemporary History
  • Jožef Stefan Institute
  • National and University Library of Slovenia
  • Slovenian Language Technologies Society

Partners

  • University of Ljubljana
  • University of Maribor
  • University of Nova Gorica
  • University of Primorska
  • ZRC SAZU
  • ZRS Koper

Repository

  • Main page
  • Contact
  • Submission Lifecycle
  • FAQ
  • About and Policies

This platform runs under the software developed for the LINDAT/CLARIAH-CZ repository for linguistics, available on GitHub

CLARIN.SI is supported by the Ministry of Education, Science and Sport of the Republic of Slovenia
under the Programme of "Research Infrastructures".