• Repository
  • About
  • Contact
  • CLARIN
  •  Login
  • English Slovenščina
  • CLARIN.SI repository
  • Search
  • CLARIN logo
  •   Browse  
    •    All of the Repository  
      •   Issue Date
      •   Authors
      •   Titles
      •   Subjects
      •   Publisher
      •   Language
      •   Type
      •   Rights Label
  •   My Account  
    •    Login
  •   General Information  
    •    Deposit
    •    Cite
    •    Submission Lifecycle
    •    FAQ
    •    About
    •    Help Desk
 

 
Selected Filters
 Author : Kuzman, Taja     Clear All
Advanced Search

Filters

Use filters to refine the search results.

Current Filters:
New Filters:

Limit your search

Author  
    • Ljubešić, Nikola (64)
    • Rupnik, Peter (56)
    • Bañón, Marta (44)
    • Esplà-Gomis, Miquel (44)
    • Forcada, Mikel L. (44)
    • García-Romero, Cristian (44)
    • Pla Sempere, Leopoldo (44)
    • Ramírez-Sánchez, Gema (44)
    • Suchomel, Vít (44)
    • Toral, Antonio (44)
    • van Noord, Rik (44)
    • Chichirau, Malina (28)
    • Galiano-Jiménez, Aarón (28)
    • Zaragoza-Bernabeu, Jaume (28)
    • van der Werff, Tobias (16)
    • Zaragoza, Jaume (16)
    • Erjavec, Tomaž (7)
    • Čibej, Jaka (7)
    • Dobrovoljc, Kaja (6)
    • ... View More
Subject  
    • web corpus (53)
    • multilingual (23)
    • parallel corpus (23)
    • automatic genre identification (11)
    • genre corpus (9)
    • manual annotation (8)
    • TEI (7)
    • CONLL-U (4)
    • dependency treebank (4)
    • named entities (4)
    • parsing (4)
    • part-of-speech tagging (4)
    • semantic role labelling (4)
    • tokenisation (4)
    • verbal multiword expressions (4)
    • Austrian Parliament (3)
    • Belgian Parliament (3)
    • Bosnian Parliament (3)
    • Bulgarian Parliament (3)
    • Catalonian Parliament (3)
    • ... View More
Language (ISO)  
    • English (28)
    • Slovenian (17)
    • Croatian (9)
    • Macedonian (8)
    • Bulgarian (6)
    • Icelandic (6)
    • Serbian (6)
    • Turkish (6)
    • Catalan (5)
    • Modern Greek (1453-) (5)
    • Albanian (4)
    • Bosnian (4)
    • Maltese (4)
    • Montenegrin (4)
    • Ukrainian (4)
    • Chakavian (1)
    • Dutch (1)
    • Spanish (1)
Type  
    • corpus (65)
    • text (65)
    • toolService (3)

Showing 1 through 10 out of 68 results

  • 1
  • 2
  • 3
  •  
  • 7
  •    
    • Sort items by
    • Relevance
    • Title Asc
    • Title Desc
    • Issue Date Asc
    •  Issue Date Desc
    •  
    • Results/page
    • 5
    •  10
    • 20
    • 40
    • 60
    • 80
    • 100

  • corpus
    CLARIN.SI data & tools
    corpus
    Multilingual IPTC Media Topic dataset EMMediaTopic 1.0
    (Jožef Stefan Institute / 2024-12-02)
    
    Author(s):
    Kuzman, Taja and Ljubešić, Nikola
     This item contains 1 file (71.3 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Genre-enriched web corpora MaCoCu-Genre
    (Jožef Stefan Institute / 2024-10-07)
    
    Author(s):
    Kuzman, Taja and Ljubešić, Nikola
     This item contains 14 files (101.43 GB).
     
    Publicly Available

  • corpus
    CLARIN.SI data & tools
    corpus
    English-Slovenian text genre dataset X-GENRE
    (Jožef Stefan Institute / 2024-09-25)
    
    Author(s):
    Kuzman, Taja and Ljubešić, Nikola
     This item contains 1 file (6.54 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • toolService
    CLARIN.SI data & tools
    toolService
    Multilingual text genre classification model X-GENRE
    (Jožef Stefan Institute / 2024-09-25)
    
    Author(s):
    Kuzman, Taja and Ljubešić, Nikola
     This item contains 1 file (779.93 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Training corpus SUK 1.1
    (Centre for Language Resources and Technologies, University of Ljubljana / 2024-08-22)
    
    Author(s):
    Arhar Holdt, Špela ; et al.show everyone Arhar Holdt, Špela ; Krek, Simon ; Dobrovoljc, Kaja ; Erjavec, Tomaž ; Gantar, Polona ; Čibej, Jaka ; Pori, Eva ; Terčon, Luka ; Munda, Tina ; Žitnik, Slavko ; Robida, Nejc ; Blagus, Neli ; Može, Sara ; Ledinek, Nina ; Holz, Nanika ; Zupan, Katja ; Kuzman, Taja ; Kavčič, Teja ; Škrjanec, Iza ; Marko, Dafne ; Jezeršek, Lucija ; Zajc, Anja
     This item contains 2 files (45.1 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Linguistically annotated multilingual comparable corpora of parliamentary debates in English ParlaMint-en.ana 4.1
    (CLARIN ERIC / 2024-06-03)
    
    Author(s):
    Kuzman, Taja ; et al.show everyone Kuzman, Taja ; Ljubešić, Nikola ; Erjavec, Tomaž ; Kopp, Matyáš ; Ogrodniczuk, Maciej ; Osenova, Petya ; Rayson, Paul ; Vidler, John ; Agerri, Rodrigo ; Agirrezabal, Manex ; Agnoloni, Tommaso ; Aires, José ; Albini, Monica ; Alkorta, Jon ; Antiba-Cartazo, Iván ; Arrieta, Ekain ; Barcala, Mario ; Bardanca, Daniel ; Barkarson, Starkaður ; Bartolini, Roberto ; Battistoni, Roberto ; Bel, Nuria ; Bonet Ramos, Maria del Mar ; Calzada Pérez, María ; Cardoso, Aida ; Çöltekin, Çağrı ; Coole, Matthew ; Darģis, Roberts ; de Does, Jesse ; de Libano, Ruben ; Depoorter, Griet ; Depuydt, Katrien ; Diwersy, Sascha ; Dodé, Réka ; Fernandez, Kike ; Fernández Rei, Elisa ; Frontini, Francesca ; Garcia, Marcos ; García Díaz, Noelia ; García Louzao, Pedro ; Gavriilidou, Maria ; Gkoumas, Dimitris ; Grigorov, Ilko ; Grigorova, Vladislava ; Haltrup Hansen, Dorte ; Iruskieta, Mikel ; Jarlbrink, Johan ; Jelencsik-Mátyus, Kinga ; Jongejan, Bart ; Kahusk, Neeme ; Kirnbauer, Martin ; Kryvenko, Anna ; Ligeti-Nagy, Noémi ; Luxardo, Giancarlo ; Magariños, Carmen ; Magnusson, Måns ; Marchetti, Carlo ; Marx, Maarten ; Meden, Katja ; Mendes, Amália ; Mochtak, Michal ; Mölder, Martin ; Montemagni, Simonetta ; Navarretta, Costanza ; Nitoń, Bartłomiej ; Norén, Fredrik Mohammadi ; Nwadukwe, Amanda ; Ojsteršek, Mihael ; Pančur, Andrej ; Papavassiliou, Vassilis ; Pereira, Rui ; Pérez Lago, María ; Piperidis, Stelios ; Pirker, Hannes ; Pisani, Marilina ; Pol, Henk van der ; Prokopidis, Prokopis ; Quochi, Valeria ; Regueira, Xosé Luís ; Rii, Andriana ; Rudolf, Michał ; Ruisi, Manuela ; Rupnik, Peter ; Schopper, Daniel ; Simov, Kiril ; Sinikallio, Laura ; Skubic, Jure ; Tamper, Minna ; Tungland, Lars Magne ; Tuominen, Jouni ; van Heusden, Ruben ; Varga, Zsófia ; Vázquez Abuín, Marta ; Venturi, Giulia ; Vidal Miguéns, Adrián ; Vider, Kadri ; Vivel Couso, Ainhoa ; Vladu, Adina Ioana ; Wissik, Tanja ; Yrjänäinen, Väinö ; Zevallos, Rodolfo ; Fišer, Darja
     This item contains 31 files (53.36 GB).
     
    Publicly Available Distributed under Creative Commons Attribution Required

  • corpus
    CLARIN.SI data & tools
    corpus
    "Choice of plausible alternatives" datasets in South Slavic dialects DIALECT-COPA
    (Jožef Stefan Institute / 2024-04-26)
    
    Author(s):
    Ljubešić, Nikola ; et al.show everyone Ljubešić, Nikola ; Kuzman, Taja ; Rupnik, Peter ; Milosavljević, Stefan ; Galant, Nada ; Benčina, Sonja ; Čibej, Jaka
     This item contains 6 files (279.69 KB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Montenegrin web corpus CLASSLA-web.cnr 1.0
    (Jožef Stefan Institute / 2024-03-26)
    
    Author(s):
    Ljubešić, Nikola ; Rupnik, Peter and Kuzman, Taja
     This item contains 2 files (1.4 GB).
     
    Publicly Available

  • corpus
    CLARIN.SI data & tools
    corpus
    Bulgarian web corpus CLASSLA-web.bg 1.0
    (Jožef Stefan Institute / 2024-03-26)
    
    Author(s):
    Ljubešić, Nikola ; Rupnik, Peter and Kuzman, Taja
     This item contains 2 files (32.1 GB).
     
    Publicly Available

  • corpus
    CLARIN.SI data & tools
    corpus
    Serbian web corpus CLASSLA-web.sr 1.0
    (Jožef Stefan Institute / 2024-03-26)
    
    Author(s):
    Ljubešić, Nikola ; Rupnik, Peter and Kuzman, Taja
     This item contains 2 files (21.58 GB).
     
    Publicly Available

  • 1
  • 2
  • 3
  •  
  • 7
  •    
    • Sort items by
    • Relevance
    • Title Asc
    • Title Desc
    • Issue Date Asc
    •  Issue Date Desc
    •  
    • Results/page
    • 5
    •  10
    • 20
    • 40
    • 60
    • 80
    • 100
 

Partners

  • Alpineon, d.o.o.
  • Amebis, d.o.o.
  • Institute of Contemporary History
  • Jožef Stefan Institute
  • National and University Library of Slovenia
  • Slovenian Language Technologies Society

Partners

  • University of Ljubljana
  • University of Maribor
  • University of Nova Gorica
  • University of Primorska
  • ZRC SAZU
  • ZRS Koper

Repository

  • Main page
  • Contact
  • Submission Lifecycle
  • FAQ
  • About and Policies

This platform runs under the software developed for the LINDAT/CLARIAH-CZ repository for linguistics, available on GitHub

CLARIN.SI is supported by the Ministry of Education, Science and Sport of the Republic of Slovenia
under the Programme of "Research Infrastructures".