• Repository
  • About
  • Contact
  • CLARIN
  •  Login
  • English Slovenščina
  • CLARIN.SI repository
  • Search
  • CLARIN logo
  •   Browse  
    •    All of the Repository  
      •   Issue Date
      •   Authors
      •   Titles
      •   Subjects
      •   Publisher
      •   Language
      •   Type
      •   Rights Label
  •   My Account  
    •    Login
  •   General Information  
    •    Deposit
    •    Cite
    •    Submission Lifecycle
    •    FAQ
    •    About
    •    Help Desk
 

 
Selected Filters
 Type : corpus     Clear All
Advanced Search

Filters

Use filters to refine the search results.

Current Filters:
New Filters:

Limit your search

Author  
    • Ljubešić, Nikola (138)
    • Erjavec, Tomaž (91)
    • Rupnik, Peter (70)
    • Kuzman, Taja (65)
    • Toral, Antonio (50)
    • Esplà-Gomis, Miquel (49)
    • Bañón, Marta (44)
    • Forcada, Mikel L. (44)
    • García-Romero, Cristian (44)
    • Pla Sempere, Leopoldo (44)
    • Ramírez-Sánchez, Gema (44)
    • Suchomel, Vít (44)
    • van Noord, Rik (44)
    • Fišer, Darja (37)
    • Chichirau, Malina (28)
    • Galiano-Jiménez, Aarón (28)
    • Zaragoza-Bernabeu, Jaume (28)
    • Arhar Holdt, Špela (25)
    • Krek, Simon (22)
    • Batanović, Vuk (18)
    • ... View More
Subject  
    • TEI (67)
    • web corpus (64)
    • parallel corpus (44)
    • manual annotation (41)
    • multilingual (39)
    • computer-mediated communication (27)
    • named entities (21)
    • parliamentary debates (21)
    • news corpus (20)
    • part-of-speech tagging (20)
    • tokenisation (18)
    • lemmatisation (16)
    • spoken corpus (16)
    • Parla-CLARIN (14)
    • word normalisation (14)
    • news comments (13)
    • Slovenian Parliament (13)
    • specialised corpus (12)
    • speech database (12)
    • speech transcription (12)
    • ... View More
Rights  
    • PUB (262)
    • ACA (24)
    • RES (1)
Language (ISO)  
    • Slovenian (165)
    • English (75)
    • Croatian (52)
    • Serbian (45)
    • Bosnian (22)
    • Bulgarian (20)
    • Spanish (16)
    • Estonian (15)
    • Russian (14)
    • French (13)
    • Hungarian (13)
    • Macedonian (13)
    • Dutch (12)
    • German (12)
    • Italian (12)
    • Czech (11)
    • Danish (11)
    • Icelandic (11)
    • Latvian (11)
    • Montenegrin (11)
    • ... View More
Type  
    • text (284)
    • audio (15)
    • video (1)
Contain Files  
    • yes (287)
    • no (13)

Showing 1 through 10 out of 300 results

  • 1
  • 2
  • 3
  •  
  • 30
  •    
    • Sort items by
    • Relevance
    •  Title Asc
    • Title Desc
    • Issue Date Asc
    • Issue Date Desc
    •  
    • Results/page
    • 5
    •  10
    • 20
    • 40
    • 60
    • 80
    • 100

  • corpus
    CLARIN.SI data & tools
    corpus
    "Choice of plausible alternatives" datasets in South Slavic dialects DIALECT-COPA
    (Jožef Stefan Institute / 2024-04-26)
    
    Author(s):
    Ljubešić, Nikola ; et al.show everyone Ljubešić, Nikola ; Kuzman, Taja ; Rupnik, Peter ; Milosavljević, Stefan ; Galant, Nada ; Benčina, Sonja ; Čibej, Jaka
     This item contains 6 files (279.69 KB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    24sata news article archive 1.0
    (Styria Media Group / 2021-04-19)
    
    Author(s):
    Purver, Matthew ; Shekhar, Ravi ; Pranjić, Marko ; Pollak, Senja and Martinc, Matej
     This item contains 2 files (1.26 GB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Noncommercial No Derivative Works

  • corpus
    CLARIN.SI data & tools
    corpus
    24sata news comment dataset 1.0
    (Styria Media Group / 2021-04-19)
    
    Author(s):
    Shekhar, Ravi ; Pranjic, Marko ; Pollak, Senja ; Pelicon, Andraž and Purver, Matthew
     This item contains 3 files (1.89 GB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Noncommercial No Derivative Works

  • corpus
    CLARIN.SI data & tools
    corpus
    Abstracts from the KAS corpus KAS-Abs 2.0
    (Faculty of Electrical Engineering and Computer Science, University of Maribor; Faculty of Computer and Information Science, University of Ljubljana / 2022-02-04)
    
    Author(s):
    Žagar, Aleš ; et al.show everyone Žagar, Aleš ; Kavaš, Matic ; Robnik-Šikonja, Marko ; Erjavec, Tomaž ; Fišer, Darja ; Ljubešić, Nikola ; Ferme, Marko ; Borovič, Mladen ; Boškovič, Borko ; Ojsteršek, Milan ; Hrovat, Goran
     This item contains 1 file (83.48 MB).
     
    Academic Use Inform Before Use Attribution Required Noncommercial

  • corpus
    CLARIN.SI data & tools
    corpus
    Albanian Spoken Corpus in Kosovo 1.0
    (University of Prishtina "Hasan Prishtina" / 2024-07-08)
    
    Author(s):
    Wasserscheidt, Philipp ; Rugova, Bardh and Baftiu, Adelajda
     This item contains 1 file (1.76 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required

  • corpus
    CLARIN.SI data & tools
    corpus
    Albanian web corpus MaCoCu-sq 1.0
    (Jožef Stefan Institute; Prompsit; Rijksuniversiteit Groningen; Universitat d'Alacant / 2023-04-20)
    
    Author(s):
    Bañón, Marta ; et al.show everyone Bañón, Marta ; Chichirau, Malina ; Esplà-Gomis, Miquel ; Forcada, Mikel L. ; Galiano-Jiménez, Aarón ; García-Romero, Cristian ; Kuzman, Taja ; Ljubešić, Nikola ; van Noord, Rik ; Pla Sempere, Leopoldo ; Ramírez-Sánchez, Gema ; Rupnik, Peter ; Suchomel, Vít ; Toral, Antonio ; Zaragoza-Bernabeu, Jaume
     This item contains 2 files (1.63 GB).
     
    Publicly Available

  • corpus
    CLARIN.SI data & tools
    corpus
    Albanian-English parallel corpus MaCoCu-sq-en 1.0
    (Jožef Stefan Institute; Prompsit; Rijksuniversiteit Groningen; Universitat d'Alacant / 2023-04-26)
    
    Author(s):
    Bañón, Marta ; et al.show everyone Bañón, Marta ; Chichirau, Malina ; Esplà-Gomis, Miquel ; Forcada, Mikel L. ; Galiano-Jiménez, Aarón ; García-Romero, Cristian ; Kuzman, Taja ; Ljubešić, Nikola ; van Noord, Rik ; Pla Sempere, Leopoldo ; Ramírez-Sánchez, Gema ; Rupnik, Peter ; Suchomel, Vít ; Toral, Antonio ; Zaragoza-Bernabeu, Jaume
     This item contains 3 files (590.81 MB).
     
    Publicly Available

  • corpus
    CLARIN.SI data & tools
    corpus
    Annotated corpus of Croatian language-related news articles MetaLangNEWS-Hr
    (ZRC SAZU; Regional Linguistic Data Initiative Centre ReLDI / 2020-10-30)
    
    Author(s):
    Bogetić, Ksenija and Batanović, Vuk
     This item contains 3 files (11.71 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Noncommercial Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Annotated corpus of Croatian language-related news comments MetaLangNEWS-COMMENTS-Hr
    (ZRC SAZU; Regional Linguistic Data Initiative Centre ReLDI / 2020-10-30)
    
    Author(s):
    Bogetić, Ksenija and Batanović, Vuk
     This item contains 3 files (14.76 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Noncommercial Share Alike

  • corpus
    CLARIN.SI data & tools
    corpus
    Annotated corpus of Macedonian language-related news articles MetaLangNEWS-Mk
    (ZRC SAZU; Regional Linguistic Data Initiative Centre ReLDI / 2022-07-27)
    
    Author(s):
    Bogetić, Ksenija ; Radošević, Petar and Batanović, Vuk
     This item contains 3 files (2.96 MB).
     
    Publicly Available Distributed under Creative Commons Attribution Required Noncommercial Share Alike

  • 1
  • 2
  • 3
  •  
  • 30
  •    
    • Sort items by
    • Relevance
    •  Title Asc
    • Title Desc
    • Issue Date Asc
    • Issue Date Desc
    •  
    • Results/page
    • 5
    •  10
    • 20
    • 40
    • 60
    • 80
    • 100
 

Partners

  • Alpineon, d.o.o.
  • Amebis, d.o.o.
  • Institute of Contemporary History
  • Jožef Stefan Institute
  • National and University Library of Slovenia
  • Slovenian Language Technologies Society

Partners

  • University of Ljubljana
  • University of Maribor
  • University of Nova Gorica
  • University of Primorska
  • ZRC SAZU
  • ZRS Koper

Repository

  • Main page
  • Contact
  • Submission Lifecycle
  • FAQ
  • About and Policies

This platform runs under the software developed for the LINDAT/CLARIAH-CZ repository for linguistics, available on GitHub

CLARIN.SI is supported by the Ministry of Education, Science and Sport of the Republic of Slovenia
under the Programme of "Research Infrastructures".