The CLASSLA-Stanza model for UD dependency parsing of spoken Slovenian 2.2
The result's identifiers
Result code in IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F26%3AQHZYBKQW" target="_blank" >RIV/00216208:11320/26:QHZYBKQW - isvavai.cz</a>
Result on the web
<a href="https://clarin.si/repository/xmlui/handle/11356/2018" target="_blank" >https://clarin.si/repository/xmlui/handle/11356/2018</a>
DOI - Digital Object Identifier
—
Alternative languages
Result language
angličtina
Original language name
The CLASSLA-Stanza model for UD dependency parsing of spoken Slovenian 2.2
Original language description
This model for UD dependency parsing of spoken Slovenian was built with the CLASSLA-Stanza tool (https://github.com/clarinsi/classla) by training on the SST treebank of spoken Slovenian (https://github.com/UniversalDependencies/UD_Slovenian-SST) combined with the SUK training corpus (http://hdl.handle.net/11356/1959) and using the CLARIN.SI-embed.sl word embeddings (http://hdl.handle.net/11356/1791) that were expanded with the MaCoCu-sl Slovene web corpus (http://hdl.handle.net/11356/1517). The estimated LAS of the parser is ~81.91.
Czech name
—
Czech description
—
Classification
Type
O - Miscellaneous
CEP classification
—
OECD FORD branch
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)
Result continuities
Project
—
Continuities
—
Others
Publication year
2025
Confidentiality
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů