Yauti: A Tool for Morphosyntactic Analysis of Nheengatu within the Universal Dependencies Framework
Identifikátory výsledku
Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F23%3AMKMQ4RH6" target="_blank" >RIV/00216208:11320/23:MKMQ4RH6 - isvavai.cz</a>
Výsledek na webu
<a href="https://sol.sbc.org.br/index.php/stil/article/view/25445" target="_blank" >https://sol.sbc.org.br/index.php/stil/article/view/25445</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.5753/stil.2023.234131" target="_blank" >10.5753/stil.2023.234131</a>
Alternativní jazyky
Jazyk výsledku
—
Název v původním jazyce
Yauti: A Tool for Morphosyntactic Analysis of Nheengatu within the Universal Dependencies Framework
Popis výsledku v původním jazyce
"This paper reports on Yauti, a rule-based morphosyntactic analyzer for the endangered Brazilian indigenous language Nheengatu. Its goal is to generate analyses in the UD framework’s CoNLL-U format. It has been developed on par with the construction of the Nheengatu treebank of the UD collection. In sentences only consisting of known and unambiguous words, the tool generally delivers good results. It obtained a LAS score of 73.2% in a version of the Nheengatu UD treebank with all 1022 sentences automatically provided with XPOS tags and a special annotation to handle non-lexicalized words."
Název v anglickém jazyce
Yauti: A Tool for Morphosyntactic Analysis of Nheengatu within the Universal Dependencies Framework
Popis výsledku anglicky
"This paper reports on Yauti, a rule-based morphosyntactic analyzer for the endangered Brazilian indigenous language Nheengatu. Its goal is to generate analyses in the UD framework’s CoNLL-U format. It has been developed on par with the construction of the Nheengatu treebank of the UD collection. In sentences only consisting of known and unambiguous words, the tool generally delivers good results. It obtained a LAS score of 73.2% in a version of the Nheengatu UD treebank with all 1022 sentences automatically provided with XPOS tags and a special annotation to handle non-lexicalized words."
Klasifikace
Druh
D - Stať ve sborníku
CEP obor
—
OECD FORD obor
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)
Návaznosti výsledku
Projekt
—
Návaznosti
—
Ostatní
Rok uplatnění
2023
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Údaje specifické pro druh výsledku
Název statě ve sborníku
"Anais do XIV Simpósio Brasileiro de Tecnologia da Informação e da Linguagem Humana"
ISBN
—
ISSN
2175-6201
e-ISSN
—
Počet stran výsledku
11
Strana od-do
135-145
Název nakladatele
SBC
Místo vydání
Brasil
Místo konání akce
Brasil
Datum konání akce
1. 1. 2023
Typ akce podle státní příslušnosti
WRD - Celosvětová akce
Kód UT WoS článku
—