Gold Data and Multiple Understanding of Discourse Relations
Identifikátory výsledku
Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F25%3A10511671" target="_blank" >RIV/00216208:11320/25:10511671 - isvavai.cz</a>
Výsledek na webu
<a href="http://dx.doi.org/10.1007/978-3-032-02551-7_22" target="_blank" >http://dx.doi.org/10.1007/978-3-032-02551-7_22</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.1007/978-3-032-02551-7_22" target="_blank" >10.1007/978-3-032-02551-7_22</a>
Alternativní jazyky
Jazyk výsledku
angličtina
Název v původním jazyce
Gold Data and Multiple Understanding of Discourse Relations
Popis výsledku v původním jazyce
Discourse relations represent a relatively ambiguous area of language that can be a challenge for computational and corpus linguis-tics. In our study, we examine the reliability of the Prague Dependency Treebank – Consolidated 2.0 by employing multiple annotations that can reveal possible different readings. It turns out that the complexity of sentence structure has a very significant influence on the distinction of the left discourse argument; in contrast, neither the mode of the text (written vs. spoken) nor the presence of attributive constructions appears to influence the variability in the interpretation of discourse structure. Furthermore, we identify the technical organization of the annotation process itself as a potential source of bias in discourse annotation.
Název v anglickém jazyce
Gold Data and Multiple Understanding of Discourse Relations
Popis výsledku anglicky
Discourse relations represent a relatively ambiguous area of language that can be a challenge for computational and corpus linguis-tics. In our study, we examine the reliability of the Prague Dependency Treebank – Consolidated 2.0 by employing multiple annotations that can reveal possible different readings. It turns out that the complexity of sentence structure has a very significant influence on the distinction of the left discourse argument; in contrast, neither the mode of the text (written vs. spoken) nor the presence of attributive constructions appears to influence the variability in the interpretation of discourse structure. Furthermore, we identify the technical organization of the annotation process itself as a potential source of bias in discourse annotation.
Klasifikace
Druh
D - Stať ve sborníku
CEP obor
—
OECD FORD obor
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)
Návaznosti výsledku
Projekt
<a href="/cs/project/GA24-11132S" target="_blank" >GA24-11132S: Neshoda v korpusové anotaci ve vztahu k víceznačnosti textu</a><br>
Návaznosti
P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)
Ostatní
Rok uplatnění
2025
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Údaje specifické pro druh výsledku
Název statě ve sborníku
28th International Conference on Text, Speech and Dialogue (Part II)
ISBN
978-3-032-02551-7
ISSN
—
e-ISSN
—
Počet stran výsledku
13
Strana od-do
250-262
Název nakladatele
Springer
Místo vydání
Cham, Switzerland
Místo konání akce
Erlangen, Germany
Datum konání akce
25. 8. 2025
Typ akce podle státní příslušnosti
WRD - Celosvětová akce
Kód UT WoS článku
—