Multiword Discourse Markers Across Languages: A Linguistic and Computational Perspective
Identifikátory výsledku
Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F26%3AHR5NSHZF" target="_blank" >RIV/00216208:11320/26:HR5NSHZF - isvavai.cz</a>
Výsledek na webu
<a href="https://onlinelibrary.wiley.com/doi/10.1111/ijal.12755" target="_blank" >https://onlinelibrary.wiley.com/doi/10.1111/ijal.12755</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.1111/ijal.12755" target="_blank" >10.1111/ijal.12755</a>
Alternativní jazyky
Jazyk výsledku
angličtina
Název v původním jazyce
Multiword Discourse Markers Across Languages: A Linguistic and Computational Perspective
Popis výsledku v původním jazyce
Discourse markers (DMs) are linguistic expressions that convey different semantic and pragmatic values, managing and organizing the structure of spoken and written discourses. They can be either single‐word or multiword expressions (MWE), made up of conjunctions, adverbs, and prepositional phrases. Although DMs are the focus of many studies, some questions regarding the interoperability of taxonomies and automatic identification and classification require further research. We aim to tackle these issues by offering a critical analysis and discussing the constitution of a multilingual corpus in 10 languages, i.e., English, Lithuanian, Bulgarian, German, Macedonian, Romanian, Hebrew, Polish, European Portuguese, and Italian. The novel two‐level annotation approach is based on (i) signaling the existence or non‐existence of DMs in a given text, and (ii) applying the ISO‐24617 standard to annotate the DMs’ discourse relation and communicative function in the corpora. Additionally, we introduce prediction models for detecting the presence of DMs within a text.
Název v anglickém jazyce
Multiword Discourse Markers Across Languages: A Linguistic and Computational Perspective
Popis výsledku anglicky
Discourse markers (DMs) are linguistic expressions that convey different semantic and pragmatic values, managing and organizing the structure of spoken and written discourses. They can be either single‐word or multiword expressions (MWE), made up of conjunctions, adverbs, and prepositional phrases. Although DMs are the focus of many studies, some questions regarding the interoperability of taxonomies and automatic identification and classification require further research. We aim to tackle these issues by offering a critical analysis and discussing the constitution of a multilingual corpus in 10 languages, i.e., English, Lithuanian, Bulgarian, German, Macedonian, Romanian, Hebrew, Polish, European Portuguese, and Italian. The novel two‐level annotation approach is based on (i) signaling the existence or non‐existence of DMs in a given text, and (ii) applying the ISO‐24617 standard to annotate the DMs’ discourse relation and communicative function in the corpora. Additionally, we introduce prediction models for detecting the presence of DMs within a text.
Klasifikace
Druh
J<sub>SC</sub> - Článek v periodiku v databázi SCOPUS
CEP obor
—
OECD FORD obor
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)
Návaznosti výsledku
Projekt
—
Návaznosti
—
Ostatní
Rok uplatnění
2025
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Údaje specifické pro druh výsledku
Název periodika
International Journal of Applied Linguistics
ISSN
0802-6106
e-ISSN
1473-4192
Svazek periodika
35
Číslo periodika v rámci svazku
4
Stát vydavatele periodika
US - Spojené státy americké
Počet stran výsledku
13
Strana od-do
2078-2090
Kód UT WoS článku
—
EID výsledku v databázi Scopus
2-s2.0-105005199556