A New Approach to Automatically Find and Fix Erroneous Labels in Dependency Parsing Treebanks

The result's identifiers

Result code in IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F21%3A10441550" target="_blank" >RIV/00216208:11320/21:10441550 - isvavai.cz</a>
Result on the web
<a href="https://verso.is.cuni.cz/pub/verso.fpl?fname=obd_publikace_handle&handle=fkq3xX7PWH" target="_blank" >https://verso.is.cuni.cz/pub/verso.fpl?fname=obd_publikace_handle&handle=fkq3xX7PWH</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.34028/iajit/18/3/12" target="_blank" >10.34028/iajit/18/3/12</a>

Alternative languages

Result language
angličtina
Original language name
A New Approach to Automatically Find and Fix Erroneous Labels in Dependency Parsing Treebanks
Original language description
Dependency Parsing (DP) is the existence of sub-term/upper-term relations between the words that make up that sentence for each sentence in the text. DP serves to produce meaningful information for high-level applications. Correct labeling of the text corpus used in DP studies is very important. There will be mistakes in the results of the studies that will be performed with the wrongly-labeled text corpus. If text corpus is labeled manually or automatically by human beings, then faulty cases will occur. As a result of the cases that may arise from human factors or annotations used for labeling, faulty labels will be on freebanks. In order to prevent these errors, detection, and correction of possible faulty labeling is very important in terms of increasing the accuracy of the studies to be carried out. Manual correction of possible faulty labels requires great effort and time. The purpose of this study is to create a model that automatically finds possible faulty labels and offers new label suggestions for faulty labels. With the help of the proposed model, it is aimed to detect and correct possible faulty labels that are included in a text corpus, and to increase consistency among the text corpus of the same language. With the help of the developed model, suggesting new labels for faulty labels by a language expert will be a great convenient for the specialist. Another advantage of the model is that the developed model provides a language-independent structure. It has succeeded in obtaining successful results in finding and correcting potentially faulty labels in experimental studies for Turkish. An increase in accuracy has been detected in studies carried out for languages other than Turkish. In investigating the accuracy of the results obtained by the system, the results were analyzed with the help of 10 different language experts.
Czech name
—
Czech description
—

Classification

Type
J<sub>imp</sub> - Article in a specialist periodical, which is included in the Web of Science database
CEP classification
—
OECD FORD branch
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)

Result continuities

Project
—
Continuities
—

Others

Publication year
2021
Confidentiality
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů

Data specific for result type

Name of the periodical
International Arab Journal of Information Technology
ISSN
1683-3198
e-ISSN
—
Volume of the periodical
18
Issue of the periodical within the volume
3
Country of publishing house
JO - JORDAN
Number of pages
9
Pages from-to
356-364
UT code for WoS article
000667208600012
EID of the result in the Scopus database
2-s2.0-85106439495

Similar results(10)

Graph-based Dependency Parser Building for Myanmar Language Skript 2015: Acquisition corpus of native speakers' Czech - transcripts of essays by students of primary and secondary schools A Weakly Supervised Data Labeling Framework for Machine Lexical Normalization in Vietnamese Social Media

What are you looking for?

Quick search

Smart search

A New Approach to Automatically Find and Fix Erroneous Labels in Dependency Parsing Treebanks

The result's identifiers

Alternative languages

Classification

Result continuities

Others

Data specific for result type

Similar results(10)

What are you looking for?

Quick search

Smart search

Result description

The result's identifiers

The result's identifiers

Alternative languages

Alternative languages

Classification

Classification

Result continuities

Result continuities

Others

Others

Data specific for result type

Data specific for result type

Similar results(10)