UDPipe 2.0 Prototype at CoNLL 2018 UD Shared Task
The result's identifiers
Result code in IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F18%3A10390208" target="_blank" >RIV/00216208:11320/18:10390208 - isvavai.cz</a>
Result on the web
<a href="http://universaldependencies.org/conll18/proceedings/pdf/K18-2020.pdf" target="_blank" >http://universaldependencies.org/conll18/proceedings/pdf/K18-2020.pdf</a>
DOI - Digital Object Identifier
—
Alternative languages
Result language
angličtina
Original language name
UDPipe 2.0 Prototype at CoNLL 2018 UD Shared Task
Original language description
UDPipe is a trainable pipeline which performs sentence segmentation, tokenization, POS tagging, lemmatization and dependency parsing. We present a prototype for UDPipe 2.0 and evaluate it in the CoNLL 2018 UD Shared Task: Multilingual Parsing from Raw Text to Universal Dependencies, which employs three metrics for submission ranking. Out of 26 participants, the prototype placed first in the MLAS ranking, third in the LAS ranking and third in the BLEX ranking. In extrinsic parser evaluation EPE 2018, the system ranked first in the overall score. The prototype utilizes an artificial neural network with a single joint model for POS tagging, lemmatization and dependency parsing, and is trained only using the CoNLL-U training data and pretrained word embeddings, contrary to both systems surpassing the prototype in the LAS and BLEX ranking in the shared task. The open-source code of the prototype is available at http://github.com/CoNLL-UD-2018/UDPipe-Future. After the shared task, we slightly refined the mo
Czech name
—
Czech description
—
Classification
Type
D - Article in proceedings
CEP classification
—
OECD FORD branch
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)
Result continuities
Project
<a href="/en/project/LM2015071" target="_blank" >LM2015071: Language Research Infrastructure in the Czech Republic</a><br>
Continuities
P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)<br>S - Specificky vyzkum na vysokych skolach
Others
Publication year
2018
Confidentiality
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Data specific for result type
Article name in the collection
Proceedings of CoNLL 2018: The SIGNLL Conference on Computational Natural Language Learning
ISBN
978-1-948087-72-8
ISSN
—
e-ISSN
neuvedeno
Number of pages
11
Pages from-to
197-207
Publisher name
Association for Computational Linguistics
Place of publication
Stroudsburg, PA, USA
Event location
Bruxelles, Belgium
Event date
Oct 31, 2018
Type of event by nationality
WRD - Celosvětová akce
UT code for WoS article
—