All

What are you looking for?

All
Projects
Results
Organizations

Quick search

  • Projects supported by TA ČR
  • Excellent projects
  • Projects with the highest public support
  • Current projects

Smart search

  • That is how I find a specific +word
  • That is how I leave the -word out of the results
  • “That is how I can find the whole phrase”

New tools for working with the ORAL series corpora of spoken Czech : AchSynku and MluvKonk

The result's identifiers

  • Result code in IS VaVaI

    <a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11210%2F15%3A10319671" target="_blank" >RIV/00216208:11210/15:10319671 - isvavai.cz</a>

  • Result on the web

    <a href="http://korpus.sk/~slovko/2015/Proceedings_Slovko_2015.pdf" target="_blank" >http://korpus.sk/~slovko/2015/Proceedings_Slovko_2015.pdf</a>

  • DOI - Digital Object Identifier

Alternative languages

  • Result language

    angličtina

  • Original language name

    New tools for working with the ORAL series corpora of spoken Czech : AchSynku and MluvKonk

  • Original language description

    This paper introduces two simple web-based tools whose aim is to make it easier to work with the ORAL series spontaneous spoken language corpora of the Czech National Corpus. Both strive to overcome and circumvent some of the limitations, either in the data themselves or in their visualization, currently faced by linguists who use them for research. AchSynku is a variant search tool which aims to compensate for the lack of lemmatization in spoken corpora by suggesting, based on a word form input by theuser, a list of variant and related forms occurring in the target corpora. MluvKonk is a visualization environment which turns single-line concordances into a multi-tier layout with one speaker per tier. This makes it easier to follow the structure of amulti-party conversation, including turn-switching and overlaps. Though ultimately destined to be superseded by more systemic solutions, both applications are under active development and feedback is welcome, because these ulterior soluti

  • Czech name

  • Czech description

Classification

  • Type

    D - Article in proceedings

  • CEP classification

    AI - Linguistics

  • OECD FORD branch

Result continuities

  • Project

    <a href="/en/project/LM2011023" target="_blank" >LM2011023: Czech National Corpus</a><br>

  • Continuities

    P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)

Others

  • Publication year

    2015

  • Confidentiality

    S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů

Data specific for result type

  • Article name in the collection

    Natural Language Processing, Corpus Linguistics, Lexicography

  • ISBN

    978-3-942303-32-3

  • ISSN

  • e-ISSN

  • Number of pages

    12

  • Pages from-to

    90-101

  • Publisher name

    RAM-Verlag

  • Place of publication

    Lüdenscheid

  • Event location

    Bratislava

  • Event date

    Oct 21, 2015

  • Type of event by nationality

    EUR - Evropská akce

  • UT code for WoS article