Vše

Co hledáte?

Vše
Projekty
Výsledky výzkumu
Subjekty

Rychlé hledání

  • Projekty podpořené TA ČR
  • Významné projekty
  • Projekty s nejvyšší státní podporou
  • Aktuálně běžící projekty

Chytré vyhledávání

  • Takto najdu konkrétní +slovo
  • Takto z výsledků -slovo zcela vynechám
  • “Takto můžu najít celou frázi”

Creating a sociologically balanced spoken corpus

Identifikátory výsledku

  • Kód výsledku v IS VaVaI

    <a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11210%2F19%3A10402534" target="_blank" >RIV/00216208:11210/19:10402534 - isvavai.cz</a>

  • Výsledek na webu

  • DOI - Digital Object Identifier

Alternativní jazyky

  • Jazyk výsledku

    angličtina

  • Název v původním jazyce

    Creating a sociologically balanced spoken corpus

  • Popis výsledku v původním jazyce

    The article presents the corpora of spoken Czech, which were created for language research and are publicly accessible. These are corpora that capture private spontaneous dialogues, therefore they were compiled according to the sociological criteria of each speaker. These corpora have been binary balanced from the beginning in the categories of gender, age and the highest achieved level of education. Later, dialect regions were added, in which the speaker spent his childhood. It is quite difficult to combine these criteria when recording longer interviews. Full balancing of all categories is accomplished in ORTOFON corpus.

  • Název v anglickém jazyce

    Creating a sociologically balanced spoken corpus

  • Popis výsledku anglicky

    The article presents the corpora of spoken Czech, which were created for language research and are publicly accessible. These are corpora that capture private spontaneous dialogues, therefore they were compiled according to the sociological criteria of each speaker. These corpora have been binary balanced from the beginning in the categories of gender, age and the highest achieved level of education. Later, dialect regions were added, in which the speaker spent his childhood. It is quite difficult to combine these criteria when recording longer interviews. Full balancing of all categories is accomplished in ORTOFON corpus.

Klasifikace

  • Druh

    D - Stať ve sborníku

  • CEP obor

  • OECD FORD obor

    60203 - Linguistics

Návaznosti výsledku

  • Projekt

    <a href="/cs/project/LM2015044" target="_blank" >LM2015044: Český národní korpus</a><br>

  • Návaznosti

    P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)

Ostatní

  • Rok uplatnění

    2019

  • Kód důvěrnosti údajů

    S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů

Údaje specifické pro druh výsledku

  • Název statě ve sborníku

    Proceedings of the international conference «Corpus Linguistics – 2019»

  • ISBN

  • ISSN

    2412-9623

  • e-ISSN

  • Počet stran výsledku

    8

  • Strana od-do

    40-47

  • Název nakladatele

    Saint Petersburg University Press

  • Místo vydání

    Petrohrad

  • Místo konání akce

    Petrohrad

  • Datum konání akce

    24. 6. 2019

  • Typ akce podle státní příslušnosti

    WRD - Celosvětová akce

  • Kód UT WoS článku