Creating a sociologically balanced spoken corpus
Identifikátory výsledku
Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11210%2F19%3A10402534" target="_blank" >RIV/00216208:11210/19:10402534 - isvavai.cz</a>
Výsledek na webu
—
DOI - Digital Object Identifier
—
Alternativní jazyky
Jazyk výsledku
angličtina
Název v původním jazyce
Creating a sociologically balanced spoken corpus
Popis výsledku v původním jazyce
The article presents the corpora of spoken Czech, which were created for language research and are publicly accessible. These are corpora that capture private spontaneous dialogues, therefore they were compiled according to the sociological criteria of each speaker. These corpora have been binary balanced from the beginning in the categories of gender, age and the highest achieved level of education. Later, dialect regions were added, in which the speaker spent his childhood. It is quite difficult to combine these criteria when recording longer interviews. Full balancing of all categories is accomplished in ORTOFON corpus.
Název v anglickém jazyce
Creating a sociologically balanced spoken corpus
Popis výsledku anglicky
The article presents the corpora of spoken Czech, which were created for language research and are publicly accessible. These are corpora that capture private spontaneous dialogues, therefore they were compiled according to the sociological criteria of each speaker. These corpora have been binary balanced from the beginning in the categories of gender, age and the highest achieved level of education. Later, dialect regions were added, in which the speaker spent his childhood. It is quite difficult to combine these criteria when recording longer interviews. Full balancing of all categories is accomplished in ORTOFON corpus.
Klasifikace
Druh
D - Stať ve sborníku
CEP obor
—
OECD FORD obor
60203 - Linguistics
Návaznosti výsledku
Projekt
<a href="/cs/project/LM2015044" target="_blank" >LM2015044: Český národní korpus</a><br>
Návaznosti
P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)
Ostatní
Rok uplatnění
2019
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Údaje specifické pro druh výsledku
Název statě ve sborníku
Proceedings of the international conference «Corpus Linguistics – 2019»
ISBN
—
ISSN
2412-9623
e-ISSN
—
Počet stran výsledku
8
Strana od-do
40-47
Název nakladatele
Saint Petersburg University Press
Místo vydání
Petrohrad
Místo konání akce
Petrohrad
Datum konání akce
24. 6. 2019
Typ akce podle státní příslušnosti
WRD - Celosvětová akce
Kód UT WoS článku
—