The English Teacher Corpus: a novel approach to learner corpus development
Identifikátory výsledku
Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11210%2F25%3A10507123" target="_blank" >RIV/00216208:11210/25:10507123 - isvavai.cz</a>
Výsledek na webu
<a href="https://verso.is.cuni.cz/pub/verso.fpl?fname=obd_publikace_handle&handle=~6E2d5tp.o" target="_blank" >https://verso.is.cuni.cz/pub/verso.fpl?fname=obd_publikace_handle&handle=~6E2d5tp.o</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.3366/cor.2025.0348" target="_blank" >10.3366/cor.2025.0348</a>
Alternativní jazyky
Jazyk výsledku
angličtina
Název v původním jazyce
The English Teacher Corpus: a novel approach to learner corpus development
Popis výsledku v původním jazyce
This study introduces the English Teacher Corpus (etc), delineating its development, technical parameters, and research prospects. This spoken learner corpus contains spontaneous and semi-spontaneous speech tasks performed by Czech teachers of English as a foreign language (efl). The tasks include a monologue, dialogue, picture-based narrative, reading-aloud assignment, and an interview conducted in the teacher's L1. Complementing this corpus is a reference counterpart featuring native English teachers based in the Czech Republic, mirroring the etc's task design. In its 12.5 hours of recorded and transcribed text, the etc consists of 76,122 tokens for the L2 and 31,898 tokens for the L1 sub-corpus. The corpus has been partly transcribed by Whisper AI and subsequently aligned using exmaralda. The etc marks a pioneering effort as the first spoken learner corpus produced by efl teachers. Its innovation extends beyond its content, as it gave rise to a developmental and pedagogical project within a university teacher-training programme.
Název v anglickém jazyce
The English Teacher Corpus: a novel approach to learner corpus development
Popis výsledku anglicky
This study introduces the English Teacher Corpus (etc), delineating its development, technical parameters, and research prospects. This spoken learner corpus contains spontaneous and semi-spontaneous speech tasks performed by Czech teachers of English as a foreign language (efl). The tasks include a monologue, dialogue, picture-based narrative, reading-aloud assignment, and an interview conducted in the teacher's L1. Complementing this corpus is a reference counterpart featuring native English teachers based in the Czech Republic, mirroring the etc's task design. In its 12.5 hours of recorded and transcribed text, the etc consists of 76,122 tokens for the L2 and 31,898 tokens for the L1 sub-corpus. The corpus has been partly transcribed by Whisper AI and subsequently aligned using exmaralda. The etc marks a pioneering effort as the first spoken learner corpus produced by efl teachers. Its innovation extends beyond its content, as it gave rise to a developmental and pedagogical project within a university teacher-training programme.
Klasifikace
Druh
J<sub>imp</sub> - Článek v periodiku v databázi Web of Science
CEP obor
—
OECD FORD obor
60203 - Linguistics
Návaznosti výsledku
Projekt
—
Návaznosti
I - Institucionalni podpora na dlouhodoby koncepcni rozvoj vyzkumne organizace
Ostatní
Rok uplatnění
2025
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Údaje specifické pro druh výsledku
Název periodika
Corpora
ISSN
1749-5032
e-ISSN
1755-1676
Svazek periodika
20
Číslo periodika v rámci svazku
3
Stát vydavatele periodika
GB - Spojené království Velké Británie a Severního Irska
Počet stran výsledku
14
Strana od-do
439-452
Kód UT WoS článku
001641558500001
EID výsledku v databázi Scopus
—