Automatic online subtitling of the Czech parliament meetings

The result's identifiers

Result code in IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F49777513%3A23520%2F06%3A00000013" target="_blank" >RIV/49777513:23520/06:00000013 - isvavai.cz</a>
Result on the web
—
DOI - Digital Object Identifier
—

Alternative languages

Result language
angličtina
Original language name
Automatic online subtitling of the Czech parliament meetings
Original language description
This paper describes a LVCSR system for automatic online subtitling (closed captioning) of TV transmissions of the Czech Parliament meetings. The recognition system is based on Hidden Markov Models, lexical trees and bigram language model. The acoustic model is trained on 40 hours of parliament speech and the language model on more than 10M tokens of parliament speech trancriptions. The first part of the article is focused on text normalization and class-based language model preparation. The second partdescribes the recognition network and its decoding with respect to real-time operation demands using up to 100k vocabulary. The third part outlines the application framework allowing generation and displaying of subtitles for any audio/video source. Finally, experimental results obtained on parliament speeches with recognition accuracy varying from 80 to 95 % (according to the discussed topic) are reported and discussed.
Czech name
Automatické online titulkování parlamentních přenosů
Czech description
Článek popisuje LVCSR systém pro automatické online titulkování TV přenosů zasedání českého parlamentu. Rozpoznávací systém je založen na skrytých markovových modelech (HMM), lexikálních stromech a bigramovém jazykovém modelu. Akustický model je natrénován na 40 hodinách parlamentních schůzí a jazykový model na více než 10M slov přepisů parlamentních schůzí. První část článku se zabývá normalizací textu a přípravou třídového jazykového modelu. Druhá část popisuje rozpoznávací síť a její dekódování s ohledem práci v reálném čase se slovníkem až 100k slov. Třetí část nastiňuje strukturu aplikace umožňující generování a zobrazování titulků pro libovolný audio/video zdroj. Závěrem jsou prezentovány a diskutovány experimentální výsledky parlamentních schůzís přesností rozpoznávání od 80 do 95 % (podle diskutovaného tématu).

Classification

Type
D - Article in proceedings
CEP classification
JD - Use of computers, robotics and its application
OECD FORD branch
—

Result continuities

Project
<a href="/en/project/1QS101470516" target="_blank" >1QS101470516: Automatic keyword spotting in audio data streams</a><br>
Continuities
P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)

Others

Publication year
2006
Confidentiality
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů

Data specific for result type

Article name in the collection
Lecture Notes in Artificial Intelligence
ISBN
3-540-39090-1
ISSN
—
e-ISSN
—
Number of pages
8
Pages from-to
—
Publisher name
Springer
Place of publication
Berlin
Event location
—
Event date
—
Type of event by nationality
—
UT code for WoS article
000241103500063

Similar results(10)

Adaptive Language Model in Automatic Online Subtitling Adaptive Language Model in Automatic Online Subtitling LIVE TV SUBTITLING - Fast 2-pass LVCSR System for Online Subtitling

What are you looking for?

Quick search

Smart search

Automatic online subtitling of the Czech parliament meetings

The result's identifiers

Alternative languages

Classification

Result continuities

Others

Data specific for result type

Similar results(10)

What are you looking for?

Quick search

Smart search

Result description

The result's identifiers

The result's identifiers

Alternative languages

Alternative languages

Classification

Classification

Result continuities

Result continuities

Others

Others

Data specific for result type

Data specific for result type

Similar results(10)