About: Automatic Online Subtitling of the Czech Parliament Meetings     Goto   Sponge   NotDistinct   Permalink

An Entity of Type : http://linked.opendata.cz/ontology/domain/vavai/Vysledek, within Data Space : linked.opendata.cz associated with source document(s)

AttributesValues
rdf:type
Description
  • Rozpoznávací systém je založen na skrytých markovových modelech (HMM), lexikálních stromech a bigramovém jazykovém modelu. Akustický model je natrénován na 40 hodinách parlamentních schůzí a jazykový model na více než 10M slov přepisů parlamentních schůzí. První část článku se zabývá normalizací textu a přípravou třídového jazykového modelu. Druhá část popisuje rozpoznávací síť a její dekódování s ohledem práci v reálném čase se slovníkem až 100k slov. Třetí část nastiňuje strukturu aplikace umožňující generován (cs)
  • This paper describes a LVCSR system for automatic online subtitling (closed captioning) of TV transmissions of the Czech Parliament meetings. The recognition system is based on Hidden Markov Models, lexical trees and bigram language model. The acoustic model is trained on 40 hours of parliament speech and the language model on more than 10M tokens of parliament speech trancriptions. The first part of the article is focused on text normalization and class-based language model preparation. The second part describes the recognition network and its decoding with respect to real-time operation demands using up to 100k vocabulary. The third part outlines the application framework allowing generation and displaying of subtitles for any audio/video source. Finally, experimental results obtained on parliament speeches with recognition accuracy varying from 80 to 95 % (according to the discussed topic) are reported and discussed.
  • This paper describes a LVCSR system for automatic online subtitling (closed captioning) of TV transmissions of the Czech Parliament meetings. The recognition system is based on Hidden Markov Models, lexical trees and bigram language model. The acoustic model is trained on 40 hours of parliament speech and the language model on more than 10M tokens of parliament speech trancriptions. The first part of the article is focused on text normalization and class-based language model preparation. The second part describes the recognition network and its decoding with respect to real-time operation demands using up to 100k vocabulary. The third part outlines the application framework allowing generation and displaying of subtitles for any audio/video source. Finally, experimental results obtained on parliament speeches with recognition accuracy varying from 80 to 95 % (according to the discussed topic) are reported and discussed. (en)
Title
  • Automatic Online Subtitling of the Czech Parliament Meetings
  • Automatic Online Subtitling of the Czech Parliament Meetings (en)
  • Automatické online titulkování parlamentních přenosů (cs)
skos:prefLabel
  • Automatic Online Subtitling of the Czech Parliament Meetings
  • Automatic Online Subtitling of the Czech Parliament Meetings (en)
  • Automatické online titulkování parlamentních přenosů (cs)
skos:notation
  • RIV/49777513:23520/06:00000013!RIV07-AV0-23520___
http://linked.open.../vavai/riv/strany
  • 501
http://linked.open...avai/riv/aktivita
http://linked.open...avai/riv/aktivity
  • P(1QS101470516)
http://linked.open...iv/cisloPeriodika
  • 0
http://linked.open...vai/riv/dodaniDat
http://linked.open...aciTvurceVysledku
http://linked.open.../riv/druhVysledku
http://linked.open...iv/duvernostUdaju
http://linked.open...titaPredkladatele
http://linked.open...dnocenehoVysledku
  • 466411
http://linked.open...ai/riv/idVysledku
  • RIV/49777513:23520/06:00000013
http://linked.open...riv/jazykVysledku
http://linked.open.../riv/klicovaSlova
  • ASR; online; subtitling; parliament; czech (en)
http://linked.open.../riv/klicoveSlovo
http://linked.open...odStatuVydavatele
  • DE - Spolková republika Německo
http://linked.open...ontrolniKodProRIV
  • [095D44D867DE]
http://linked.open...i/riv/nazevZdroje
  • Lecture Notes in Artificial Intelligence
http://linked.open...in/vavai/riv/obor
http://linked.open...ichTvurcuVysledku
http://linked.open...cetTvurcuVysledku
http://linked.open...vavai/riv/projekt
http://linked.open...UplatneniVysledku
http://linked.open...iv/tvurceVysledku
  • Kanis, Jakub
  • Pražák, Aleš
  • Psutka, Josef
  • Müller, Luděk
  • Hoidekr, Jan
issn
  • 0302-9743
number of pages
http://localhost/t...ganizacniJednotka
  • 23520
Faceted Search & Find service v1.16.118 as of Jun 21 2024


Alternative Linked Data Documents: ODE     Content Formats:   [cxml] [csv]     RDF   [text] [turtle] [ld+json] [rdf+json] [rdf+xml]     ODATA   [atom+xml] [odata+json]     Microdata   [microdata+json] [html]    About   
This material is Open Knowledge   W3C Semantic Web Technology [RDF Data] Valid XHTML + RDFa
OpenLink Virtuoso version 07.20.3240 as of Jun 21 2024, on Linux (x86_64-pc-linux-gnu), Single-Server Edition (126 GB total memory, 77 GB memory in use)
Data on this page belongs to its respective rights holders.
Virtuoso Faceted Browser Copyright © 2009-2024 OpenLink Software