chared: Character Encoding Detection with a Known Language
The result's identifiers
Result code in IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216224%3A14330%2F11%3A00050165" target="_blank" >RIV/00216224:14330/11:00050165 - isvavai.cz</a>
Result on the web
—
DOI - Digital Object Identifier
—
Alternative languages
Result language
angličtina
Original language name
chared: Character Encoding Detection with a Known Language
Original language description
chared is a system which can detect character encoding of a text document provided the language of the document is known. The system supports a wide range of languages and the most commonly used character encodings. We explain the details of the algorithm, describe the process of creating models for various languages and present results of an evaluation on a collection of Web pages.
Czech name
—
Czech description
—
Classification
Type
D - Article in proceedings
CEP classification
IN - Informatics
OECD FORD branch
—
Result continuities
Project
<a href="/en/project/GAP401%2F10%2F0792" target="_blank" >GAP401/10/0792: Temporal aspects of knowledge and information</a><br>
Continuities
P - Projekt vyzkumu a vyvoje financovany z verejnych zdroju (s odkazem do CEP)<br>S - Specificky vyzkum na vysokych skolach
Others
Publication year
2011
Confidentiality
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Data specific for result type
Article name in the collection
RASLAN 2011
ISBN
978-80-263-0077-9
ISSN
—
e-ISSN
—
Number of pages
5
Pages from-to
125-129
Publisher name
Tribun EU
Place of publication
Brno, Czech Republic
Event location
Karlova Studánka, Czech Republic
Event date
Jan 1, 2011
Type of event by nationality
CST - Celostátní akce
UT code for WoS article
—