Nalazite se na CroRIS probnoj okolini. Ovdje evidentirani podaci neće biti pohranjeni u Informacijskom sustavu znanosti RH. Ako je ovo greška, CroRIS produkcijskoj okolini moguće je pristupi putem poveznice www.croris.hr
izvor podataka: crosbi !

Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian (CROSBI ID 280410)

Prilog u časopisu | izvorni znanstveni rad | međunarodna recenzija

Ignatoski, Matea ; Lerga, Jonatan ; Stanković, Ljubiša ; Daković, Miloš Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian // Mathematics, 2020 (2020), 8; 1059, 14. doi: 10.3390/math8071059

Podaci o odgovornosti

Ignatoski, Matea ; Lerga, Jonatan ; Stanković, Ljubiša ; Daković, Miloš

engleski

Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian

The rapid growth in the amount of data in the digital world leads to the need for data compression, and so forth, reducing the number of bits needed to represent a text file, an image, audio, or video content. Compressing data saves storage capacity and speeds up data transmission. In this paper, we focus on the text compression and provide a comparison of algorithms (in particular, entropy-based arithmetic and dictionary-based Lempel–Ziv–Welch (LZW) methods) for text compression in different languages (Croatian, Finnish, Hungarian, Czech, Italian, French, German, and English). The main goal is to answer a question: ”How does the language of a text affect the compression ratio?” The results indicated that the compression ratio is affected by the size of the language alphabet, and size or type of the text. For example, The European Green Deal was compressed by 75.79%, 76.17%, 77.33%, 76.84%, 73.25%, 74.63%, 75.14%, and 74.51% using the LZW algorithm, and by 72.54%, 71.47%, 72.87%, 73.43%, 69.62%, 69.94%, 72.42% and 72% using the arithmetic algorithm for the English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian versions, respectively.

Arithmetic ; Lempel–Ziv–Welch (LZW) ; Text compression ; Encoding ; English ; German ; French ; Italian ; Czech ; Hungarian ; Finnish ; Croatian

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

Podaci o izdanju

2020 (8)

2020.

1059

14

objavljeno

2227-7390

10.3390/math8071059

Trošak objave rada u otvorenom pristupu

Povezanost rada

Računarstvo

Poveznice
Indeksiranost