Pregled bibliografske jedinice broj: 1069509
Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian
Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian // Mathematics, 2020 (2020), 8; 1059, 14 doi:10.3390/math8071059 (međunarodna recenzija, članak, znanstveni)
CROSBI ID: 1069509 Za ispravke kontaktirajte CROSBI podršku putem web obrasca
Naslov
Comparison of Entropy and Dictionary Based Text
Compression in English, German, French, Italian,
Czech, Hungarian, Finnish, and Croatian
Autori
Ignatoski, Matea ; Lerga, Jonatan ; Stanković, Ljubiša ; Daković, Miloš
Izvornik
Mathematics (2227-7390) 2020
(2020), 8;
1059, 14
Vrsta, podvrsta i kategorija rada
Radovi u časopisima, članak, znanstveni
Ključne riječi
Arithmetic ; Lempel–Ziv–Welch (LZW) ; Text compression ; Encoding ; English ; German ; French ; Italian ; Czech ; Hungarian ; Finnish ; Croatian
Sažetak
The rapid growth in the amount of data in the digital world leads to the need for data compression, and so forth, reducing the number of bits needed to represent a text file, an image, audio, or video content. Compressing data saves storage capacity and speeds up data transmission. In this paper, we focus on the text compression and provide a comparison of algorithms (in particular, entropy-based arithmetic and dictionary-based Lempel–Ziv–Welch (LZW) methods) for text compression in different languages (Croatian, Finnish, Hungarian, Czech, Italian, French, German, and English). The main goal is to answer a question: ”How does the language of a text affect the compression ratio?” The results indicated that the compression ratio is affected by the size of the language alphabet, and size or type of the text. For example, The European Green Deal was compressed by 75.79%, 76.17%, 77.33%, 76.84%, 73.25%, 74.63%, 75.14%, and 74.51% using the LZW algorithm, and by 72.54%, 71.47%, 72.87%, 73.43%, 69.62%, 69.94%, 72.42% and 72% using the arithmetic algorithm for the English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian versions, respectively.
Izvorni jezik
Engleski
Znanstvena područja
Računarstvo
POVEZANOST RADA
Projekti:
IP-2020-02-4358
HRZZ-IP-2018-01-3739 - Sustav potpore odlučivanju za zeleniju i sigurniju plovidbu brodova (DESSERT) (Prpić-Oršić, Jasna, HRZZ - 2018-01) ( CroRIS)
Ustanove:
Tehnički fakultet, Rijeka,
Sveučilište u Rijeci
Profili:
Jonatan Lerga
(autor)
Citiraj ovu publikaciju:
Časopis indeksira:
- Current Contents Connect (CCC)
- Web of Science Core Collection (WoSCC)
- Science Citation Index Expanded (SCI-EXP)
- SCI-EXP, SSCI i/ili A&HCI
- Scopus