Pretražite po imenu i prezimenu autora, mentora, urednika, prevoditelja

Napredna pretraga

Pregled bibliografske jedinice broj: 1069509

Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian


Ignatoski, Matea; Lerga, Jonatan; Stanković, Ljubiša; Daković, Miloš
Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian // Mathematics, 2020 (2020), 8; 1059, 14 doi:10.3390/math8071059 (međunarodna recenzija, članak, znanstveni)


CROSBI ID: 1069509 Za ispravke kontaktirajte CROSBI podršku putem web obrasca

Naslov
Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian

Autori
Ignatoski, Matea ; Lerga, Jonatan ; Stanković, Ljubiša ; Daković, Miloš

Izvornik
Mathematics (2227-7390) 2020 (2020), 8; 1059, 14

Vrsta, podvrsta i kategorija rada
Radovi u časopisima, članak, znanstveni

Ključne riječi
Arithmetic ; Lempel–Ziv–Welch (LZW) ; Text compression ; Encoding ; English ; German ; French ; Italian ; Czech ; Hungarian ; Finnish ; Croatian

Sažetak
The rapid growth in the amount of data in the digital world leads to the need for data compression, and so forth, reducing the number of bits needed to represent a text file, an image, audio, or video content. Compressing data saves storage capacity and speeds up data transmission. In this paper, we focus on the text compression and provide a comparison of algorithms (in particular, entropy-based arithmetic and dictionary-based Lempel–Ziv–Welch (LZW) methods) for text compression in different languages (Croatian, Finnish, Hungarian, Czech, Italian, French, German, and English). The main goal is to answer a question: ”How does the language of a text affect the compression ratio?” The results indicated that the compression ratio is affected by the size of the language alphabet, and size or type of the text. For example, The European Green Deal was compressed by 75.79%, 76.17%, 77.33%, 76.84%, 73.25%, 74.63%, 75.14%, and 74.51% using the LZW algorithm, and by 72.54%, 71.47%, 72.87%, 73.43%, 69.62%, 69.94%, 72.42% and 72% using the arithmetic algorithm for the English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian versions, respectively.

Izvorni jezik
Engleski

Znanstvena područja
Računarstvo



POVEZANOST RADA


Projekti:
IP-2020-02-4358
HRZZ-IP-2018-01-3739 - Sustav potpore odlučivanju za zeleniju i sigurniju plovidbu brodova (DESSERT) (Prpić-Oršić, Jasna, HRZZ - 2018-01) ( CroRIS)

Ustanove:
Tehnički fakultet, Rijeka,
Sveučilište u Rijeci

Profili:

Avatar Url Jonatan Lerga (autor)

Poveznice na cjeloviti tekst rada:

doi www.mdpi.com

Citiraj ovu publikaciju:

Ignatoski, Matea; Lerga, Jonatan; Stanković, Ljubiša; Daković, Miloš
Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian // Mathematics, 2020 (2020), 8; 1059, 14 doi:10.3390/math8071059 (međunarodna recenzija, članak, znanstveni)
Ignatoski, M., Lerga, J., Stanković, L. & Daković, M. (2020) Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian. Mathematics, 2020 (8), 1059, 14 doi:10.3390/math8071059.
@article{article, author = {Ignatoski, Matea and Lerga, Jonatan and Stankovi\'{c}, Ljubi\v{s}a and Dakovi\'{c}, Milo\v{s}}, year = {2020}, pages = {14}, DOI = {10.3390/math8071059}, chapter = {1059}, keywords = {Arithmetic, Lempel–Ziv–Welch (LZW), Text compression, Encoding, English, German, French, Italian, Czech, Hungarian, Finnish, Croatian}, journal = {Mathematics}, doi = {10.3390/math8071059}, volume = {2020}, number = {8}, issn = {2227-7390}, title = {Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian}, keyword = {Arithmetic, Lempel–Ziv–Welch (LZW), Text compression, Encoding, English, German, French, Italian, Czech, Hungarian, Finnish, Croatian}, chapternumber = {1059} }
@article{article, author = {Ignatoski, Matea and Lerga, Jonatan and Stankovi\'{c}, Ljubi\v{s}a and Dakovi\'{c}, Milo\v{s}}, year = {2020}, pages = {14}, DOI = {10.3390/math8071059}, chapter = {1059}, keywords = {Arithmetic, Lempel–Ziv–Welch (LZW), Text compression, Encoding, English, German, French, Italian, Czech, Hungarian, Finnish, Croatian}, journal = {Mathematics}, doi = {10.3390/math8071059}, volume = {2020}, number = {8}, issn = {2227-7390}, title = {Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian}, keyword = {Arithmetic, Lempel–Ziv–Welch (LZW), Text compression, Encoding, English, German, French, Italian, Czech, Hungarian, Finnish, Croatian}, chapternumber = {1059} }

Časopis indeksira:


  • Current Contents Connect (CCC)
  • Web of Science Core Collection (WoSCC)
    • Science Citation Index Expanded (SCI-EXP)
    • SCI-EXP, SSCI i/ili A&HCI
  • Scopus


Citati:





    Contrast
    Increase Font
    Decrease Font
    Dyslexic Font