• Skip to primary navigation
  • Skip to main content
  • Skip to primary sidebar
  • Skip to footer

Center for Artificial Intelligence and Cybersecurity – AIRI

  • Home
  • About Us
    • Center Activities
    • Vision, Mission and Goals
    • Center Faculty
    • Steering Committee
    • Press
  • Research
    • Scientific Projects
    • Research Papers
  • Laboratories
    • Machine Learning
    • Natural Speech & Language Processing
    • Blockchain Technology
    • Information Processing & Pattern Recognition
    • AI in Medicine
    • Data Mining
    • Computer Vision
    • Complex Networks
    • Human-Computer Interaction
    • Maritime Cybersecurity
    • Autonomous Navigation
    • AI in Mechatronics
    • AI in Education
    • Hybrid Computational Methods
    • Drug Design
    • Legal Aspects of AI
    • Ethically Aligned AI
    • Cultural Complexity
    • Trustworthy and Explainable AI
  • Collaboration
    • Industry Collaboration
    • Industry Projects
    • International Collaboration
  • News
  • Contact
  • Login

Comparison of Entropy and Dictionary Based Text Compression in English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian

01.07.2020

The rapid growth in the amount of data in the digital world leads to the need for data compression, and so forth, reducing the number of bits needed to represent a text file, an image, audio, or video content. Compressing data saves storage capacity and speeds up data transmission. In this paper, we focus on the text compression and provide a comparison of algorithms (in particular, entropy-based arithmetic and dictionary-based Lempel–Ziv–Welch (LZW) methods) for text compression in different languages (Croatian, Finnish, Hungarian, Czech, Italian, French, German, and English). The main goal is to answer a question: ”How does the language of a text affect the compression ratio?” The results indicated that the compression ratio is affected by the size of the language alphabet, and size or type of the text. For example, The European Green Deal was compressed by 75.79%, 76.17%, 77.33%, 76.84%, 73.25%, 74.63%, 75.14%, and 74.51% using the LZW algorithm, and by 72.54%, 71.47%, 72.87%, 73.43%, 69.62%, 69.94%, 72.42% and 72% using the arithmetic algorithm for the English, German, French, Italian, Czech, Hungarian, Finnish, and Croatian versions, respectively.

Authors:
Matea Ignatoski, Jonatan Lerga, Ljubiša Stanković, Miloš Daković
Journal:
Mathematics
Publishing date:
01.07.2020
View original article

Primary Sidebar

Latest Projects

Advanced Data Analysis Using Digital Signal Processing and Machine Learning Techniques

Compound Flooding in Coastal Rivers in Present and Future Climate

Data Processing on Graphs

North Adriatic Hydrogen Valley

Data Governance and Intellectual Property Governance in Common European Data Spaces – DGIP-CEDS

Latest Research Papers

Forecasting the Trajectory of Personal Watercrafts Using Models Based on Recurrent Neural Networks

A System for Real-Time Detection of Abandoned Luggage

Enhancing Biophysical Muscle Fatigue Model in the Dynamic Context of Soccer

Pravna tehnologija (Legal Tech) i njezina (ne)prikladnost za zamjenu pravne struke

Regression-Based Machine Learning Approaches for Estimating Discharge from Water Levels in Microtidal Rivers

Latest News

Arian Skoki defended his doctoral thesis “Data-Driven Assessment of Player Performance and Recovery in Soccer”

Anna Maria Mihel defended her PhD dissertation topic

Prof. dr. sc. Renato Filjar participated at the meeting of the 31st National Space-Based Positioning, Navigation and Timing US Advisory Board

Presentation of the NPOO project Peoplet

Ana Vranković Lacković defended her doctoral thesis

We provide the expertise for solving real world problems using AI

If your company wants to implement artificial intelligence in your products or services, or increase your level of cybersecurity, our multidisciplinary team of scientists is your ideal partner.

Contact us

Footer

Center for Artificial Intelligence and Cybersecurity
  • jlerga@airi.uniri.hr
  • +385 51 406 500

University of Rijeka

University of Rijeka

About the Center

  • About Us
  • News
  • Privacy Policy
  • Contact

Center Activities

  • Laboratories
  • Scientific Projects
  • Industry Projects
  • Research Papers
  • Industry Collaboration
  • International Collaboration

Footer bottom left

© 2020 Center for Artificial Intelligence and Cybersecurity, all rights reserved.

Designed & developed by Nela Dunato Art & Design