An intuitive IPT-IPA based deep learning approach for video compression techniques.

Saved in:
Bibliographic Details
Title: An intuitive IPT-IPA based deep learning approach for video compression techniques.
Authors: Naik, Mudhavath Ramesh1 (AUTHOR), Kumar, Jayendra1 (AUTHOR) jkumar.ece@nitjsr.ac.in, Yadav, Arvind R.2 (AUTHOR)
Source: Multimedia Tools & Applications. Dec2025, Vol. 84 Issue 42, p50437-50470. 34p.
Subjects: Video compression, Deep learning, Statistical accuracy, Prediction models, Lossy data compression, Generative adversarial networks
Abstract: The acquisition and processing of digital images and videos has become increasingly common in modern times owing to the expansion of internet services and their associated portable devices. With the limited processing capacity and memory constraints of such devices, the need for innovative compression algorithms emerges. In contemporary research, Researchers have widely adopted Deep Learning (DL) models to enhance image and video compression techniques. Despite the wide usage, the performance and memory constraints of analytical structures are predominant with the DL models that are affected by loss estimation, pruning, and memory localization. So, to analyse the estimation and minimization of memory features in DL models, an Iterative Predictive Transform (IPA) and an Intuitive Prediction Algorithm (IPT) technique have been proposed, facilitating the reduction of compression loss with improved loss punning for DL methods. The proposed transformation and prediction algorithm optimizes weight parameters by using a log sigmoid function to lower the error rate and provide superior accuracy for visual compression. The efficiency of the proposed compression techniques has been evaluated using several performance metrics, including Structural Similarity Index Measure (SSIM), Mean Square Error (MSE), Peak Signal-to-Noise Ratio (PSNR), Compression Ratio (CR), and classification accuracy. Experimental results indicate that the IPA-IPT-based methods specifically IPA-AE, IPA-DL, and IPT-Convolutional Neural Network (CNN) models—achieved an exceptional accuracy of 99.998%, with compression and multiplication factors of 17 and 10, respectively. These outcomes demonstrate that the proposed approaches significantly outperform existing state-of-the-art algorithms based on Generative Adversarial Networks (GANs). [ABSTRACT FROM AUTHOR]
Copyright of Multimedia Tools & Applications is the property of Springer Nature and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
Database: Engineering Source
FullText Text:
  Availability: 0
Header DbId: egs
DbLabel: Engineering Source
An: 190506892
AccessLevel: 6
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: An intuitive IPT-IPA based deep learning approach for video compression techniques.
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Naik%2C+Mudhavath+Ramesh%22">Naik, Mudhavath Ramesh</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Kumar%2C+Jayendra%22">Kumar, Jayendra</searchLink><relatesTo>1</relatesTo> (AUTHOR)<i> jkumar.ece@nitjsr.ac.in</i><br /><searchLink fieldCode="AR" term="%22Yadav%2C+Arvind+R%2E%22">Yadav, Arvind R.</searchLink><relatesTo>2</relatesTo> (AUTHOR)
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="JN" term="%22Multimedia+Tools+%26+Applications%22">Multimedia Tools & Applications</searchLink>. Dec2025, Vol. 84 Issue 42, p50437-50470. 34p.
– Name: Subject
  Label: Subjects
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Video+compression%22">Video compression</searchLink><br /><searchLink fieldCode="DE" term="%22Deep+learning%22">Deep learning</searchLink><br /><searchLink fieldCode="DE" term="%22Statistical+accuracy%22">Statistical accuracy</searchLink><br /><searchLink fieldCode="DE" term="%22Prediction+models%22">Prediction models</searchLink><br /><searchLink fieldCode="DE" term="%22Lossy+data+compression%22">Lossy data compression</searchLink><br /><searchLink fieldCode="DE" term="%22Generative+adversarial+networks%22">Generative adversarial networks</searchLink>
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: The acquisition and processing of digital images and videos has become increasingly common in modern times owing to the expansion of internet services and their associated portable devices. With the limited processing capacity and memory constraints of such devices, the need for innovative compression algorithms emerges. In contemporary research, Researchers have widely adopted Deep Learning (DL) models to enhance image and video compression techniques. Despite the wide usage, the performance and memory constraints of analytical structures are predominant with the DL models that are affected by loss estimation, pruning, and memory localization. So, to analyse the estimation and minimization of memory features in DL models, an Iterative Predictive Transform (IPA) and an Intuitive Prediction Algorithm (IPT) technique have been proposed, facilitating the reduction of compression loss with improved loss punning for DL methods. The proposed transformation and prediction algorithm optimizes weight parameters by using a log sigmoid function to lower the error rate and provide superior accuracy for visual compression. The efficiency of the proposed compression techniques has been evaluated using several performance metrics, including Structural Similarity Index Measure (SSIM), Mean Square Error (MSE), Peak Signal-to-Noise Ratio (PSNR), Compression Ratio (CR), and classification accuracy. Experimental results indicate that the IPA-IPT-based methods specifically IPA-AE, IPA-DL, and IPT-Convolutional Neural Network (CNN) models—achieved an exceptional accuracy of 99.998%, with compression and multiplication factors of 17 and 10, respectively. These outcomes demonstrate that the proposed approaches significantly outperform existing state-of-the-art algorithms based on Generative Adversarial Networks (GANs). [ABSTRACT FROM AUTHOR]
– Name: AbstractSuppliedCopyright
  Label:
  Group: Ab
  Data: <i>Copyright of Multimedia Tools & Applications is the property of Springer Nature and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.)
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=190506892
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1007/s11042-025-21132-2
    Languages:
      – Code: eng
        Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 34
        StartPage: 50437
    Subjects:
      – SubjectFull: Video compression
        Type: general
      – SubjectFull: Deep learning
        Type: general
      – SubjectFull: Statistical accuracy
        Type: general
      – SubjectFull: Prediction models
        Type: general
      – SubjectFull: Lossy data compression
        Type: general
      – SubjectFull: Generative adversarial networks
        Type: general
    Titles:
      – TitleFull: An intuitive IPT-IPA based deep learning approach for video compression techniques.
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Naik, Mudhavath Ramesh
      – PersonEntity:
          Name:
            NameFull: Kumar, Jayendra
      – PersonEntity:
          Name:
            NameFull: Yadav, Arvind R.
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 27
              M: 12
              Text: Dec2025
              Type: published
              Y: 2025
          Identifiers:
            – Type: issn-print
              Value: 13807501
          Numbering:
            – Type: volume
              Value: 84
            – Type: issue
              Value: 42
          Titles:
            – TitleFull: Multimedia Tools & Applications
              Type: main
ResultId 1