TAIS-Net: Time adaptive implicit sampling diffusion model for arbitrary-scale UAV video super-resolution.

Saved in:
Bibliographic Details
Title: TAIS-Net: Time adaptive implicit sampling diffusion model for arbitrary-scale UAV video super-resolution.
Authors: Li, Wenke1 (AUTHOR) ieulwk@163.com, Dai, Chenguang1 (AUTHOR), Fan, Huixin1 (AUTHOR), Zhao, Zhen1 (AUTHOR), Li, Kangrui1 (AUTHOR), Sun, Yifan1 (AUTHOR), Zhang, Yongsheng1 (AUTHOR), Wang, Longguang2 (AUTHOR), Wang, Hanyun1,3 (AUTHOR) wanghanyun@mail.sysu.edu.cn
Source: ISPRS Journal of Photogrammetry & Remote Sensing. Aug2026, Vol. 238, p176-190. 15p.
Subjects: Optical flow, Implicit functions, Probabilistic generative models, High resolution imaging
Abstract: Emerging applications urgently require unmanned aerial vehicle (UAV) videos with high clarity and rich structural details. However, limitations in imaging sensors and non-ideal conditions often result in blurred videos lacking sufficient details. Recent studies indicate diffusion models have strong potential for natural scene video super-resolution (SR). However, the application of diffusion models to UAV video SR remains underexplored, particularly in the context of arbitrary-scale reconstruction. To tackle the challenges in UAV video SR at arbitrary scales, this work introduces a time adaptive implicit sampling diffusion model TAIS-Net. First, to achieve robust temporal alignment under large camera motions and texture scarcity commonly encountered in UAV videos, we use a reference frame alignment module to compensate previously reconstructed frames by incorporating motion cues estimated from a pre-trained optical flow model. Second, to enable scale-flexible reconstruction while preserving fine geometric details for UAV videos under arbitrary upscaling factors, we introduce an implicit denoising U-Net to learn latent features by leveraging implicit neural representations and a temporal conditioning module. Third, to reduce the prediction error propagation during sampling, we introduce a historical sample gain module that dynamically corrects and refines latent features at each sampling step, thereby suppressing temporal artifacts such as flickering or drifting. Finally, an arbitrary-scale implicit decoder is used to reconstruct high-resolution videos directly from these enhanced latent features, avoiding scale-dependent blurring while preserving high-frequency details in UAV videos. To validate the effectiveness of TAIS-Net, we construct two visible and one infrared UAV video SR datasets. The experimental results on these datasets demonstrate that TAIS-Net attains competitive performance relative to existing methods in terms of reconstruction quality, perceptual quality and temporal consistency. Moreover, experiments on various motion intensities and additional Gaussian blur beyond conventional bicubic downsampling degradation without requiring model retraining also demonstrate the superiority of our TAIS-Net in maintaining perceptual and temporal consistency under real-world UAV scenarios. The code and datasets are available at https://github.com/cyber-lwk/TAIS-Net. [ABSTRACT FROM AUTHOR]
Copyright of ISPRS Journal of Photogrammetry & Remote Sensing is the property of Elsevier B.V. and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
Database: Engineering Source
FullText Text:
  Availability: 0
Header DbId: egs
DbLabel: Engineering Source
An: 194367884
AccessLevel: 6
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: TAIS-Net: Time adaptive implicit sampling diffusion model for arbitrary-scale UAV video super-resolution.
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Li%2C+Wenke%22">Li, Wenke</searchLink><relatesTo>1</relatesTo> (AUTHOR)<i> ieulwk@163.com</i><br /><searchLink fieldCode="AR" term="%22Dai%2C+Chenguang%22">Dai, Chenguang</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Fan%2C+Huixin%22">Fan, Huixin</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Zhao%2C+Zhen%22">Zhao, Zhen</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Li%2C+Kangrui%22">Li, Kangrui</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Sun%2C+Yifan%22">Sun, Yifan</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Zhang%2C+Yongsheng%22">Zhang, Yongsheng</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Wang%2C+Longguang%22">Wang, Longguang</searchLink><relatesTo>2</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Wang%2C+Hanyun%22">Wang, Hanyun</searchLink><relatesTo>1,3</relatesTo> (AUTHOR)<i> wanghanyun@mail.sysu.edu.cn</i>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="JN" term="%22ISPRS+Journal+of+Photogrammetry+%26+Remote+Sensing%22">ISPRS Journal of Photogrammetry & Remote Sensing</searchLink>. Aug2026, Vol. 238, p176-190. 15p.
– Name: Subject
  Label: Subjects
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Optical+flow%22">Optical flow</searchLink><br /><searchLink fieldCode="DE" term="%22Implicit+functions%22">Implicit functions</searchLink><br /><searchLink fieldCode="DE" term="%22Probabilistic+generative+models%22">Probabilistic generative models</searchLink><br /><searchLink fieldCode="DE" term="%22High+resolution+imaging%22">High resolution imaging</searchLink>
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: Emerging applications urgently require unmanned aerial vehicle (UAV) videos with high clarity and rich structural details. However, limitations in imaging sensors and non-ideal conditions often result in blurred videos lacking sufficient details. Recent studies indicate diffusion models have strong potential for natural scene video super-resolution (SR). However, the application of diffusion models to UAV video SR remains underexplored, particularly in the context of arbitrary-scale reconstruction. To tackle the challenges in UAV video SR at arbitrary scales, this work introduces a time adaptive implicit sampling diffusion model TAIS-Net. First, to achieve robust temporal alignment under large camera motions and texture scarcity commonly encountered in UAV videos, we use a reference frame alignment module to compensate previously reconstructed frames by incorporating motion cues estimated from a pre-trained optical flow model. Second, to enable scale-flexible reconstruction while preserving fine geometric details for UAV videos under arbitrary upscaling factors, we introduce an implicit denoising U-Net to learn latent features by leveraging implicit neural representations and a temporal conditioning module. Third, to reduce the prediction error propagation during sampling, we introduce a historical sample gain module that dynamically corrects and refines latent features at each sampling step, thereby suppressing temporal artifacts such as flickering or drifting. Finally, an arbitrary-scale implicit decoder is used to reconstruct high-resolution videos directly from these enhanced latent features, avoiding scale-dependent blurring while preserving high-frequency details in UAV videos. To validate the effectiveness of TAIS-Net, we construct two visible and one infrared UAV video SR datasets. The experimental results on these datasets demonstrate that TAIS-Net attains competitive performance relative to existing methods in terms of reconstruction quality, perceptual quality and temporal consistency. Moreover, experiments on various motion intensities and additional Gaussian blur beyond conventional bicubic downsampling degradation without requiring model retraining also demonstrate the superiority of our TAIS-Net in maintaining perceptual and temporal consistency under real-world UAV scenarios. The code and datasets are available at https://github.com/cyber-lwk/TAIS-Net. [ABSTRACT FROM AUTHOR]
– Name: AbstractSuppliedCopyright
  Label:
  Group: Ab
  Data: <i>Copyright of ISPRS Journal of Photogrammetry & Remote Sensing is the property of Elsevier B.V. and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.)
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=194367884
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1016/j.isprsjprs.2026.04.060
    Languages:
      – Code: eng
        Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 15
        StartPage: 176
    Subjects:
      – SubjectFull: Optical flow
        Type: general
      – SubjectFull: Implicit functions
        Type: general
      – SubjectFull: Probabilistic generative models
        Type: general
      – SubjectFull: High resolution imaging
        Type: general
    Titles:
      – TitleFull: TAIS-Net: Time adaptive implicit sampling diffusion model for arbitrary-scale UAV video super-resolution.
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Li, Wenke
      – PersonEntity:
          Name:
            NameFull: Dai, Chenguang
      – PersonEntity:
          Name:
            NameFull: Fan, Huixin
      – PersonEntity:
          Name:
            NameFull: Zhao, Zhen
      – PersonEntity:
          Name:
            NameFull: Li, Kangrui
      – PersonEntity:
          Name:
            NameFull: Sun, Yifan
      – PersonEntity:
          Name:
            NameFull: Zhang, Yongsheng
      – PersonEntity:
          Name:
            NameFull: Wang, Longguang
      – PersonEntity:
          Name:
            NameFull: Wang, Hanyun
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 08
              Text: Aug2026
              Type: published
              Y: 2026
          Identifiers:
            – Type: issn-print
              Value: 09242716
          Numbering:
            – Type: volume
              Value: 238
          Titles:
            – TitleFull: ISPRS Journal of Photogrammetry & Remote Sensing
              Type: main
ResultId 1