Enhanced Cross-Modal Hashing via Hybrid Distillation and Structural Refinement.
Saved in:
| Title: | Enhanced Cross-Modal Hashing via Hybrid Distillation and Structural Refinement. |
|---|---|
| Authors: | Liu, Xiaoqing1 ft_liuxiaoqing@mail.scut.edu.cn, Yu, Zhiwen2 zhwyu@scut.edu.cn, Yang, Kaixiang2 yangkx@scut.edu.cn, Yu, Jun3 yujun@hit.edu.cn, Zeng, Huanqiang4 zeng0043@hqu.edu.cn, Philip Chen, C. L.2 Philip.Chen@ieee.org |
| Source: | IEEE Transactions on Image Processing. 2025, Vol. 34, p7138-7151. 14p. |
| Subjects: | Machine learning, Hashing, Machine theory, Message authentication codes, Electronic file management |
| Abstract: | Since cross-modal hashing requires minimal storage and computation, it is becoming increasingly popular with the exponential growth of multimedia content on the internet. However, the lack of accurate supervisory data has curtailed the effectiveness of unsupervised hashing techniques. Conversely, supervised hashing strategies necessitate considerable human and financial resources for data annotation. To address this limitation, we propose a novel semi-supervised cross-modal hashing method called Enhanced Cross-Modal Hashing via Hybrid Distillation and Structural Refinement (HDSR). Specifically, we first learn the features of inter-modal and inter-instance similarity relationships through pointwise semantic alignment and listwise similarity partial order learning, respectively, to extract refined structural representations from partially labeled data. Secondly, by fusing inter-modal similarity to construct higher-order affinity matrices, we precisely delineate the semantic correlation information across cross-modal data, facilitating stable self-supervised training of unlabeled data through the application of momentum fusion strategies. Finally, the refined structural representation of labeled data is transferred into unlabeled branches through hybrid distillation, enhancing the performance of cross-modal hash learning by generating compact and accurate hash codes. The proposed HDSR is compared with several state-of-the-art deep cross-modal hashing methods on three widely used benchmark databases, and the experimental results verify its efficiency and superiority. [ABSTRACT FROM AUTHOR] |
| Copyright of IEEE Transactions on Image Processing is the property of IEEE and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.) | |
| Database: | Engineering Source |
| FullText | Text: Availability: 0 |
|---|---|
| Header | DbId: egs DbLabel: Engineering Source An: 191897310 AccessLevel: 6 PubType: Academic Journal PubTypeId: academicJournal PreciseRelevancyScore: 0 |
| IllustrationInfo | |
| Items | – Name: Title Label: Title Group: Ti Data: Enhanced Cross-Modal Hashing via Hybrid Distillation and Structural Refinement. – Name: Author Label: Authors Group: Au Data: <searchLink fieldCode="AR" term="%22Liu%2C+Xiaoqing%22">Liu, Xiaoqing</searchLink><relatesTo>1</relatesTo><i> ft_liuxiaoqing@mail.scut.edu.cn</i><br /><searchLink fieldCode="AR" term="%22Yu%2C+Zhiwen%22">Yu, Zhiwen</searchLink><relatesTo>2</relatesTo><i> zhwyu@scut.edu.cn</i><br /><searchLink fieldCode="AR" term="%22Yang%2C+Kaixiang%22">Yang, Kaixiang</searchLink><relatesTo>2</relatesTo><i> yangkx@scut.edu.cn</i><br /><searchLink fieldCode="AR" term="%22Yu%2C+Jun%22">Yu, Jun</searchLink><relatesTo>3</relatesTo><i> yujun@hit.edu.cn</i><br /><searchLink fieldCode="AR" term="%22Zeng%2C+Huanqiang%22">Zeng, Huanqiang</searchLink><relatesTo>4</relatesTo><i> zeng0043@hqu.edu.cn</i><br /><searchLink fieldCode="AR" term="%22Philip+Chen%2C+C%2E+L%2E%22">Philip Chen, C. L.</searchLink><relatesTo>2</relatesTo><i> Philip.Chen@ieee.org</i> – Name: TitleSource Label: Source Group: Src Data: <searchLink fieldCode="JN" term="%22IEEE+Transactions+on+Image+Processing%22">IEEE Transactions on Image Processing</searchLink>. 2025, Vol. 34, p7138-7151. 14p. – Name: Subject Label: Subjects Group: Su Data: <searchLink fieldCode="DE" term="%22Machine+learning%22">Machine learning</searchLink><br /><searchLink fieldCode="DE" term="%22Hashing%22">Hashing</searchLink><br /><searchLink fieldCode="DE" term="%22Machine+theory%22">Machine theory</searchLink><br /><searchLink fieldCode="DE" term="%22Message+authentication+codes%22">Message authentication codes</searchLink><br /><searchLink fieldCode="DE" term="%22Electronic+file+management%22">Electronic file management</searchLink> – Name: Abstract Label: Abstract Group: Ab Data: Since cross-modal hashing requires minimal storage and computation, it is becoming increasingly popular with the exponential growth of multimedia content on the internet. However, the lack of accurate supervisory data has curtailed the effectiveness of unsupervised hashing techniques. Conversely, supervised hashing strategies necessitate considerable human and financial resources for data annotation. To address this limitation, we propose a novel semi-supervised cross-modal hashing method called Enhanced Cross-Modal Hashing via Hybrid Distillation and Structural Refinement (HDSR). Specifically, we first learn the features of inter-modal and inter-instance similarity relationships through pointwise semantic alignment and listwise similarity partial order learning, respectively, to extract refined structural representations from partially labeled data. Secondly, by fusing inter-modal similarity to construct higher-order affinity matrices, we precisely delineate the semantic correlation information across cross-modal data, facilitating stable self-supervised training of unlabeled data through the application of momentum fusion strategies. Finally, the refined structural representation of labeled data is transferred into unlabeled branches through hybrid distillation, enhancing the performance of cross-modal hash learning by generating compact and accurate hash codes. The proposed HDSR is compared with several state-of-the-art deep cross-modal hashing methods on three widely used benchmark databases, and the experimental results verify its efficiency and superiority. [ABSTRACT FROM AUTHOR] – Name: AbstractSuppliedCopyright Label: Group: Ab Data: <i>Copyright of IEEE Transactions on Image Processing is the property of IEEE and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.) |
| PLink | https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=191897310 |
| RecordInfo | BibRecord: BibEntity: Identifiers: – Type: doi Value: 10.1109/TIP.2025.3613942 Languages: – Code: eng Text: English PhysicalDescription: Pagination: PageCount: 14 StartPage: 7138 Subjects: – SubjectFull: Machine learning Type: general – SubjectFull: Hashing Type: general – SubjectFull: Machine theory Type: general – SubjectFull: Message authentication codes Type: general – SubjectFull: Electronic file management Type: general Titles: – TitleFull: Enhanced Cross-Modal Hashing via Hybrid Distillation and Structural Refinement. Type: main BibRelationships: HasContributorRelationships: – PersonEntity: Name: NameFull: Liu, Xiaoqing – PersonEntity: Name: NameFull: Yu, Zhiwen – PersonEntity: Name: NameFull: Yang, Kaixiang – PersonEntity: Name: NameFull: Yu, Jun – PersonEntity: Name: NameFull: Zeng, Huanqiang – PersonEntity: Name: NameFull: Philip Chen, C. L. IsPartOfRelationships: – BibEntity: Dates: – D: 01 M: 01 Text: 2025 Type: published Y: 2025 Identifiers: – Type: issn-print Value: 10577149 Numbering: – Type: volume Value: 34 Titles: – TitleFull: IEEE Transactions on Image Processing Type: main |
| ResultId | 1 |