Tran‐GCN: A Transformer‐Enhanced Graph Convolutional Network for Person Re‐Identification in Monitoring Videos.

Saved in:
Bibliographic Details
Title: Tran‐GCN: A Transformer‐Enhanced Graph Convolutional Network for Person Re‐Identification in Monitoring Videos.
Authors: Hong, Xiaobin1 (AUTHOR), Adam, Tarmizi1 (AUTHOR) Tarmizi.adam@utm.my, Ghazali, Masitah1 (AUTHOR)
Source: IET Computer Vision (Wiley-Blackwell). Jan2025, Vol. 19 Issue 1, p1-14. 14p.
Subjects: Computer vision, Pose estimation (Computer vision), Deep learning, Feature extraction, Graph neural networks, Transformer models, Video surveillance
Abstract: Person re‐identification (Re‐ID) has gained popularity in computer vision, enabling cross‐camera pedestrian recognition. Although the development of deep learning has provided a robust technical foundation for person Re‐ID research, most existing person Re‐ID methods overlook the potential relationships among local person features, failing to adequately address the impact of pedestrian pose variations and local body parts occlusion. Therefore, we propose a transformer‐enhanced graph convolutional network (Tran‐GCN) model to improve person re‐identification performance in monitoring videos. The model comprises four key components: (1) a pose estimation learning branch is utilised to estimate pedestrian pose information and inherent skeletal structure data, extracting pedestrian key point information; (2) a transformer learning branch learns the global dependencies between fine‐grained and semantically meaningful local person features; (3) a convolution learning branch uses the basic ResNet architecture to extract the person's fine‐grained local features; and (4) a Graph convolutional module (GCM) integrates local feature information, global feature information and body information for more effective person identification after fusion. Quantitative and qualitative analysis experiments conducted on three different datasets (Market‐1501, DukeMTMC‐ReID and MSMT17) demonstrate that the Tran‐GCN model can more accurately capture discriminative person features in monitoring videos, significantly improving identification accuracy. [ABSTRACT FROM AUTHOR]
Copyright of IET Computer Vision (Wiley-Blackwell) is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
Database: Engineering Source
FullText Text:
  Availability: 0
Header DbId: egs
DbLabel: Engineering Source
An: 190526555
AccessLevel: 6
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Tran‐GCN: A Transformer‐Enhanced Graph Convolutional Network for Person Re‐Identification in Monitoring Videos.
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Hong%2C+Xiaobin%22">Hong, Xiaobin</searchLink><relatesTo>1</relatesTo> (AUTHOR)<br /><searchLink fieldCode="AR" term="%22Adam%2C+Tarmizi%22">Adam, Tarmizi</searchLink><relatesTo>1</relatesTo> (AUTHOR)<i> Tarmizi.adam@utm.my</i><br /><searchLink fieldCode="AR" term="%22Ghazali%2C+Masitah%22">Ghazali, Masitah</searchLink><relatesTo>1</relatesTo> (AUTHOR)
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="JN" term="%22IET+Computer+Vision+%28Wiley-Blackwell%29%22">IET Computer Vision (Wiley-Blackwell)</searchLink>. Jan2025, Vol. 19 Issue 1, p1-14. 14p.
– Name: Subject
  Label: Subjects
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Computer+vision%22">Computer vision</searchLink><br /><searchLink fieldCode="DE" term="%22Pose+estimation+%28Computer+vision%29%22">Pose estimation (Computer vision)</searchLink><br /><searchLink fieldCode="DE" term="%22Deep+learning%22">Deep learning</searchLink><br /><searchLink fieldCode="DE" term="%22Feature+extraction%22">Feature extraction</searchLink><br /><searchLink fieldCode="DE" term="%22Graph+neural+networks%22">Graph neural networks</searchLink><br /><searchLink fieldCode="DE" term="%22Transformer+models%22">Transformer models</searchLink><br /><searchLink fieldCode="DE" term="%22Video+surveillance%22">Video surveillance</searchLink>
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: Person re‐identification (Re‐ID) has gained popularity in computer vision, enabling cross‐camera pedestrian recognition. Although the development of deep learning has provided a robust technical foundation for person Re‐ID research, most existing person Re‐ID methods overlook the potential relationships among local person features, failing to adequately address the impact of pedestrian pose variations and local body parts occlusion. Therefore, we propose a transformer‐enhanced graph convolutional network (Tran‐GCN) model to improve person re‐identification performance in monitoring videos. The model comprises four key components: (1) a pose estimation learning branch is utilised to estimate pedestrian pose information and inherent skeletal structure data, extracting pedestrian key point information; (2) a transformer learning branch learns the global dependencies between fine‐grained and semantically meaningful local person features; (3) a convolution learning branch uses the basic ResNet architecture to extract the person's fine‐grained local features; and (4) a Graph convolutional module (GCM) integrates local feature information, global feature information and body information for more effective person identification after fusion. Quantitative and qualitative analysis experiments conducted on three different datasets (Market‐1501, DukeMTMC‐ReID and MSMT17) demonstrate that the Tran‐GCN model can more accurately capture discriminative person features in monitoring videos, significantly improving identification accuracy. [ABSTRACT FROM AUTHOR]
– Name: AbstractSuppliedCopyright
  Label:
  Group: Ab
  Data: <i>Copyright of IET Computer Vision (Wiley-Blackwell) is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.)
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=190526555
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1049/cvi2.70025
    Languages:
      – Code: eng
        Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 14
        StartPage: 1
    Subjects:
      – SubjectFull: Computer vision
        Type: general
      – SubjectFull: Pose estimation (Computer vision)
        Type: general
      – SubjectFull: Deep learning
        Type: general
      – SubjectFull: Feature extraction
        Type: general
      – SubjectFull: Graph neural networks
        Type: general
      – SubjectFull: Transformer models
        Type: general
      – SubjectFull: Video surveillance
        Type: general
    Titles:
      – TitleFull: Tran‐GCN: A Transformer‐Enhanced Graph Convolutional Network for Person Re‐Identification in Monitoring Videos.
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Hong, Xiaobin
      – PersonEntity:
          Name:
            NameFull: Adam, Tarmizi
      – PersonEntity:
          Name:
            NameFull: Ghazali, Masitah
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 01
              Text: Jan2025
              Type: published
              Y: 2025
          Identifiers:
            – Type: issn-print
              Value: 17519632
          Numbering:
            – Type: volume
              Value: 19
            – Type: issue
              Value: 1
          Titles:
            – TitleFull: IET Computer Vision (Wiley-Blackwell)
              Type: main
ResultId 1