{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,24]],"date-time":"2026-01-24T11:02:05Z","timestamp":1769252525237,"version":"3.49.0"},"reference-count":65,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2023,7,14]],"date-time":"2023-07-14T00:00:00Z","timestamp":1689292800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"STI 2030\u2014Major Projects","award":["2022ZD0208700"],"award-info":[{"award-number":["2022ZD0208700"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["61702491"],"award-info":[{"award-number":["61702491"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2023,11,30]]},"abstract":"<jats:p>Person re-identification (re-ID) aims to match pedestrian pairs captured from different cameras. Recently, various attribute-based models have been proposed to combine the pedestrian attribute as an auxiliary semantic information to learn a more discriminative pedestrian representation. However, these methods usually directly concatenate the visual branch and attribute branch embeddings as the final pedestrian representation, which ignores the semantic relation between the pedestrian revealed by attribute similarity. To capture and explore such semantic relation, we propose a unified pedestrian representation framework, called Visual Attribute Graph Embedding Network (VAGEN), to simultaneously learn attribute and visual representation. We unify the visual embedding and attribute similarity into a Visual Attribute Graph, where pedestrian is considered as a node and attribute similarity as an edge. Then, we learn graph node embedding to generate pedestrian representation through Graph Neural Network. Except for this unified representation for visual and attribute embeddings, VAGEN also conducts implicitly hard example mining for visual similar false-positive results, which has not been explored yet among existing attribute-based methods. We conduct extensive empirical studies on several person re-ID datasets to evaluate our proposed algorithm from different aspects. The results show that our proposed method outperforms state-of-the-art techniques with considerable margins.<\/jats:p>","DOI":"10.1145\/3487044","type":"journal-article","created":{"date-parts":[[2023,3,27]],"date-time":"2023-03-27T12:12:28Z","timestamp":1679919148000},"page":"1-20","update-policy":"https:\/\/2.zoppoz.workers.dev:443\/https\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":11,"title":["Learning Semantic Representation on Visual Attribute Graph for Person Re-identification and Beyond"],"prefix":"10.1145","volume":"19","author":[{"ORCID":"https:\/\/2.zoppoz.workers.dev:443\/https\/orcid.org\/0000-0001-8108-6886","authenticated-orcid":false,"given":"Geyu","family":"Tang","sequence":"first","affiliation":[{"name":"Institute of Microelectronics, Chinese Academy of Sciences, and University of Chinese Academy of Sciences, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/2.zoppoz.workers.dev:443\/https\/orcid.org\/0000-0002-4660-8092","authenticated-orcid":false,"given":"Xingyu","family":"Gao","sequence":"additional","affiliation":[{"name":"Institute of Microelectronics, Chinese Academy of Sciences, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/2.zoppoz.workers.dev:443\/https\/orcid.org\/0000-0002-4989-7109","authenticated-orcid":false,"given":"Zhenyu","family":"Chen","sequence":"additional","affiliation":[{"name":"Big Data Center, State Grid Corporation of China, and China Electric Power Research Institute, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,7,14]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"3908","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Ahmed Ejaz","year":"2015","unstructured":"Ejaz Ahmed, Michael Jones, and Tim K. Marks. 2015. An improved deep learning architecture for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3908\u20133916."},{"key":"e_1_3_1_3_2","first-page":"0","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops","author":"Bao Liqiang","year":"2019","unstructured":"Liqiang Bao, Bingpeng Ma, Hong Chang, and Xilin Chen. 2019. Masked graph attention network for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 0\u20130."},{"key":"e_1_3_1_4_2","article-title":"Spectral networks and locally connected networks on graphs","author":"Bruna Joan","year":"2013","unstructured":"Joan Bruna, Wojciech Zaremba, Arthur Szlam, and Yann LeCun. 2013. Spectral networks and locally connected networks on graphs. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1312.6203.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1312.6203"},{"key":"e_1_3_1_5_2","first-page":"2109","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Chang Xiaobin","year":"2018","unstructured":"Xiaobin Chang, Timothy M. Hospedales, and Tao Xiang. 2018. Multi-level factorisation net for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 2109\u20132118."},{"key":"e_1_3_1_6_2","first-page":"8649","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Chen Dapeng","year":"2018","unstructured":"Dapeng Chen, Dan Xu, Hongsheng Li, Nicu Sebe, and Xiaogang Wang. 2018. Group consistent similarity learning via deep crf for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 8649\u20138658."},{"key":"e_1_3_1_7_2","doi-asserted-by":"crossref","first-page":"94","DOI":"10.1016\/j.patcog.2018.05.007","article-title":"Deep feature learning via structured graph Laplacian embedding for person re-identification","volume":"82","author":"Cheng De","year":"2018","unstructured":"De Cheng, Yihong Gong, Xiaojun Chang, Weiwei Shi, Alexander Hauptmann, and Nanning Zheng. 2018. Deep feature learning via structured graph Laplacian embedding for person re-identification. Pattern Recogn. 82 (2018), 94\u2013104.","journal-title":"Pattern Recogn."},{"key":"e_1_3_1_8_2","first-page":"3691","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Dai Zuozhuo","year":"2019","unstructured":"Zuozhuo Dai, Mingqiang Chen, Xiaodong Gu, Siyu Zhu, and Ping Tan. 2019. Batch DropBlock network for person re-identification and beyond. In Proceedings of the IEEE International Conference on Computer Vision. 3691\u20133701."},{"key":"e_1_3_1_9_2","first-page":"3844","volume-title":"Advances in Neural Information Processing Systems","author":"Defferrard Micha\u00ebl","year":"2016","unstructured":"Micha\u00ebl Defferrard, Xavier Bresson, and Pierre Vandergheynst. 2016. Convolutional neural networks on graphs with fast localized spectral filtering. In Advances in Neural Information Processing Systems. MIT Press, 3844\u20133852."},{"key":"e_1_3_1_10_2","first-page":"248","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Deng Jia","year":"2009","unstructured":"Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009. Imagenet: A large-scale hierarchical image database. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. IEEE, 248\u2013255."},{"issue":"9","key":"e_1_3_1_11_2","doi-asserted-by":"crossref","first-page":"1627","DOI":"10.1109\/TPAMI.2009.167","article-title":"Object detection with discriminatively trained part-based models","volume":"32","author":"Felzenszwalb Pedro F.","year":"2009","unstructured":"Pedro F. Felzenszwalb, Ross B. Girshick, David McAllester, and Deva Ramanan. 2009. Object detection with discriminatively trained part-based models. IEEE Trans. Pattern Anal. Mach. Intell. 32, 9 (2009), 1627\u20131645.","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"e_1_3_1_12_2","first-page":"8295","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","volume":"33","author":"Fu Yang","year":"2019","unstructured":"Yang Fu, Yunchao Wei, Yuqian Zhou, Honghui Shi, Gao Huang, Xinchao Wang, Zhiqiang Yao, and Thomas Huang. 2019. Horizontal pyramid matching for person re-identification. In Proceedings of the AAAI Conference on Artificial Intelligence, Vol. 33. 8295\u20138302."},{"key":"e_1_3_1_13_2","first-page":"1024","volume-title":"Advances in Neural Information Processing Systems","author":"Hamilton Will","year":"2017","unstructured":"Will Hamilton, Zhitao Ying, and Jure Leskovec. 2017. Inductive representation learning on large graphs. In Advances in Neural Information Processing Systems. MIT Press, 1024\u20131034."},{"key":"e_1_3_1_14_2","doi-asserted-by":"crossref","first-page":"2040","DOI":"10.1145\/3240508.3240550","volume-title":"Proceedings of the 26th ACM International Conference on Multimedia","author":"Han Kai","year":"2018","unstructured":"Kai Han, Jianyuan Guo, Chao Zhang, and Mingjian Zhu. 2018. Attribute-aware attention model for fine-grained representation learning. In Proceedings of the 26th ACM International Conference on Multimedia. 2040\u20132048."},{"key":"e_1_3_1_15_2","first-page":"770","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"He Kaiming","year":"2016","unstructured":"Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016. Deep residual learning for image recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 770\u2013778."},{"key":"e_1_3_1_16_2","article-title":"In defense of the triplet loss for person re-identification","author":"Hermans Alexander","year":"2017","unstructured":"Alexander Hermans, Lucas Beyer, and Bastian Leibe. 2017. In defense of the triplet loss for person re-identification. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1703.07737.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1703.07737"},{"issue":"2","key":"e_1_3_1_17_2","doi-asserted-by":"crossref","first-page":"181","DOI":"10.1080\/00031305.1998.10480559","article-title":"Violin plots: A box plot-density trace synergism","volume":"52","author":"Hintze Jerry L.","year":"1998","unstructured":"Jerry L. Hintze and Ray D. Nelson. 1998. Violin plots: A box plot-density trace synergism. Amer. Stat. 52, 2 (1998), 181\u2013184.","journal-title":"Amer. Stat."},{"key":"e_1_3_1_18_2","first-page":"4700","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Huang Gao","year":"2017","unstructured":"Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q. Weinberger. 2017. Densely connected convolutional networks. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 4700\u20134708."},{"key":"e_1_3_1_19_2","article-title":"Batch normalization: Accelerating deep network training by reducing internal covariate shift","author":"Ioffe Sergey","year":"2015","unstructured":"Sergey Ioffe and Christian Szegedy. 2015. Batch normalization: Accelerating deep network training by reducing internal covariate shift. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1502.03167.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1502.03167"},{"key":"e_1_3_1_20_2","article-title":"AttKGCN: Attribute knowledge graph convolutional network for person re-identification","author":"Jiang Bo","year":"2019","unstructured":"Bo Jiang, Xixi Wang, and Jin Tang. 2019. AttKGCN: Attribute knowledge graph convolutional network for person re-identification. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1911.10544.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1911.10544"},{"key":"e_1_3_1_21_2","first-page":"736","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201918)","author":"Kim Wonsik","year":"2018","unstructured":"Wonsik Kim, Bhavya Goyal, Kunal Chawla, Jungmin Lee, and Keunjoo Kwon. 2018. Attention-based ensemble for deep metric learning. In Proceedings of the European Conference on Computer Vision (ECCV\u201918). 736\u2013751."},{"key":"e_1_3_1_22_2","article-title":"Adam: A method for stochastic optimization","author":"Kingma Diederik P.","year":"2014","unstructured":"Diederik P. Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1412.6980.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1412.6980"},{"key":"e_1_3_1_23_2","article-title":"Semi-supervised classification with graph convolutional networks","author":"Kipf Thomas N.","year":"2016","unstructured":"Thomas N. Kipf and Max Welling. 2016. Semi-supervised classification with graph convolutional networks. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1609.02907.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1609.02907"},{"issue":"7553","key":"e_1_3_1_24_2","doi-asserted-by":"crossref","first-page":"436","DOI":"10.1038\/nature14539","article-title":"Deep learning","volume":"521","author":"LeCun Yann","year":"2015","unstructured":"Yann LeCun, Yoshua Bengio, and Geoffrey Hinton. 2015. Deep learning. Nature 521, 7553 (2015), 436\u2013444.","journal-title":"Nature"},{"key":"e_1_3_1_25_2","article-title":"A survey of open-world person re-identification","author":"Leng Qingming","year":"2019","unstructured":"Qingming Leng, Mang Ye, and Qi Tian. 2019. A survey of open-world person re-identification. IEEE Trans. Circ. Syst. Video Technol. (2019).","journal-title":"IEEE Trans. Circ. Syst. Video Technol."},{"issue":"1","key":"e_1_3_1_26_2","doi-asserted-by":"crossref","first-page":"97","DOI":"10.1109\/TSP.2018.2879624","article-title":"Cayleynets: Graph convolutional neural networks with complex rational spectral filters","volume":"67","author":"Levie Ron","year":"2018","unstructured":"Ron Levie, Federico Monti, Xavier Bresson, and Michael M. Bronstein. 2018. Cayleynets: Graph convolutional neural networks with complex rational spectral filters. IEEE Trans. Signal Process. 67, 1 (2018), 97\u2013109.","journal-title":"IEEE Trans. Signal Process."},{"key":"e_1_3_1_27_2","first-page":"384","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Li Dangwei","year":"2017","unstructured":"Dangwei Li, Xiaotang Chen, Zhang Zhang, and Kaiqi Huang. 2017. Learning deep context-aware features over body and latent parts for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 384\u2013393."},{"key":"e_1_3_1_28_2","doi-asserted-by":"crossref","first-page":"107016","DOI":"10.1016\/j.patcog.2019.107016","article-title":"Attributes-aided part detection and refinement for person re-identification","volume":"97","author":"Li Shuzhao","year":"2020","unstructured":"Shuzhao Li, Huimin Yu, and Roland Hu. 2020. Attributes-aided part detection and refinement for person re-identification. Pattern Recogn. 97 (2020), 107016.","journal-title":"Pattern Recogn."},{"key":"e_1_3_1_29_2","doi-asserted-by":"crossref","first-page":"151","DOI":"10.1016\/j.patcog.2019.06.006","article-title":"Improving person re-identification by attribute and identity learning","volume":"95","author":"Lin Yutian","year":"2019","unstructured":"Yutian Lin, Liang Zheng, Zhedong Zheng, Yu Wu, Zhilan Hu, Chenggang Yan, and Yi Yang. 2019. Improving person re-identification by attribute and identity learning. Pattern Recogn. 95 (2019), 151\u2013161.","journal-title":"Pattern Recogn."},{"key":"e_1_3_1_30_2","first-page":"1096","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Liu Ziwei","year":"2016","unstructured":"Ziwei Liu, Ping Luo, Shi Qiu, Xiaogang Wang, and Xiaoou Tang. 2016. Deepfashion: Powering robust clothes recognition and retrieval with rich annotations. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 1096\u20131104."},{"issue":"10","key":"e_1_3_1_31_2","doi-asserted-by":"crossref","first-page":"2597","DOI":"10.1109\/TMM.2019.2958756","article-title":"A strong baseline and batch normalization neck for deep person re-identification","volume":"22","author":"Luo Hao","year":"2020","unstructured":"Hao Luo, Wei Jiang, Youzhi Gu, Fuxu Liu, Xingyu Liao, Shenqi Lai, and Jianyang Gu. 2020. A strong baseline and batch normalization neck for deep person re-identification. IEEE Trans. Multimedia 22, 10 (2020), 2597\u20132609.","journal-title":"IEEE Trans. Multimedia"},{"key":"e_1_3_1_32_2","doi-asserted-by":"crossref","first-page":"53","DOI":"10.1016\/j.patcog.2019.05.028","article-title":"AlignedReID++: Dynamically matching local information for person re-identification","volume":"94","author":"Luo Hao","year":"2019","unstructured":"Hao Luo, Wei Jiang, Xuan Zhang, Xing Fan, Jingjing Qian, and Chi Zhang. 2019. AlignedReID++: Dynamically matching local information for person re-identification. Pattern Recogn. 94 (2019), 53\u201361.","journal-title":"Pattern Recogn."},{"key":"e_1_3_1_33_2","first-page":"2579","article-title":"Visualizing data using t-SNE","volume":"9","author":"Maaten Laurens van der","year":"2008","unstructured":"Laurens van der Maaten and Geoffrey Hinton. 2008. Visualizing data using t-SNE. J. Mach. Learn. Res. 9(Nov.2008), 2579\u20132605.","journal-title":"J. Mach. Learn. Res."},{"key":"e_1_3_1_34_2","first-page":"5399","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Qian Xuelin","year":"2017","unstructured":"Xuelin Qian, Yanwei Fu, Yu-Gang Jiang, Tao Xiang, and Xiangyang Xue. 2017. Multi-scale deep learning architectures for person re-identification. In Proceedings of the IEEE International Conference on Computer Vision. 5399\u20135408."},{"key":"e_1_3_1_35_2","first-page":"17","volume-title":"Proceedings of the European Conference on Computer Vision","author":"Ristani Ergys","year":"2016","unstructured":"Ergys Ristani, Francesco Solera, Roger Zou, Rita Cucchiara, and Carlo Tomasi. 2016. Performance measures and a data set for multi-target, multi-camera tracking. In Proceedings of the European Conference on Computer Vision. Springer, 17\u201335."},{"issue":"1","key":"e_1_3_1_36_2","doi-asserted-by":"crossref","first-page":"61","DOI":"10.1109\/TNN.2008.2005605","article-title":"The graph neural network model","volume":"20","author":"Scarselli Franco","year":"2008","unstructured":"Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini. 2008. The graph neural network model. IEEE Trans. Neural Netw. 20, 1 (2008), 61\u201380.","journal-title":"IEEE Trans. Neural Netw."},{"key":"e_1_3_1_37_2","first-page":"20","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops","author":"Schumann Arne","year":"2017","unstructured":"Arne Schumann and Rainer Stiefelhagen. 2017. Person re-identification by deep learning attribute-complementary information. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 20\u201328."},{"key":"e_1_3_1_38_2","first-page":"486","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201918)","author":"Shen Yantao","year":"2018","unstructured":"Yantao Shen, Hongsheng Li, Shuai Yi, Dapeng Chen, and Xiaogang Wang. 2018. Person re-identification with deep similarity-guided graph neural network. In Proceedings of the European Conference on Computer Vision (ECCV\u201918). 486\u2013504."},{"key":"e_1_3_1_39_2","article-title":"Very deep convolutional networks for large-scale image recognition","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1409.1556.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1409.1556"},{"key":"e_1_3_1_40_2","first-page":"1179","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Song Chunfeng","year":"2018","unstructured":"Chunfeng Song, Yan Huang, Wanli Ouyang, and Liang Wang. 2018. Mask-guided contrastive attention model for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 1179\u20131188."},{"issue":"3","key":"e_1_3_1_41_2","doi-asserted-by":"crossref","first-page":"714","DOI":"10.1109\/72.572108","article-title":"Supervised neural networks for the classification of structures","volume":"8","author":"Sperduti Alessandro","year":"1997","unstructured":"Alessandro Sperduti and Antonina Starita. 1997. Supervised neural networks for the classification of structures. IEEE Trans. Neural Netw. 8, 3 (1997), 714\u2013735.","journal-title":"IEEE Trans. Neural Netw."},{"key":"e_1_3_1_42_2","first-page":"3960","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Su Chi","year":"2017","unstructured":"Chi Su, Jianing Li, Shiliang Zhang, Junliang Xing, Wen Gao, and Qi Tian. 2017. Pose-driven deep convolutional model for person re-identification. In Proceedings of the IEEE International Conference on Computer Vision. 3960\u20133969."},{"key":"e_1_3_1_43_2","first-page":"393","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Sun Yifan","year":"2019","unstructured":"Yifan Sun, Qin Xu, Yali Li, Chi Zhang, Yikang Li, Shengjin Wang, and Jian Sun. 2019. Perceive where to focus: Learning visibility-aware part-level features for partial person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 393\u2013402."},{"key":"e_1_3_1_44_2","first-page":"480","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201918)","author":"Sun Yifan","year":"2018","unstructured":"Yifan Sun, Liang Zheng, Yi Yang, Qi Tian, and Shengjin Wang. 2018. Beyond part models: Person retrieval with refined part pooling (and a strong convolutional baseline). In Proceedings of the European Conference on Computer Vision (ECCV\u201918). 480\u2013496."},{"key":"e_1_3_1_45_2","first-page":"7134","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Tay Chiat-Pin","year":"2019","unstructured":"Chiat-Pin Tay, Sharmili Roy, and Kim-Hui Yap. 2019. Aanet: Attribute attention network for person re-identifications. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 7134\u20137143."},{"key":"e_1_3_1_46_2","doi-asserted-by":"crossref","first-page":"274","DOI":"10.1145\/3240508.3240552","volume-title":"Proceedings of the 26th ACM International Conference on Multimedia","author":"Wang Guanshuo","year":"2018","unstructured":"Guanshuo Wang, Yufeng Yuan, Xiong Chen, Jiwei Li, and Xi Zhou. 2018. Learning discriminative features with multiple granularities for person re-identification. In Proceedings of the 26th ACM International Conference on Multimedia. 274\u2013282."},{"issue":"3","key":"e_1_3_1_47_2","doi-asserted-by":"crossref","first-page":"513","DOI":"10.1109\/TCSVT.2016.2586851","article-title":"Deeplist: Learning deep features with adaptive listwise constraint for person reidentification","volume":"27","author":"Wang Jin","year":"2016","unstructured":"Jin Wang, Zheng Wang, Changxin Gao, Nong Sang, and Rui Huang. 2016. Deeplist: Learning deep features with adaptive listwise constraint for person reidentification. IEEE Trans. Circ. Syst. Video Technol. 27, 3 (2016), 513\u2013524.","journal-title":"IEEE Trans. Circ. Syst. Video Technol."},{"issue":"4","key":"e_1_3_1_48_2","doi-asserted-by":"crossref","first-page":"219","DOI":"10.1049\/trit.2018.1001","article-title":"Survey on person re-identification-based on deep learning","volume":"3","author":"Wang Kejun","year":"2018","unstructured":"Kejun Wang, Haolin Wang, Meichen Liu, Xianglei Xing, and Tian Han. 2018. Survey on person re-identification-based on deep learning. CAAI Trans. Intell. Technol. 3, 4 (2018), 219\u2013227.","journal-title":"CAAI Trans. Intell. Technol."},{"key":"e_1_3_1_49_2","article-title":"Deep graph library: Towards efficient and scalable deep learning on graphs","author":"Wang Minjie","year":"2019","unstructured":"Minjie Wang, Lingfan Yu, Da Zheng, Quan Gan, Yu Gai, Zihao Ye, Mufei Li, Jinjing Zhou, Qi Huang, Chao Ma et\u00a0al. 2019. Deep graph library: Towards efficient and scalable deep learning on graphs. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1909.01315.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1909.01315"},{"key":"e_1_3_1_50_2","unstructured":"Peter Welinder Steve Branson Takeshi Mita Catherine Wah Florian Schroff Serge Belongie and Pietro Perona. 2010. Caltech-UCSD birds 200."},{"key":"e_1_3_1_51_2","doi-asserted-by":"crossref","first-page":"8821","DOI":"10.1109\/TIP.2020.3001693","article-title":"Adaptive graph representation learning for video person re-identification","volume":"29","author":"Wu Yiming","year":"2020","unstructured":"Yiming Wu, Omar El Farouk Bourahla, Xi Li, Fei Wu, Qi Tian, and Xue Zhou. 2020. Adaptive graph representation learning for video person re-identification. IEEE Trans. Image Process. 29 (2020), 8821\u20138830.","journal-title":"IEEE Trans. Image Process."},{"key":"e_1_3_1_52_2","article-title":"A comprehensive survey on graph neural networks","author":"Wu Zonghan","year":"2019","unstructured":"Zonghan Wu, Shirui Pan, Fengwen Chen, Guodong Long, Chengqi Zhang, and Philip S. Yu. 2019. A comprehensive survey on graph neural networks. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1901.00596.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1901.00596"},{"key":"e_1_3_1_53_2","first-page":"2899","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Yan Yichao","year":"2020","unstructured":"Yichao Yan, Jie Qin, Jiaxin Chen, Li Liu, Fan Zhu, Ying Tai, and Ling Shao. 2020. Learning multi-granular hypergraphs for video-based person re-identification. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 2899\u20132908."},{"key":"e_1_3_1_54_2","first-page":"2158","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Yan Yichao","year":"2019","unstructured":"Yichao Yan, Qiang Zhang, Bingbing Ni, Wendong Zhang, Minghao Xu, and Xiaokang Yang. 2019. Learning context graph for person search. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 2158\u20132167."},{"key":"e_1_3_1_55_2","first-page":"3289","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Yang Jinrui","year":"2020","unstructured":"Jinrui Yang, Wei-Shi Zheng, Qize Yang, Ying-Cong Chen, and Qi Tian. 2020. Spatial-temporal graph convolutional network for video-based person re-identification. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 3289\u20133299."},{"key":"e_1_3_1_56_2","first-page":"814","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Yuan Yuhui","year":"2017","unstructured":"Yuhui Yuan, Kuiyuan Yang, and Chao Zhang. 2017. Hard-aware deeply cascaded embedding. In Proceedings of the IEEE International Conference on Computer Vision. 814\u2013823."},{"key":"e_1_3_1_57_2","article-title":"Deep learning on graphs: A survey","author":"Zhang Ziwei","year":"2018","unstructured":"Ziwei Zhang, Peng Cui, and Wenwu Zhu. 2018. Deep learning on graphs: A survey. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1812.04202.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1812.04202"},{"key":"e_1_3_1_58_2","first-page":"3219","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Zhao Liming","year":"2017","unstructured":"Liming Zhao, Xi Li, Yueting Zhuang, and Jingdong Wang. 2017. Deeply-learned part-aligned representations for person re-identification. In Proceedings of the IEEE International Conference on Computer Vision. 3219\u20133228."},{"key":"e_1_3_1_59_2","first-page":"1116","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Zheng Liang","year":"2015","unstructured":"Liang Zheng, Liyue Shen, Lu Tian, Shengjin Wang, Jingdong Wang, and Qi Tian. 2015. Scalable person re-identification: A benchmark. In Proceedings of the IEEE International Conference on Computer Vision. 1116\u20131124."},{"key":"e_1_3_1_60_2","article-title":"Person re-identification: Past, present, and future","author":"Zheng Liang","year":"2016","unstructured":"Liang Zheng, Yi Yang, and Alexander G. Hauptmann. 2016. Person re-identification: Past, present, and future. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1610.02984.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1610.02984"},{"key":"e_1_3_1_61_2","first-page":"5735","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Zheng Meng","year":"2019","unstructured":"Meng Zheng, Srikrishna Karanam, Ziyan Wu, and Richard J. Radke. 2019. Re-identification with consistent attentive siamese networks. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 5735\u20135744."},{"key":"e_1_3_1_62_2","first-page":"3754","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Zheng Zhedong","year":"2017","unstructured":"Zhedong Zheng, Liang Zheng, and Yi Yang. 2017. Unlabeled samples generated by GAN improve the person re-identification baseline in vitro. In Proceedings of the IEEE International Conference on Computer Vision. 3754\u20133762."},{"issue":"10","key":"e_1_3_1_63_2","doi-asserted-by":"crossref","first-page":"3037","DOI":"10.1109\/TCSVT.2018.2873599","article-title":"Pedestrian alignment network for large-scale person re-identification","volume":"29","author":"Zheng Zhedong","year":"2018","unstructured":"Zhedong Zheng, Liang Zheng, and Yi Yang. 2018. Pedestrian alignment network for large-scale person re-identification. IEEE Trans. Circ. Syst. Video Technol. 29, 10 (2018), 3037\u20133045.","journal-title":"IEEE Trans. Circ. Syst. Video Technol."},{"key":"e_1_3_1_64_2","article-title":"Graph neural networks: A review of methods and applications","author":"Zhou Jie","year":"2018","unstructured":"Jie Zhou, Ganqu Cui, Zhengyan Zhang, Cheng Yang, Zhiyuan Liu, Lifeng Wang, Changcheng Li, and Maosong Sun. 2018. Graph neural networks: A review of methods and applications. Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1812.08434.","journal-title":"Retrieved from https:\/\/2.zoppoz.workers.dev:443\/https\/arXiv:1812.08434"},{"key":"e_1_3_1_65_2","first-page":"3702","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Zhou Kaiyang","year":"2019","unstructured":"Kaiyang Zhou, Yongxin Yang, Andrea Cavallaro, and Tao Xiang. 2019. Omni-scale feature learning for person re-identification. In Proceedings of the IEEE International Conference on Computer Vision. 3702\u20133712."},{"key":"e_1_3_1_66_2","first-page":"3741","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Zhou Sanping","year":"2017","unstructured":"Sanping Zhou, Jinjun Wang, Jiayun Wang, Yihong Gong, and Nanning Zheng. 2017. Point to set similarity-based deep feature learning for person re-identification. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3741\u20133750."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/dl.acm.org\/doi\/10.1145\/3487044","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/dl.acm.org\/doi\/pdf\/10.1145\/3487044","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:18:47Z","timestamp":1750191527000},"score":1,"resource":{"primary":{"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/dl.acm.org\/doi\/10.1145\/3487044"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,7,14]]},"references-count":65,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2023,11,30]]}},"alternative-id":["10.1145\/3487044"],"URL":"https:\/\/2.zoppoz.workers.dev:443\/https\/doi.org\/10.1145\/3487044","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,7,14]]},"assertion":[{"value":"2020-10-09","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-02-26","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-07-14","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}