Sitemap
A list of all the posts and pages found on the site. For you robots out there is an XML version available for digesting as well.
Pages
Posts
portfolio
publications
High-resolution representations for labeling pixels and regions
Published in ArXiv Preprint, 2019

Recommended citation: Sun Ke*, Yang Zhao*, Borui Jiang*, Tianheng Cheng*, Bin Xiao, Dong Liu, Yadong Mu, Xinggang Wang, Wenyu Liu, and Jingdong Wang. High-resolution representations for labeling pixels and regions. ArXiv Preprint, 2019.
Download Paper
Deep high-resolution representation learning for visual recognition
Published in IEEE Transactions on Pattern Analysis and Machine Intelligence (IEEE TPAMI), 2020

Recommended citation: Jingdong Wang, Ke Sun, Tianheng Cheng, Borui Jiang, Chaorui Deng, Yang Zhao, Dong Liu, Yadong Mu, Mingkui Tan, Xinggang Wang, Wenyu Liu and Bin Xiao. Deep high-resolution representation learning for visual recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence (IEEE TPAMI), 2020.
Download Paper
MobileFAN: transferring deep hidden representation for face alignment
Published in Pattern Recognition (PR), 2020

Recommended citation: Yang Zhao, Yifan Liu, Chunhua Shen, Yongsheng Gao, and Shengwu Xiong. MobileFAN: transferring deep hidden representation for face alignment. Pattern Recognition (PR), 2020.
Download Paper
Patchy Image Structure Classification Using Multi-Orientation Region Transform
Published in In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), 2020

Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, Shengwu Xiong, and Xiaohui Yuan. Patchy Image Structure Classification Using Multi-Orientation Region Transform. In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), 2020.
Download Paper
Learning Discriminative Region Representation for Person Retrieval
Published in Pattern Recognition (PR), 2021

Recommended citation: Yang Zhao, Xiaohan Yu, Yongsheng Gao and Chunhua Shen. Learning Discriminative Region Representation for Person Retrieval. Pattern Recognition (PR), 2021.
Download Paper
Benchmark Platform for Ultra-Fine-Grained Visual Categorization Beyond Human Performance
Published in In Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021

Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, Xiaohui Yuan and Shengwu Xiong. Benchmark Platform for Ultra-Fine-Grained Visual Categorization Beyond Human Performance. In Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021.
Download Paper
MaskCOV: A Random Mask Covariance Network for Ultra-Fine-Grained Visual Categorization
Published in Pattern Recognition (PR), 2021

Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, and Shengwu Xiong. MaskCOV: A Random Mask Covariance Network for Ultra-Fine-Grained Visual Categorization. Pattern Recognition (PR), 2021.
Download Paper
Learning Deep Part-Aware Embedding for Person Retrieval
Published in Pattern Recognition (PR), 2021

Recommended citation: Yang Zhao, Chunhua Shen, Xiaohan Yu, Hao Chen, Yongsheng Gao, and Shengwu Xiong. Learning Deep Part-Aware Embedding for Person Retrieval. Pattern Recognition (PR), 2021.
Download Paper
Gait-Assisted Video Person Retrieval
Published in IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022
Recommended citation: Yang Zhao, Xinlong Wang, Xiaohan Yu, Chunlei Liu, Yongsheng Gao. Gait-Assisted Video Person Retrieval. IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022.
Download Paper
SPARE: Self-Supervised Part Erasing for Ultra-Fine-Grained Visual Categorization
Published in Pattern Recognition (PR), 2022

Recommended citation: Xiaohan Yu, Yang Zhao, and Yongsheng Gao. SPARE: Self-Supervised Part Erasing for Ultra-Fine-Grained Visual Categorization. Pattern Recognition (PR), 2022.
Download Paper
RB-Net: Training Highly Accurate and Efficient Binary Neural Networks with Reshaped Point-wise Convolution and Balanced Activation
Published in IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022

Recommended citation: Chunlei Liu, Wenrui Ding, Peng Chen, Bohan Zhuang, Yufeng Wang, Yang Zhao, Baochang Zhang, Yuqi Han. RB-Net: Training Highly Accurate and Efficient Binary Neural Networks with Reshaped Point-wise Convolution and Balanced Activation. IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022.
Download Paper
Mix-ViT: Mixing Attentive Vision Transformer for Ultra-Fine-Grained Visual Categorization
Published in Pattern Recognition (PR), 2023
Recommended citation: Xiaohan Yu, Jun Wang, Yang Zhao, and Yongsheng Gao. Mix-ViT: Mixing Attentive Vision Transformer for Ultra-Fine-Grained Visual Categorization. Pattern Recognition (PR), 2023.
Download Paper
Motion Avatar: Generate Human and Animal Avatars with Arbitrary Motion
Published in The 35th British Machine Vision Conference (BMVC), 2024
![]()
Recommended citation: Zeyu Zhang*, Yiran Wang*, Biao Wu*, Shuo Chen, Zhiyuan Zhang, Shiya Huang, Wenbo Zhang, Meng Fang, Ling Chen, Yang Zhao. Motion Avatar: Generate Human and Animal Avatars with Arbitrary Motion. The 35th British Machine Vision Conference (BMVC 2024), Glasgow, UK.
Download Paper
A Landmark-based Approach for Instability Prediction in Distal Radius Fractures
Published in The IEEE International Symposium on Biomedical Imaging (ISBI), 2024

Recommended citation: Yang Zhao, Zhibin Liao, Yunxiang Liu, Koen Oude Nijhuis, Britt Barvelink, Jasper Prijs, Joost Colaris, Mathieu Wijffels, Max Reijman, Zeyu Zhang, Minh-Son To, Ruurd Jaarsma, Job Doornberg, Johan Verjans. A Landmark-based Approach for Instability Prediction in Distal Radius Fractures. The IEEE International Symposium on Biomedical Imaging (IEEE ISBI 2024), Athens, Greece, 2024.
Download Paper
Self-Supervised Lie Algebra Representation Learning via Optimal Canonical Metric
Published in The IEEE Transactions on Neural Networks and Learning Systems (IEEE TNNLS), 2024

Recommended citation: Xiaohan Yu, Zicheng Pan, Yang Zhao, Yongsheng Gao. Self-Supervised Lie Algebra Representation Learning via Optimal Canonical Metric. The IEEE Transactions on Neural Networks and Learning Systems (TNNLS), 2024.
Download Paper
Occluded Person Retrieval with Hierarchical Feature Optimization
Published in 18th IEEE International Conference on Automatic Face and Gesture Recognition (FG), (Oral) (Best reviewed papers), 2024

Recommended citation: Yang Zhao, Pengcheng Zhang, Xiaohan Yu, Zhibin Liao, Johan Verjans, Xiao Bai, Wei Xiang. Occluded Person Retrieval with Hierarchical Feature Optimization. 18th IEEE International Conference on Automatic Face and Gesture Recognition (IEEE FG 2024), Istanbul, Turkey, 2024. (Oral) (Best reviewed papers).
Download Paper
MedDet: Generative Adversarial Distillation for Efficient Cervical Disc Herniation Detection
Published in 2024 IEEE International Conference on Bioinformatics and Biomedicine (BIBM), 2024
<!– Direct-link publication template.
Peddet: adaptive spectral optimization for multimodal pedestrian detection
Published in the 28th European Conference on Artificial Intelligence (ECAI), 2025
<!– Direct-link publication template.
Projectedex: Enhancing generation in explainable ai for prostate cancer
Published in 2025 IEEE 38th International Symposium on Computer-Based Medical Systems (CBMS), 2025
<!– Direct-link publication template.
Msdet: Receptive field enhanced multiscale detection for tiny pulmonary nodule
Published in 2025 IEEE International Conference on Multimedia and Expo (ICME), 2025
<!– Direct-link publication template.
MedConv: convolutions beat transformers on long-tailed bone density prediction
Published in 2025 International Joint Conference on Neural Networks (IJCNN), 2025
<!– Direct-link publication template.
Contrastive Lie Algebra Learning for Ultra-Fine-Grained Visual Categorization
Published in Proceedings of the 33rd ACM International Conference on Multimedia (MM), 2025
<!– Direct-link publication template.
CIT: Rethinking class-incremental semantic segmentation with a Class Independent Transformation
Published in Pattern Recognition (PR), 2025
<!– Direct-link publication template.
Presentagent: Multimodal agent for presentation video generation
Published in Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2025
<!– Direct-link publication template.
Advancing federated domain generalization in ophthalmology: vision enhancement and consistency assurance for multicenter fundus image segmentation
Published in Pattern Recognition (PR), 2026
<!– Direct-link publication template.
Vasevqa: Multimodal agent and benchmark for ancient greek pottery
Published in Findings of the Association for Computational Linguistics (EACL), 2026
<!– Direct-link publication template.
VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery
Published in The Fourteenth International Conference on Learning Representations (ICLR), 2026
<!– Direct-link publication template.
