Sitemap
A list of all the posts and pages found on the site. For you robots out there is an XML version available for digesting as well.
Pages
Posts
portfolio
publications
High-resolution representations for labeling pixels and regions
Published in ArXiv Preprint, 2019, 2019

Recommended citation: Sun Ke*, Yang Zhao*, Borui Jiang*, Tianheng Cheng*, Bin Xiao, Dong Liu, Yadong Mu, Xinggang Wang, Wenyu Liu, and Jingdong Wang. High-resolution representations for labeling pixels and regions. ArXiv Preprint, 2019.
Download Paper
Deep high-resolution representation learning for visual recognition
Published in IEEE Transactions on Pattern Analysis and Machine Intelligence (IEEE TPAMI), 2020, 2020

Recommended citation: Jingdong Wang, Ke Sun, Tianheng Cheng, Borui Jiang, Chaorui Deng, Yang Zhao, Dong Liu, Yadong Mu, Mingkui Tan, Xinggang Wang, Wenyu Liu and Bin Xiao. Deep high-resolution representation learning for visual recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence (IEEE TPAMI), 2020.
Download Paper
MobileFAN: transferring deep hidden representation for face alignment
Published in Pattern Recognition (PR), 2020, 2020

Recommended citation: Yang Zhao, Yifan Liu, Chunhua Shen, Yongsheng Gao, and Shengwu Xiong. MobileFAN: transferring deep hidden representation for face alignment. Pattern Recognition (PR), 2020.
Download Paper
Patchy Image Structure Classification Using Multi-Orientation Region Transform
Published in In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), 2020, 2020

Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, Shengwu Xiong, and Xiaohui Yuan. Patchy Image Structure Classification Using Multi-Orientation Region Transform. In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), 2020.
Download Paper
EAR-NET: Error Attention Refining Network For Retinal Vessel Segmentation
Published in Digital Image Computing: Techniques and Applications (DICTA), 2021, 2021

Recommended citation: Jun Wang, Yang Zhao, Linglong Qian, Xiaohan Yu and Yongsheng Gao. EAR-NET: Error Attention Refining Network For Retinal Vessel Segmentation. Digital Image Computing: Techniques and Applications (DICTA), 2021.
Download Paper
Learning Discriminative Region Representation for Person Retrieval
Published in Pattern Recognition (PR), 2021, 2021

Recommended citation: Yang Zhao, Xiaohan Yu, Yongsheng Gao and Chunhua Shen. Learning Discriminative Region Representation for Person Retrieval. Pattern Recognition (PR), 2021.
Download Paper
Benchmark Platform for Ultra-Fine-Grained Visual Categorization Beyond Human Performance
Published in In Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021, 2021

Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, Xiaohui Yuan and Shengwu Xiong. Benchmark Platform for Ultra-Fine-Grained Visual Categorization Beyond Human Performance. In Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021.
Download Paper
MaskCOV: A Random Mask Covariance Network for Ultra-Fine-Grained Visual Categorization
Published in Pattern Recognition (PR), 2021, 2021

Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, and Shengwu Xiong. MaskCOV: A Random Mask Covariance Network for Ultra-Fine-Grained Visual Categorization. Pattern Recognition (PR), 2021.
Download Paper
Learning Deep Part-Aware Embedding for Person Retrieval
Published in Pattern Recognition (PR), 2021, 2021

Recommended citation: Yang Zhao, Chunhua Shen, Xiaohan Yu, Hao Chen, Yongsheng Gao, and Shengwu Xiong. Learning Deep Part-Aware Embedding for Person Retrieval. Pattern Recognition (PR), 2021.
Download Paper
Learning Deep Asymmetric Tolerant Part Representation
Published in IEEE Transactions on Artificial Intelligence (IEEE TAI), 2022, 2022
Recommended citation: Xiaohan Yu, Yang Zhao, Yongsheng Gao, Shengwu Xiong. Learning Deep Asymmetric Tolerant Part Representation. IEEE Transactions on Artificial Intelligence (IEEE TAI), 2022.
Download Paper
Gait-Assisted Video Person Retrieval
Published in IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022, 2022
Recommended citation: Yang Zhao, Xinlong Wang, Xiaohan Yu, Chunlei Liu, Yongsheng Gao. Gait-Assisted Video Person Retrieval. IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022.
Download Paper
SPARE: Self-Supervised Part Erasing for Ultra-Fine-Grained Visual Categorization
Published in Pattern Recognition (PR), 2022, 2022

Recommended citation: Xiaohan Yu, Yang Zhao, and Yongsheng Gao. SPARE: Self-Supervised Part Erasing for Ultra-Fine-Grained Visual Categorization. Pattern Recognition (PR), 2022.
Download Paper
RB-Net: Training Highly Accurate and Efficient Binary Neural Networks with Reshaped Point-wise Convolution and Balanced Activation
Published in IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022, 2022

Recommended citation: Chunlei Liu, Wenrui Ding, Peng Chen, Bohan Zhuang, Yufeng Wang, Yang Zhao, Baochang Zhang, Yuqi Han. RB-Net: Training Highly Accurate and Efficient Binary Neural Networks with Reshaped Point-wise Convolution and Balanced Activation. IEEE Transactions on Circuits and Systems for Video Technology (IEEE TCSVT), 2022.
Download Paper
Normal periocular anthropometric measurements in an Australian population
Published in International Ophthalmology, 2023, 2023

Recommended citation: Khizar Rana, Mark B. Beecher, Carmelo Caltabiano, Yang Zhao, Johan Verjans, Dinesh Selva. Normal periocular anthropometric measurements in an Australian population. International Ophthalmology, 2023.
Download Paper
Mix-ViT: Mixing Attentive Vision Transformer for Ultra-Fine-Grained Visual Categorization
Published in Pattern Recognition (PR), 2023, 2023
Recommended citation: Xiaohan Yu, Jun Wang, Yang Zhao, and Yongsheng Gao. Mix-ViT: Mixing Attentive Vision Transformer for Ultra-Fine-Grained Visual Categorization. Pattern Recognition (PR), 2023.
Download Paper
Motion Avatar: Generate Human and Animal Avatars with Arbitrary Motion
Published in The 35th British Machine Vision Conference (BMVC 2024), Glasgow, UK, 2024
![]()
Recommended citation: Zeyu Zhang*, Yiran Wang*, Biao Wu*, Shuo Chen, Zhiyuan Zhang, Shiya Huang, Wenbo Zhang, Meng Fang, Ling Chen, Yang Zhao. Motion Avatar: Generate Human and Animal Avatars with Arbitrary Motion. The 35th British Machine Vision Conference (BMVC 2024), Glasgow, UK.
Download Paper
A Landmark-based Approach for Instability Prediction in Distal Radius Fractures
Published in The IEEE International Symposium on Biomedical Imaging (IEEE ISBI 2024), Athens, Greece, 2024, 2024

Recommended citation: Yang Zhao, Zhibin Liao, Yunxiang Liu, Koen Oude Nijhuis, Britt Barvelink, Jasper Prijs, Joost Colaris, Mathieu Wijffels, Max Reijman, Zeyu Zhang, Minh-Son To, Ruurd Jaarsma, Job Doornberg, Johan Verjans. A Landmark-based Approach for Instability Prediction in Distal Radius Fractures. The IEEE International Symposium on Biomedical Imaging (IEEE ISBI 2024), Athens, Greece, 2024.
Download Paper
Facial and Mandibular Landmark Tracking with Habitual Head Posture Estimation using Linear and Fiducial Markers
Published in Healthcare Technology Letters, 2024, 2024
![]()
Recommended citation: Farhan Saad, Taseef Farook, Saif Ahmed, Yang Zhao, Zhibin Liao, Johan Verjans, James Dudley. Facial and Mandibular Landmark Tracking with Habitual Head Posture Estimation using Linear and Fiducial Markers. Healthcare Technology Letters, 2024.
Download Paper
Self-Supervised Lie Algebra Representation Learning via Optimal Canonical Metric
Published in The IEEE Transactions on Neural Networks and Learning Systems (TNNLS), 2024, 2024

Recommended citation: Xiaohan Yu, Zicheng Pan, Yang Zhao, Yongsheng Gao. Self-Supervised Lie Algebra Representation Learning via Optimal Canonical Metric. The IEEE Transactions on Neural Networks and Learning Systems (TNNLS), 2024.
Download Paper
Occluded Person Retrieval with Hierarchical Feature Optimization
Published in 18th IEEE International Conference on Automatic Face and Gesture Recognition (IEEE FG 2024), Istanbul, Turkey, 2024. (Oral) (Best reviewed papers), 2024

Recommended citation: Yang Zhao, Pengcheng Zhang, Xiaohan Yu, Zhibin Liao, Johan Verjans, Xiao Bai, Wei Xiang. Occluded Person Retrieval with Hierarchical Feature Optimization. 18th IEEE International Conference on Automatic Face and Gesture Recognition (IEEE FG 2024), Istanbul, Turkey, 2024. (Oral) (Best reviewed papers).
Download Paper
Detection, classification, and characterization of proximal humerus fractures on plain radiographs
Published in The bone & joint journal, 2024
<!– Direct-link publication template.
Convolutional neural networks can accurately detect but not classify proximal humerus fractures
Published in Navigating new waters in proximal humerus fractures, 2024
<!– Direct-link publication template.
MedDet: Generative Adversarial Distillation for Efficient Cervical Disc Herniation Detection
Published in 2024 IEEE International Conference on Bioinformatics and Biomedicine (BIBM), 2024
<!– Direct-link publication template.
Artificial intelligence to automate assessment of ocular and periocular measurements
Published in European Journal of Ophthalmology, 2025
<!– Direct-link publication template.
Peddet: adaptive spectral optimization for multimodal pedestrian detection
Published in the 28th European Conference on Artificial Intelligence. (ECAI2025), 2025
<!– Direct-link publication template.
Sss: Semi-supervised sam-2 with efficient prompting for medical imaging segmentation
Published in Biomedical Signal Processing and Control, 2025
<!– Direct-link publication template.
Projectedex: Enhancing generation in explainable ai for prostate cancer
Published in 2025 IEEE 38th International Symposium on Computer-Based Medical Systems (CBMS), 2025
<!– Direct-link publication template.
Msdet: Receptive field enhanced multiscale detection for tiny pulmonary nodule
Published in 2025 IEEE International Conference on Multimedia and Expo (ICME), 2025
<!– Direct-link publication template.
MedConv: convolutions beat transformers on long-tailed bone density prediction
Published in 2025 International Joint Conference on Neural Networks (IJCNN), 2025
<!– Direct-link publication template.
Contrastive Lie Algebra Learning for Ultra-Fine-Grained Visual Categorization
Published in Proceedings of the 33rd ACM International Conference on Multimedia (MM 2025), 2025
<!– Direct-link publication template.
CIT: Rethinking class-incremental semantic segmentation with a Class Independent Transformation
Published in Pattern Recognition, 2025
<!– Direct-link publication template.
Presentagent: Multimodal agent for presentation video generation
Published in Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, 2025
<!– Direct-link publication template.
Medical artificial intelligence for early detection of lung cancer: A survey
Published in Engineering Applications of Artificial Intelligence, 2025
<!– Direct-link publication template.
Open-source convolutional neural network to classify distal radial fractures according to the AO/OTA classification on plain radiographs
Published in European Journal of Trauma and Emergency Surgery, 2025
<!– Direct-link publication template.
An open source convolutional neural network to detect and localize distal radius fractures on plain radiographs
Published in European Journal of Trauma and Emergency Surgery, 2025
<!– Direct-link publication template.
Advancing federated domain generalization in ophthalmology: vision enhancement and consistency assurance for multicenter fundus image segmentation
Published in Pattern Recognition, 2026
<!– Direct-link publication template.
Vasevqa: Multimodal agent and benchmark for ancient greek pottery
Published in Findings of the Association for Computational Linguistics: EACL 2026, 2026
<!– Direct-link publication template.
VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery
Published in The Fourteenth International Conference on Learning Representations (ICLR 2026), 2026
<!– Direct-link publication template.
Dynamic domain adaptation-driven physics-informed graph representation learning for ac-opf
Published in Sustainable Energy, Grids and Networks, 2026
<!– Direct-link publication template.
Doei: Dual optimization of embedding information for attention-enhanced class activation maps
Published in Neurocomputing, 2026
<!– Direct-link publication template.
