I am a Ph.D. student in the School of Computer Science at Northwestern Polytechnical University, Xi’an, China, supervised by Prof. Lei Zhang at ASGO. My research focuses on computer vision and generative AI, particularly diffusion models and synthetic data for visual recognition.
I received my M.S. in Computer Technology from Northwestern Polytechnical University in 2024, advised by Prof. Lei Zhang, and my B.S. in Computer Science and Technology from Yanshan University in 2021.
I am open to research discussions and collaborations. Please feel free to contact me by email.
🔥 News
- 2026.01: 🎉 SRA was accepted to ICLR 2026.
- 2025.11: 🎉 JoDiffusion was accepted to AAAI 2026.
- 2025.04: 🎉 PFCD was accepted to IJCAI 2025.
- 2025.02: 🎉 lbGen was accepted to CVPR 2025.
📝 Selected Publications
* Equal contribution. See Google Scholar for the full publication list.
JoDiffusion: Jointly Diffusing Image with Pixel-Level Annotations for Semantic Segmentation Promotion
Haoyu Wang, Lei Zhang, Wenrui Liu, Dengyang Jiang, Wei Wei, Chen Ding
- A joint diffusion framework for semantic segmentation dataset generation using an annotation VAE for latent alignment and a boundary mode-based mask optimization strategy.
No Other Representation Component Is Needed: Diffusion Transformers Can Provide Representation Guidance by Themselves
Dengyang Jiang, Mengmeng Wang, Liuzhuozheng Li, Lei Zhang, Haoyu Wang, Wei Wei, Guang Dai, Yanning Zhang, Jingdong Wang
- Enhances Diffusion Transformers’ representation and generation through self-representation alignment via self-distillation, eliminating external components.
Prompt-Free Conditional Diffusion for Multi-object Image Augmentation
Haoyu Wang, Lei Zhang, Wei Wei, Chen Ding, Yanning Zhang
- A framework for multi-object image augmentation using local-global semantic fusion and a reward model-based counting loss.
Low-Biased General Annotated Dataset Generation
Dengyang Jiang*, Haoyu Wang*, Lei Zhang, Wei Wei, Guang Dai, Mengmeng Wang, Jingdong Wang, Yanning Zhang
- A framework generating low-biased annotated datasets using a fine-tuned diffusion model with bi-level semantic alignment and quality assurance for enhanced backbone generalization.
Adapt Anything: Tailor Any Image Classifier across Domains And Categories Using Text-to-Image Diffusion Models
Weijie Chen*, Haoyu Wang*, Shicai Yang, Lei Zhang, Wei Wei, Yanning Zhang, Luojun Lin, Di Xie, Yueting Zhuang
- Uses text-to-image diffusion models to create synthetic data, enabling image classifier adaptation across domains and categories without real-world source data.
Glocal energy-based learning for few-shot open-set recognition
Haoyu Wang*, Guansong Pang*, Peng Wang*, Lei Zhang, Wei Wei, Yanning Zhang
- A novel energy-based model for few-shot open-set recognition using global and local features.
📖 Education & Experience
- 2024.03 – present: Northwestern Polytechnical University, Ph.D. student in Computer Science and Technology.
- 2023.06 – 2023.09: Hikvision Research Institute, Research Intern.
- 2021.09 – 2024.03: Northwestern Polytechnical University, M.S. in Computer Technology.
- 2017.09 – 2021.06: Yanshan University, B.S. in Computer Science and Technology.
💻 Academic Services
- Conference reviewer: ICCV 2025, NeurIPS 2025, AAAI 2026.
- Journal reviewer: TBC, PR, JSTARS.