I am currently a Research Fellow in Learning and Vision Laboratory (LV-Lab) at the National University of Singapore (NUS), working with Prof. Shuicheng Yan. I received my Ph.D. degree from Fudan University (FDU) in 2026, where I was affiliated with the Institute of Trustworthy Embodied AI (复旦大学可信具身智能研究院), under the supervision of Prof. Yu-Gang Jiang and Prof. Zuxuan Wu. To date, I have published dozens of papers in top-tier international venues.

Since 2019, my research has focused on AIGC and Embodied AI, with particular interests in controllable diffusion models, world models, and embodied intelligence. My long-term vision is to develop world models that enable embodied agents to understand, reason about, and purposefully interact with the physical world.

Please feel free to contact me at zhaohaoyu [at] nus.edu.sg if you are interested in internship or research collaboration opportunities.

🔥 News

📝 Selected Publications

  • Haoyu Zhao, Zihao Zhang, et al. CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation. ACM MM, 2026.

  • Zihao Zhang*, Haoyu Zhao*, et al. SPEED: One-Step Pixel Diffusion for High-quality Video Frame Interpolation. ACM MM, 2026. (* Equal contribution)

  • Haoyu Zhao, Zuxuan Wu, Yu-Gang Jiang. UniCam: Taming Unified Diffusion Models in Noise Space for Camera-controllable Video Rendering. IJCV, 2026.

  • Haoyu Zhao, Zhongang Qi, et al. DynamiCtrl: Rethinking the Basic Structure and the Role of Text for High-quality Human Image Animation. TMM, 2026. (Media Coverage: CVer)

  • Haoyu Zhao, Jiaxi Gu, et al. CameraNoise: Enabling Faithful Camera Control in Video Diffusion through Geometry-Flow-Guided Noise Warping. ICML, 2026. (Media Coverage: 腾讯广告技术)

  • Haoyu Zhao, Yuang Zhang, et al. DCDM: Divide-and-conquer diffusion models for consistency-preserving video generation. 🏆 AAAI’26 CVM Challenging Top Team, 2026. (Media Coverage: 智猩猩AI)

  • Qingping Zheng, Bo Huang, Yang Liu, Haoyu Zhao, et al. ReFocusEraser: Refocusing for small object removal with robust context-shadow repair. ICLR, 2026.

  • Haoyu Zhao, Jiaxi Gu, et al. LSTD: Long short-term temporal diffusion for video generation. TMM, 2025. (Media Coverage: 机器之心)

  • Zihao Zhang, Haoran Chen, Haoyu Zhao, et al. Eden: Enhanced diffusion for high-quality large-motion video frame interpolation. CVPR, 2025.

  • Cong Wang, Panwen Hu, Haoyu Zhao, et al. Uniadapter: All-in-one control for flexible video generation. TCSVT, 2025.

  • Haoyu Zhao, Tianyi Lu, et al. MagDiff: Multi-alignment diffusion for high-fidelity video generation and editing. ECCV, 2024.

  • Haoyu Zhao, Weidong Min, et al. Scene-adaptive crowd counting method based on meta learning with dual-input network DMNet. Frontiers of Computer Science (FCS), 2023.

  • Haoyu Zhao, Weidong Min, et al. Memory-efficient document layout analysis method using LD-net. Multimedia Tools and Applications (MTAP), 2023.

  • Qi Wang, Yuling Zhong, Weidong Min, Haoyu Zhao, et al. Dual similarity pre-training and domain difference encouragement learning for vehicle re-identification in the wild. PR, 2023.

  • Haoyu Zhao, Qi Wang, et al. Need Only One More Point (NOOMP): Perspective adaptation crowd counting in complex scenes. TMM, 2022.

  • Zitai Wei, Weidong Min, Qi Wang, Qian Liu, Haoyu Zhao. ECNFP: Edge-constrained network using a feature pyramid for image inpainting. Expert Systems with Applications (ESWA), 2022.

  • Haoyu Zhao, Weidong Min, et al. SPACE: Finding key-speaker in complex multi-person scenes. IEEE Transactions on Emerging Topics in Computing (TETC), 2021.

  • Haoyu Zhao, Weidong Min, et al. MSR-FAN: Multi-scale residual feature-aware network for crowd counting. IET IP, 2021.

  • Haoyu Zhao, Weidong Min, et al. Illumination-enhanced crowd counting based on IC-Net in low lighting conditions. ICIG, 2021.

  • Qi Wang, Weidong Min, Qing Han, Qian Liu, Cheng Zha, Haoyu Zhao, et al. Inter-domain adaptation label for data augmentation in vehicle re-identification. TMM, 2021.

🎖 Honors and Awards

  • AAAI’26 CVM Top Team (Role: PI), 2026

  • Outstanding Student Award, Fudan University, 2024

  • Outstanding Student Leader, Fudan University, 2024

  • Academic Scholarship (First Class), Fudan University, awarded multiple times, 2022-2026

  • National Scholarship, Ministry of Education, China (Top 1%), 2022

💼 Work Experience

  • 2026.07 - Present, Research Fellow, LV-Lab, National University of Singapore (NUS).

  • 2025.05 - 2026.06, Research Intern (Qingyun Talent Program), Advertising Marketing Services (AMS), Tencent. Conducted research on digital human and contributed to their deployment in Tencent Ads Miaosi.

  • 2023.03 - 2025.03, Research Intern, Huawei Noah’s Ark Lab, Shanghai. Worked on AIGC research and its deployment in Huawei Mate-series phones.

📋 Academic Services

Reviewer for IJCV, TIP, TMM, KBS, TCSVT, NeurIPS, ICLR, ICML, CVPR, AAAI, ACM MM, ACL, WACV etc.


Updated on July, 2026.

Visitor Map

Live visitor tracking.