I’m Zhuoran Zhao, a third-year PhD student in Computational Media and Arts at HKUST(GZ), supervised by Prof. Pan Hui and Prof. Anyi Rao. I am currently a visiting student at MMLab@HKUST. My research interests are mainly in Video Generation / Editing and 3D Computer Vision. Before that, I received my master’s degree in Computer Science (AI Specialization) at National University of Singapore (NUS), supervised by Prof. Angela Yao and Dr. Linlin Yang. I received my B.Eng degree in Software Engineering from South China University of Technology (SCUT), working with Prof. Junying Chen.

Recent News

  • Sep 2026: Our paper Mask Forcing is now available on arXiv!
  • Sep 2026: Released SolarWM — open data and scalable training for long-horizon video world models.
  • May 2026: One paper is accepted to ICML 2026.
  • Feb 2026: One paper is accepted to CVPR 2026.
  • Jan 2026: Two papers are accepted to ICLR 2026. See you in Rio de Janeiro!
  • June 2025: I will attend the Experimental Model Auditing via Controllable Synthesis workshop and Rhobin workshop at CVPR as “invited poster”!
  • Feb 2025: One paper is accepted to CVPR 2025. See you in Nashville!
  • Sep 2024: One paper is accepted to SIGGRAPH Asia 2024 Educator’s Forum. See you in Tokyo!
  • Jul 2024: One paper is accepted to SIGGRAPH Asia 2024.
  • Jul 2023: One paper is accepted to GCPR 2023.
  • June 2023: One paper is accepted to AAAI 2023 Summer Symposium - AI x Metaverse, with Best Paper Award.

Selected Papers

Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout

Zhuoran Zhao, Shengju Qian, Tongtong Liang, Xianghao Kong, Songchun Zhang, Junchao Huang, Guian Fang, Xin Wang, Pan Hui, Anyi Rao

Preprint, 2026.

[Paper][Project Page] [Code]

SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models

Junchao Huang, Guian Fang, Shengju Qian, Xianghao Kong, Zhuoran Zhao, Wei Huang, Yihua Du, Zixin Zhang, Justin Cui, Yuchao Gu, Yukang Chen, Xinting Hu, Tianyu He, Shaoshuai Shi, Zhuotao Tian, Xin Wang, Mike Zheng Shou, Li Jiang

Technical Report, 2026.

[Paper][Project Page] [Code] [Data]

Threshold-Guided Optimization for Visual Generative Models

Jinbin Bai, Yu Lei, Qingyu Shi, Aosong Feng, Yi Xin, Zhuoran Zhao, Fei Shen, Kaidong Yu, Xiangtai Li

ICML, 2026.

[Paper]

Composing Concepts from Images and Videos via Concept-prompt Binding

Xianghao Kong, Zeyu Zhang, Yuwei Guo, Zhuoran Zhao, Songchun Zhang, Anyi Rao

CVPR, 2026. (highlight)

[Paper] [Project Page] [Code]

SesaHand: Enhancing 3D Hand Reconstruction via Controllable Generation with Semantic and Structural Alignment

Zhuoran Zhao, Xianghao Kong, Linlin Yang, Zheng Wei, Pan Hui, Anyi Rao

ICLR, 2026.

[Openreview] [Project Page]

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model

Qingyu Shi, Jinbin Bai, Zhuoran Zhao, Wenhao Chai, Kaidong Yu, Jianzong Wu, Shuangyong Song, Yunhai Tong, Xiangtai Li, Xuelong Li, Shuicheng Yan

ICLR, 2026.

[Paper] [Code]

Analyzing the Synthetic-to-Real Domain Gap in 3D Hand Pose Estimation

Zhuoran Zhao, Linlin Yang, Pengzhan Sun, Pan Hui, Angela Yao

CVPR, 2025.

[Paper] [Code]

HiFiHR Image

InfNeRF: Towards Infinite Scale NeRF Rendering with O(log n) Space Complexity

Jiabin Liang, Lanqing Zhang, Zhuoran Zhao, Xiangyu Xu

SIGGRAPH Asia, 2024.

[Paper] [Project Page] [Code]

HiFiHR Image

HiFiHR: Enhancing 3D Hand Reconstruction from a Single Image via High-Fidelity Texture

Jiayin Zhu, Zhuoran Zhao, Linlin Yang, Angela Yao

German Conference on Pattern Recognition (GCPR), 2023.

[Paper] [Code]

HiFiHR Image

Taming Diffusion Models for Music-driven Conducting Motion Generation

Zhuoran Zhao*, Jinbin Bai*, Delong Chen, Debang Wang, Yubo Pan

AAAI Symposia Proceedings, 2023.

[Paper] [Code]


Experience

HiFiHR Image

Sea AI Lab

Research Engineer Intern

Topic: Large-scale 3D Scene Reconstruction

Aug 2023 - Dec 2023


HiFiHR Image

Tencent

Frontend Developer

July 2021 - July 2022

Academic Service

  • Reviewer for CVPR, AAAI, ECCV, Siggraph Asia, Pacific Graphics