About
My research spans image and video generation, diffusion models, multi-modal learning, and visual representation learning.
At Tencent, I focus on post-training the Hunyuan multi-modal foundation models — aligning large-scale generative models with human preference.
Research
Reward Modeling & RL for Vision Generation — aligning text-to-image and text-to-video diffusion models with human preference. I build reward models that capture what people actually value in a generated image or video, then post-train generative policies against them with GRPO, NFT, OPD and ReFL.
Publications
13 papers · 2019–2025* equal contribution. My name in bold.
Honors
- 2020
- 1st of the Objects365 Challenge
- 2019
- 1st of the NuScenes 3D Detection, CVPR WAD Workshop
- 2017
- Outstanding Graduates of Liaoning Province
- 2015–2016
- 1st of the Chinese Mathematics Competitions (1/6000+ in Liaoning)
- 2014–2016
- National Scholarship (Top 0.2% students in China)