Staff Researcher

Zhengkai Jiang

Tencent Hunyuan · Multi-modal Generative AI

About

My research spans image and video generation, diffusion models, multi-modal learning, and visual representation learning.

At Tencent, I focus on post-training the Hunyuan multi-modal foundation models — aligning large-scale generative models with human preference.

Research

Reward Modeling & RL for Vision Generation — aligning text-to-image and text-to-video diffusion models with human preference. I build reward models that capture what people actually value in a generated image or video, then post-train generative policies against them with GRPO, NFT, OPD and ReFL.

Publications

13 papers · 2019–2025

* equal contribution. My name in bold.

Honors

2020
1st of the Objects365 Challenge
2019
1st of the NuScenes 3D Detection, CVPR WAD Workshop
2017
Outstanding Graduates of Liaoning Province
2015–2016
1st of the Chinese Mathematics Competitions (1/6000+ in Liaoning)
2014–2016
National Scholarship (Top 0.2% students in China)