I am a Principal Researcher at Tencent TEG / Hunyuan Foundation Model Department, where I lead R2V and video editing for the video foundation model. My work also covers video omni-modal understanding post-training and Agentic RL. Previously, I worked at Alibaba / Alimama, NetEase Youdao, and SenseTime. My research interests span video generation, multimodal understanding, AIGC, and multi-agent systems.
I received my M.S. in Computer Software and Theory from the Institute of Computing Technology, Chinese Academy of Sciences (ICT, CAS), advised by Prof. Jianfeng Zhan, and my B.Eng. in Software Engineering from Nankai University.
M.S. in Computer Software and Theory
B.Eng. in Software Engineering