I am a fourth-year Ph.D. candidate in the Gaoling School of Artificial Intelligence at Renmin University of China, advised by Prof. Zhiwu Lu. I was a visiting student in the College of Computing and Data Science at Nanyang Technological University, advised by Prof. Hanwang Zhang. Before my doctoral studies, I received my B.E. from the School of Software at Dalian University of Technology.
My research and industry experience spans large multimodal models, reinforcement learning for reasoning, and multi-task optimization, covering the full R&D pipeline from data construction and filtering to algorithm design, model training, and evaluation. I have published multiple first-author papers at ICLR, WWW, and UAI, with the broader goal of building more capable and versatile multimodal models. In industry, I work extensively on data construction and filtering for vision-language pretraining of Zhipu AI's GLM foundation models. Previously, as a core technical contributor on the startup team at Metabrain AGI, I helped develop ChatImg and the Awaker series of Chinese large multimodal models.