I am a Ph.D. student at The Hong Kong Polytechnic University, advised by Prof. Hongxia Yang. Previously, I received my MSc in Artificial Intelligence from Dalian University of Technology, advised by Prof. Huchuan Lu.
My research focuses on how generative models learn, reason, and act efficiently. I study language model training and inference, with a particular interest in diffusion language models and parallel decoding. I also work on multimodal agents for GUI interaction and autonomous driving, connecting visual understanding with reasoning and action generation. My earlier work explored controllable video generation and the evaluation of vision-language models in challenging driving scenes.