I'm an undergraduate student at Fudan University, expecting to graduate in 2028.
My research interests lie in vision-language models (e.g. video understanding, grounding) and generative models (e.g. t2v t2i, world model) with special focus on agentic methods for them. I also read about mechinterp. and representation learning.
Beyond research, I enjoy films and indie music.

