About me
I am a Ph.D. student in Artificial Intelligence at Zhejiang University, where I am a member of REAL Lab, advised by Dr. Yongliang Shen and Dr. Jian Shao. I expect to graduate around 2027. Before that, I received my B.S. from Beijing Normal University.
My research centers on large language model (LLM) reasoning and AI agents, with a focus on how to make reasoning models think more effectively and efficiently. My work spans long-horizon and infinite-horizon reasoning, reinforcement learning for reasoning, reward modeling and verification, and agents that act reliably in real environments. To date I have published 15 CCF-A papers, including 8 as first or co-first author.
Alongside my Ph.D., I have worked as an intern on foundation-model teams in industry, including Meituan (LongCat), Ant Group (Ling / Ring), and currently the Tencent Hunyuan VLM Post-Training Team, on large-scale pre-training, post-training, and reinforcement learning for reasoning models.
You can find my work on Google Scholar, Semantic Scholar, and OpenReview. Feel free to reach out at yanyuchen@zju.edu.cn if you would like to discuss research or potential collaborations.
Research Interests
- LLM Reasoning: long-horizon and infinite-horizon reasoning, efficient and adaptive reasoning, chain-of-thought tuning
- Reinforcement Learning for LLMs: reward modeling, verification, policy optimization
- AI Agents: tool use, GUI and mobile agents, agents grounded in environments
News
- 2026.06 InftyThink+ and Milestone-Guided Policy Learning accepted to ICML 2026.
- 2026.04 Joined the Tencent Hunyuan VLM Post-Training Team as an intern.
- 2026.01 InftyThink, MathFimer, VerifyBench, and SpatialLadder accepted to ICLR 2026.
- 2025.12 Test-Time RL for GUI Grounding accepted to AAAI 2026.
- 2025.09 Mind the Gap and Self-Braking Tuning accepted to NeurIPS 2025 (both co-first author).
