Zhonghao Yan

Zhonghao Yan

Master Student

Beijing University of Posts and Telecommunications

ScholarGitHub

Research Interests

Embodied AI
Multimodal Large Language Model
AI for X

About

I am currently a master student at Beijing University of Posts and Telecommunications, advised by Prof. Kongming Liang and Prof. Zhanyu Ma at the PRIS-CV Lab. I earned my Bachelor's degree in a joint program between Beijing University of Posts and Telecommunications and Queen Mary University of London in 2024.

My current research focuses on Embodied AI. I am currently seeking PhD opportunities for Fall 2027. If you are interested in working with me, please feel free to email me :)

News

2026-09

LaST-R1 accepted to NeurIPS 2026 (CCF-A)! 🎉

2026-07

One paper accepted to EMNLP 2026 (CCF-B)! 🎉

2026-07

LingT2I accepted to ACMMM 2026 (CCF-A)! 🎉

2026-06

Hepto-LLaVA accepted to MICCAI 2026 (CCF-B)! 🎉

2026-01

Joined Simplexity Robotics as a Research Intern, led by Peng Jia 🧑‍💻

2026-01

Joined Peking University as a Research Assistant, supervised by Shanghang Zhang 🧑‍💻

2026-01

Joined Shanghai AI Lab as a Remote Research Assistant, supervised by Weijia Li and Conghui He 🧑‍💻

2025-12

Released Step-GUI! 🎉

2025-11

MedReasoner accepted to AAAI 2026 (CCF-A)! 🎉

2025-09

Joined StepFun as a Research Intern, led by Zheng Ge and Xiangyu Zhang 🧑‍💻

2025-06

TRIG accepted to ICCV 2025 (CCF-A)! 🎉

2025-01

PGP-SAM accepted to ISBI 2025! 🎉

2024-09

Joined BUPT AI, PRIS-CV as a Master Student, advised by Kongming Liang and Zhanyu Ma 👨‍🎓

2024-06

Selected as an outstanding graduate of Beijing! 🎉

2024-06

One paper accepted to MICCAI 2024 (CCF-B)! 🎉

Research Trajectory

Multimodal evaluation, reasoning and post-training for MLLMs, and new foundation model architectures and post-training techniques for Embodied AI.

First / Co-first / Project LeadNon-first Author
01

Evaluation

Developing benchmarks for AIGC and MLLMs.

02

Multimodal Large Language Models

Exploring reasoning abilities and post-training methods for MLLMs.

03

Embodied AI

Exploring new foundation model architectures and post-training techniques for Embodied AI.

Selected Publications

View All →

*: Equal Contribution, †: Project Lead, ‡: Corresponding Author

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning

Hao Chen*, Jiaming Liu*†, Zhonghao Yan*, Nuowei Han*, Renrui Zhang†, Chenyang Gu, Jialin Gao, Ziyu Guo, Siyuan Qian, Yinxi Wang, Peng Jia, Chi-Wing Fu, Shanghang Zhang‡, Pheng-Ann Heng

NeurIPS2026 (CCF-A)ProjectarXivPaper
MedReasoner: Reinforcement Learning Drives Reasoning Grounding from Clinical Thought to Pixel-Level Precision

MedReasoner: Reinforcement Learning Drives Reasoning Grounding from Clinical Thought to Pixel-Level Precision

Zhonghao Yan*, Muxi Diao*, Yuxuan Yang, Ruoyan Jing, Jiayuan Xu, Kaizhou Zhang, Lele Yang, Yanxi Liu, Kongming Liang‡, Zhanyu Ma

AAAI2026 (CCF-A)ProjectarXivPaper
Trade-offs in image generation: How do different dimensions interact?

Trade-offs in image generation: How do different dimensions interact?

Sicheng Zhang*, Binzhu Xie*, Zhonghao Yan*, Yuli Zhang, Donghao Zhou, Xiaofei Chen, Shi Qiu, Jiaqi Liu, Guoyang Xie‡, Zhichao Lu

ICCV2025 (CCF-A)arXivPaper