I am a first-year Ph.D. student at
The Chinese University of Hong Kong, Shenzhen,
supervised by Prof.
Benyou Wang.
My research focuses on exploring trustworthy medical large language models (LLMs) and multimodal large models (MLLMs), as well as exploring multimodal generative for healthcare applications.
I like collaborating with
Claude Code
and
CodeX
. They are my good partners! I am a perfectionist. I want to solve the problems in my hands in perfect ways, which is, however, not achievable in many situations.
Jan. 2026 One paper accepted by ICLR 2026, congrats to all co-authors! See you at Brazil!
May 2025 Two papers accepted by ACL 2025, congrats to all co-authors!
Apr. 2025 Our team won the gold medal in the AIMO-2 competition, ranking 14th out of 2213!
Publications
† Equal contribution. * Corresponding author. The leading papers are highlighted.
If you would like to view my other publications, you are welcome to access them on my Google Scholar.
RareLens: Towards End-to-End Rare Disease Care via Aligning Divergent Large Language Model Reasoning Xi Chen,
Hongru Zhou,
Shiyu Feng,
Hanyu Zhou,
Huahui Yi,
Rongsheng Wang,
Tiancheng He,
Kun Wang,
Pingping Liu,
Qiankun Li,
Sicheng Lin,
Huiying Ou,
Xiaohong Zheng,
Tianying Zang,
Zhuohang Wu,
Leheng Jiang,
Kexin Cao,
Wenhan Zhang,
ChengYi Li,
Zhiyang Wang,
Songlin Li,
Benyou Wang,
Ningbei Yin,
Shaoting Zhang,
Weili Fu*,
Jian Li*,
Kang Li* Homepage
/
Human Evaluation
/
Code Under Review
RareLens provides trustworthy, actionable decision support across the full rare-disease care pathway, from early risk screening and diagnosis to treatment, efficacy evaluation, and prognosis.
GameCraft-Bench: Can Agents Build Playable Games End-to-End in a Real Game Engine? Tongxu Luo†,
Rongsheng Wang†,
Jiaxi Bi†,
Chenming Xu†,
Zhengyang Tang†,
Jianlong Chen,
Juhao Liang,
Ke Ji,
Shuqi Guo,
Yuhao Du,
Fan Bu,
Wenyu Du,
Xiaotong Zhang,
Kyle Li,
Shaobo Wang,
Linfeng Zhang,
Yuxuan Liu,
Xin Lai,
Chenxin Li,
Yiduo Guo,
Zhexin Zhang,
Xinyuan Wang,
Tianyi Bai,
Ziniu Li,
Benyou Wang* Homepage / Code / Paper Preprint
GameCraft-Bench evaluates whether agents can build playable games end-to-end in a real game engine.