I am a fourth-year Ph.D. student at Duke University advised by Prof. Pan Xu. My primary research interests focus on reinforcement learning (RL), especially reinforcement learning with verifiable rewards (RLVR) for improving the reasoning capabilities of large language models. I also work on off-dynamics RL and RL applications in healthcare. More broadly, I am interested in incorporating RL techniques into foundation models to improve their reasoning, adaptation, and alignment with human preferences.
Powered by Jekyll and Minimal Light theme.