Hi! I’m Kaixuan Ji, a third-year Ph.D. student in Computer Science at UCLA, fortunately advised by Professor Quanquan Gu. Before coming to UCLA, I completed my undergraduate studies in Department of Computer Science and Technology at Tsinghua University, where I was fortunate to work with Professors Jie Tang and Juanzi Li. My current research explores reinforcement learning theory and its role in training large language models. Here is my latest CV.
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
Kaixuan Ji*, Qiwei Di*, Heyang Zhao, Qingyue Zhao, Quanquan Gu, NeurIPS 2026
Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability
Qingyue Zhao*, Kaixuan Ji*, Heyang Zhao, Quanquan Gu, NeurIPS 2026, Spotlight Presentation