Publications

Peer-reviewed papers and preprints on language model reasoning, reinforcement learning, and knowledge-enhanced systems.

2026

  1. Preprint
    DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes
    Caijun Xu, Changyi Xiao, Zhongyuan Peng, and Yixin Cao
    arXiv preprint arXiv:2605.28421, May 2026
  2. Preprint
    Reinforcement Learning with Conditional Expectation Reward
    Changyi Xiao, Caijun Xu, and Yixin Cao
    arXiv preprint arXiv:2603.10624, Mar 2026
  3. Preprint
    CoDiQ: Test-Time Scaling for Controllable Difficult Question Generation
    Zhongyuan Peng, Caijun Xu, Changyi Xiao, Shibo Hong, Eli Zhang, Stephen Huang, and Yixin Cao
    arXiv preprint arXiv:2602.01660, Feb 2026
  4. ACL’26 Findings
    SCALER: Synthetic Scalable Adaptive Learning Environment for Reasoning
    Caijun Xu, Changyi Xiao, Zhongyuan Peng, Xinrun Wang, and Yixin Cao
    In Findings of the Association for Computational Linguistics: ACL 2026, Jan 2026

2024

  1. CIKM’24
    Exploring High-Order User Preference with Knowledge Graph for Recommendation
    Caijun Xu, Fuwei Zhang, Zhao Zhang, Fuzhen Zhuang, and Rui Liu
    In Proceedings of the 33rd ACM International Conference on Information and Knowledge Management, Jan 2024