Selected Publication
|
💪 Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning
Dylan Zhang, Yufeng Xu, Haojin Wang, Qingzhi Chen, Hao Peng
Proceedings of ICML, 2026
arxiv
|
📊 Distribution Prompting: Understanding the Expressivity of Language Models Through the Next-Token Distributions They Can Produce
Haojin Wang, Zining Zhu, Freda Shi
Proceedings of EMNLP, 2025
paper,
arxiv
|
Misc
|
|
I am a huge fan of Arsenal. COYG!
|
|