💬 About Me I am a researcher at StepFun, working on post-training and long-horizon tasks. My research interests include reinforcement learning, post-training, and optimization.