πŸ’¬ About Me

  • I am a researcher on the Foundation Team at StepFun, working on post-training for the Step model family.
  • My research interests include reinforcement learning, post-training, and optimization for foundation models.
  • We are hiring researchers and interns across pre-training, post-training, multimodal, and speech, with opportunities in algorithms, infrastructure, and data. Feel free to send your resume and research interests to my email for a referral.

πŸ“– Educations

  • 2022.09 - 2026.06: B.Eng, Computer Science (Turing Honors Class), Renmin University of China, advisor Mingyang Yi.

πŸ’» Experiences

  • 2026.06 - now: Researcher, LLM Post-training Team, StepFun, Beijing.
  • 2026.01 - 2026.06: Research Intern, AI Infra and Data Team, JD.com, Beijing.
  • 2025.03 - 2025.09: Research Intern, Reinforcement Learning Team, Zhongguancun Academy, Beijing.
  • 2024.09 - 2025.03: Development Intern, LLM Pre-training Team, PixVerse, Beijing.

πŸ“ Publications

PrePrints

πŸŽ– Selected Honors and Awards

  • 2026.06: Special Award, Sa Shixuan Elite Fund, Renmin University of China (30,000 CNY)
  • 2026.05: Research Funding, Anonymous Personal Sponsorship (5,000 USDT)
  • 2022.11: Gold Medal, The 2022 ICPC Asia Shenyang Regional Contest