Moonshot AI (月之暗面)
Intern RL Team
Kimi Coding Models
Contributed to the post-training of Kimi coding models across data construction, reinforcement learning, experimentation, evaluation, and model tuning, with a focus on complex software engineering and long-horizon agent tasks.
- Data synthesis: Designed synthetic training data for long-tail code tasks, multi-turn PRD/SOP scenarios, and code instruction following.
- RL training: Extended internal RL training infrastructure and ran training, tuning, and evaluation experiments.
