Internships

Research and engineering internships spanning LLM post-training, agents, retrieval-augmented generation, and 3D vision.

to Present

Moonshot AI (月之暗面)

Intern RL Team

Kimi Coding Models

Contributed to the post-training of Kimi coding models across data construction, reinforcement learning, experimentation, evaluation, and model tuning, with a focus on complex software engineering and long-horizon agent tasks.

  • Data synthesis: Designed synthetic training data for long-tail code tasks, multi-turn PRD/SOP scenarios, and code instruction following.
  • RL training: Extended internal RL training infrastructure and ran training, tuning, and evaluation experiments.
  • Reinforcement Learning
  • Agentic Coding
  • Data Synthesis
to

Ant Group

Machine Learning Algorithm Intern 百灵&数字科技

Agentic LLM

Worked on an agentic LLM training pipeline spanning synthetic trajectory generation, data validation, and training and evaluation environment engineering.

  • Trajectory synthesis: Designed task-synthesis workflows for multi-step tool use and decision-making trajectories.
  • Data validation: Built a filtering pipeline for sample cleaning, trajectory deduplication, and rule-based validation to support versioned dataset iterations.
  • Environment engineering: Deployed and validated agent training and evaluation environments, including dependency management, execution debugging, and regression checks.
  • Agentic Training
  • Tool Use
  • Data Curation
to

Institute of Artificial Intelligence, Hefei Comprehensive National Science Center

Machine Game Intelligence Group

Large Language Models

Built a government-service question-answering system using Qwen and retrieval-augmented generation, with support for multi-turn dialogue, contextual understanding, and traceable answers.

  • 12345 question answering: Integrated Qwen with a RAG pipeline for the public-service hotline knowledge domain.
  • Qwen
  • RAG
  • Multi-turn Dialogue
to

Tsinghua Shenzhen International Graduate School

IVG@SZ

3D Human Reconstruction

Reimplemented the ACTOR and ROMP 3D human reconstruction projects in MindSpore, covering core model components, training, and evaluation.

  • Model reimplementation: Delivered MindSpore implementations of two established 3D human reconstruction methods.
  • Open-source contribution: The resulting projects were made available through MindSpore's official community.
  • MindSpore
  • 3D Human Reconstruction
  • Model Reimplementation