Biography

  • I am a Senior Researcher at Tencent Hunyuan. I earned my M.S. degree from Peking University in 2020 and B.S. degree from HFUT.
  • At Hunyuan, my research mainly focuses on CUA/Code Agent, RL, and Synthetic Data.
  • At Microsoft, I co-founded WizardLM project, which contributed the SOTA LLMs WizardLM, WizardCoder and WizardMath, I also created widely adopted methods Evol-Instruct, RLEIF and Arena-Learning.
  • We are hiring research interns! If you have strong experiences in LLMs and are willing to work in Hunyuan, feel free to shoot me your resume.

News

  • [Jun 2026] Inroduce two RL papers STARE for Policy Entropy Stability and VeriEvol for Multimodal Reasoning!
  • [Mar 2026] Our two RL papers RubricBench and Beyond Length Scaling are featured on Huggingface Daily Paper #4 and Huggingface Daily Paper #6 respectively!
  • [Jan 2026] AgentMath accepted by ICLR 2026!
  • [Aug 2025] We release Hunyuan-Large-Vision which ranks #5 globally on LMArena-Vision and is the #1 VLM of China.
  • [May 2025] We release Hunyuan-TurboS which ranks #7 globally on LMArena-Text and is the #2 LLM of China.
  • [May 2025] AgentGen accepted by KDD 2025!
  • [May 2025] WarriorCoder accepted by ACL 2025!
  • [Sep 2024] WizardArena accepted by NeurIPS 2024!
  • [Oct 2024] Wizard models achieves 3M+ HF downloads, 9K+ Github stars and ranks #4 globally (#1 opensource) on LMSYS Arena!
  • [Jun 2024] We release WizardLM-2, which ranks #3 in monthly token usage on OpenRouter, just behind Gemini 1.5 Pro and Llama 3.
  • [Aug 2023] We release WizardMath. Accepted by ICLR 2025 as an Oral paper!
  • [Jun 2023] We release WizardCoder. Accepted by ICLR 2024!
  • [Apr 2023] We release WizardLM. Accepted by ICLR 2024! Project link: WizardLM .
  • 2 papers accepted by ACL 2023!
  • 1 paper accepted by EMNLP 2022!
  • 1 paper accepted by NAACL 2022 as an Oral paper!
  • 2 papers accepted by ACL 2022!
  • 1 paper accepted by EMNLP 2019!

Selected Publications [Google Scholar]

(*: Equal contribution, #: The intern I mentored)

Agentic Model:

RLVR & RLHF:

General LLM & VLM:

Open-source Projects

(showing only those I’m the project lead):

  • WizardLM LLMs Family: WizardLM, WizardCoder, WizardMath
  • STARE Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability
  • Llama-X Open Academic Research on Improving LLaMA to SOTA LLM
  • MMDialog A Large-scale Multi-turn Dataset Towards VLM
  • PromDA Prompt-based Data Augmentation for Low-Resource NLU Tasks

Experiences

  • Dec. 2024 - Now, Senior Researcher, Tencent Hunyuan X.
  • July. 2020 - Dec. 2024, Research Scientist, Microsoft AI.
  • Sept. 2018 - June. 2020, Research Intern, Microsoft XiaoIce.

Academic Services

Program Committee for

  • ICLR 2025, 2026
  • NeurIPS 2024
  • NAACL 2024
  • ACL 2023, 2026
  • EACL 2023
  • KDD 2022, 2023
  • EMNLP 2022
  • COLING 2022