ICLR 2026 Oral
News
-
Oct 2026
KV-cache compression is not a free lunch: when n tokens are compressed into one cache entry, models spontaneously learn to keep only a single content-agnostic slot—inducing what we call phase sensitivity. We study this theoretically and empirically.
-
Jan 2026
Paper “ERC loss for MoEs”, an auxiliary loss for MoE autonomy, accepted to ICLR 2026 Oral
-
Oct 2025
Awarded National Scholarship for Doctoral Students Ranked 1st in GSAI
-
Sept 2025
Paper “PolarQuant”, effective post-RoPE KV-cache quantization, accepted to NeurIPS 2025.
-
July 2025
Joined ByteDance Seed Top Seed Intern
-
July 2025
Paper “HoPE”, on why partial RoPE works, received the ACL 2025 SAC Highlights Award
-
May 2025
Paper “Autonomy-of-Experts Models”, self-selecting MoE experts, accepted to ICML 2025.
Honors and Awards
-
2025
National Scholarship for Doctoral Students1st-Ranked in GSAI
-
2025
ByteDance Top Seed Intern
-
2025
ACL 2025 SAC Highlights Paper AwardTop 1.5%
-
2025
CIE-Tencent Doctoral Student Research Incentive ProgramHunYuan Large Language Model Special Project · 1 of 17 selected individuals nationwide
-
2024
CCF-Tencent Rhino-Bird Elite Talent Program1 of 50 selected individuals nationwide
-
2023 – 2025
Outstanding Innovative Talents Cultivation Funded ProgramsRenmin University of China
Academic Services
-
Area Chair
EMNLP, ACLACL Rolling Review (ARR)
-
Reviewer
ICML Gold, ICLR, NeurIPS
Internships
- 2025.07 – Now
-
2024.05 – 2025.07
Tencent Hunyuan, mentored by Ruobing Xie. We conducted a series of work on autonomous MoE experts.
-
2023.09 – 2024.05
Alibaba, Tongyi Lab.
-
2023.03 – 2023.09
Microsoft Research, Machine Learning Area, mentored by Xu Tan. I am deeply grateful to Xu Tan for his patient and rigorous mentorship, which laid the foundation for my growth as a researcher. Our collaborative efforts on the Muzic project boast 5k stars on GitHub.
Recent Publications
NeurIPS 2025
ICML 2025