Research Engineer on the Scaling Team at
Tencent HY, working on
model architecture. I received my Ph.D. in Computer Science
and Technology from Tsinghua University.
// 01 about
I am a Research Engineer on the Scaling Team at
Tencent HY, where I work on model
architecture.
I received my Ph.D. in Computer Science and Technology from
Tsinghua University in June 2026.
I was fortunate to be advised by Prof.
Shaoping Ma, Prof.
Yiqun Liu, and Prof.
Qingyao Ai. My research
focuses on efficient, scalable model architectures and the systems that
make them practical.
// 02 research
01
Model Architecture
Co-designing algorithms and infrastructure for capable models that remain efficient at scale.
02
Efficient Attention
Sparse attention, linear attention, and architecture-level techniques for reducing inference cost.
03
Scalable Context
Extending useful context while balancing model quality, memory, throughput, and latency.
04
Low-Precision Training
Developing stable low-precision methods that reduce training cost while preserving model quality at scale.
// 03 news
2026.08
Hy4 preview was released. As one of its model-architecture contributors, I helped bring our research IndexCache into production in Hy4, enabling a cost-efficient 1M-token context.
2026.06
I received my Ph.D. in Computer Science and Technology from Tsinghua University.
2026.03
IndexCache was released, accelerating sparse attention via cross-layer index reuse.
2026.02
The GLM-5 Technical Report was released. I am one of the core contributors to model architecture.