I am currently a Senior Research Scientist at ByteDance/TikTok, working on video synthesis and generation.
Previously, I was a Postdoc Fellow supervised by Professor
Hao (Richard) Zhang in
GrUVi Lab at Simon Fraser University (SFU), Canada.
During my postdoc, I worked on 3D shape reconstruction and content creation.
I received my PhD degree from Peking University, where I worked on Computer Vision and Computer Graphics.
During my PhD, I focused on applying deep generative models to analyze and synthesize 2D geometric data
such as glyphs, fonts, and layouts. I was supervised by Prof.
Zhouhui Lian and Prof.
Jianguo Xiao.
News
Oct 2026SplitMoE is accepted by NeurIPS 2026 as a spotlight paper!
A one-line projection that filters critic error from distribution-matching updates, stabilizing few-step video and joint audio-video diffusion distillation without extra losses or networks.
An adversarial formulation of distribution matching that learns log-density ratios with two discriminator heads for fast image, video, and joint audio-video generation.