- Post-Training of Modern LLMs
Modern LLM post-training, from preference learning to reinforcement learning with verifiable rewards.
6 min - One-step Generation in the Post Diffusion Era
Training-based routes toward one-step generation for diffusion and flow models, covering distillation, Consistency Models, CTM, MeanFlow, DMD, and Drifting Models.
12 min - Understanding Diffusion Models in Two Perspectives
DDPM and score-based SDEs as two routes to reverse diffusion, contrasting their shared structure and modeling differences.
5 min
Back