Back to projects

Diffusion / Post-training

Self-Evolving Diffusion

Qwen-VL and on-policy self-distillation for self-evolving diffusion models.

2026 Research project In progress

An ongoing project exploring how diffusion models can improve through their own generated experience. The current direction combines Qwen-VL with on-policy self-distillation to study self-evolving training loops for multimodal generation and reasoning.

The project is still in progress; more details and release materials will be added as the experiments mature.