🔑 Key features:
Learnable policy network for adaptive modulation of token generation
Adversarial reward model for improved quality and diversity
Significantly reduced inference time compared to diffusion models
📊 Impressive results on ImageNet, MSCOCO, and CC3M datasets!