RL post-training offers a steerable alternative to CFG for image generation
Original titleA new way to steer image generation models (using RL instead of CFG).
AISummary
Researchers introduced a simple, sample-efficient online reinforcement learning technique for post-training image generation models. It is presented as a possible steerable alternative to classifier-free guidance (CFG) that can be driven by any scalar reward, including human preference.
Source: NVIDIA AI Developer · x.comPublished · added here