What Is RLCD? Reinforcement Learning from Contrast Distillation
A plain-language explainer on RLCD, a way of aligning language models by learning from contrasting outputs rather than human ratings alone.
From the site: RLCD is a method developed to adjust language models to human preferences without using human feedback data. This approach aims to address…
Why I recommend it: Good background reading if you want to understand how the models you use are actually steered.
