RLCD: Reinforcement Learning from Contrastive Distillation for LM Alignment

Kevin Yang, Daniel Klein, Asli Celikyilmaz, Nanyun Peng, Yuandong Tian
2/12/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 1 weaker match is hidden by default — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{yang2026rlcd,
  title  = {RLCD: Reinforcement Learning from Contrastive Distillation for LM Alignment},
  author = {Kevin Yang and Daniel Klein and Asli Celikyilmaz and Nanyun Peng and Yuandong Tian},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:f4b5ed96ca25c0f61778a26d03cdd6c4b946b5ea},
  journal = {ICLR 2024 2024}
}

Discussion