Tailoring Self-Rationalizers with Multi-Reward Distillation

Sahana Ramnath, Brihi Joshi, Skyler Hallinan, Ximing Lu, Liunian Harold Li +4 more
2/12/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 1 weaker match is hidden by default — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{ramnath2026tailoring,
  title  = {Tailoring Self-Rationalizers with Multi-Reward Distillation},
  author = {Sahana Ramnath and Brihi Joshi and Skyler Hallinan and Ximing Lu and Liunian Harold Li and Aaron Chan and Jack Hessel and Yejin Choi and Xiang Ren},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:da3657410a97cde3ebaea4d0378d80e182367f22},
  journal = {ICLR 2024 2024}
}

Discussion