Weighted-Reward Preference Optimization for Implicit Model Fusion

Ziyi Yang, Fanqi Wan, Longguang Zhong, Tianyuan Shi, Xiaojun Quan
2/12/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 2 weaker matches are hidden by default — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{yang2026weightedreward,
  title  = {Weighted-Reward Preference Optimization for Implicit Model Fusion},
  author = {Ziyi Yang and Fanqi Wan and Longguang Zhong and Tianyuan Shi and Xiaojun Quan},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:fd70a52a9db002f22794b60fb29e132c38c8bad0},
  journal = {ICLR 2025 2025}
}

Discussion