ParamΔ for Direct Mixing: Post-Train Large Language Model At Zero Cost
Sheng Cao, Mingrui Wu, Karthik Prasad, Yuandong Tian, Zechun Liu
2/12/2026
No confident code match yet
We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 2 weaker matches are hidden by default — verify before relying on them.
No code implementations found yet.
Know of an implementation? Let us know in the comments below!
@article{cao2026param,
title = {ParamΔ for Direct Mixing: Post-Train Large Language Model At Zero Cost},
author = {Sheng Cao and Mingrui Wu and Karthik Prasad and Yuandong Tian and Zechun Liu},
year = {2026},
url = {https://api.semanticscholar.org/CorpusID:f67cac37888b1c5448733ed52504b17cc809cf78},
journal = {ICLR 2025 2025}
}