Value-aligned Behavior Cloning for Offline Reinforcement Learning via Bi-level Optimization
Xingyu Jiang, Ning Gao, Xiuhui Zhang, Hongkun Dou, Yue Deng
2/12/2026
No confident code match yet
We couldn't find an author-owned or strongly-evidenced community implementation for this paper. Any repos shown below are weak matches — verify before relying on them.
No code implementations found yet.
Know of an implementation? Let us know in the comments below!
@article{jiang2026valuealigned,
title = {Value-aligned Behavior Cloning for Offline Reinforcement Learning via Bi-level Optimization},
author = {Xingyu Jiang and Ning Gao and Xiuhui Zhang and Hongkun Dou and Yue Deng},
year = {2026},
url = {https://api.semanticscholar.org/CorpusID:f2d8bac0f6f921cd4631377805c04d108bb60a0d},
journal = {ICLR 2025 2025}
}