Value-aligned Behavior Cloning for Offline Reinforcement Learning via Bi-level Optimization

Xingyu Jiang, Ning Gao, Xiuhui Zhang, Hongkun Dou, Yue Deng
2/12/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. Any repos shown below are weak matches — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{jiang2026valuealigned,
  title  = {Value-aligned Behavior Cloning for Offline Reinforcement Learning via Bi-level Optimization},
  author = {Xingyu Jiang and Ning Gao and Xiuhui Zhang and Hongkun Dou and Yue Deng},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:f2d8bac0f6f921cd4631377805c04d108bb60a0d},
  journal = {ICLR 2025 2025}
}

Discussion