Cross-Modal Contextualized Diffusion Models for Text-Guided Visual Generation and Editing
Ling Yang, Zhilong Zhang, Zhaochen Yu, Jingwei Liu, Minkai Xu +2 more
2/12/2026
No confident code match yet
We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 4 weaker matches are hidden by default — verify before relying on them.
No code implementations found yet.
Know of an implementation? Let us know in the comments below!
@article{yang2026crossmodal,
title = {Cross-Modal Contextualized Diffusion Models for Text-Guided Visual Generation and Editing},
author = {Ling Yang and Zhilong Zhang and Zhaochen Yu and Jingwei Liu and Minkai Xu and Stefano Ermon and Bin Cui},
year = {2026},
doi = {10.48550/arXiv.2402.16627},
url = {https://doi.org/10.48550/arXiv.2402.16627},
journal = {ICLR 2024 2024}
}