Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads

Siqi Kou, Jiachun Jin, Zhihong Liu, Chang Liu, Ye Ma +4 more
2/14/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 1 weaker match is hidden by default — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{kou2026orthus,
  title  = {Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads},
  author = {Siqi Kou and Jiachun Jin and Zhihong Liu and Chang Liu and Ye Ma and Jian Jia and Quan Chen and Peng Jiang and Zhijie Deng},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:d9cf91264c90b38cc6d9a19ab43dcfb1f2dc5690},
  journal = {ICML 2025 2025}
}

Discussion