PT-T2I/V: An Efficient Proxy-Tokenized Diffusion Transformer for Text-to-Image/Video-Task

Jing Wang, Ao Ma, Jiasong Feng, Dawei Leng, Yuhui Yin +1 more
2/12/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 1 weaker match is hidden by default — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{wang2026pttiv,
  title  = {PT-T2I/V: An Efficient Proxy-Tokenized Diffusion Transformer for Text-to-Image/Video-Task},
  author = {Jing Wang and Ao Ma and Jiasong Feng and Dawei Leng and Yuhui Yin and Xiaodan Liang},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:e03bd875d5a23a940b9bde8419f3b26848629a60},
  journal = {ICLR 2025 2025}
}

Discussion