AIpparel: A Multimodal Foundation Model for Digital Garments

Kiyohiro Nakayama, Jan Ackermann, Timur Levent Kesdogan, Yang Zheng, Maria Korosteleva +4 more
2/10/2026

Abstract

Apparel is essential to human life, offering protection, mirroring cultural identities, and showcasing personal style. Yet, the creation of garments remains a time-consuming process, largely due to the manual work involved in designing them. To simplify this process, we introduce AIpparel, a multimodal foundation model for generating and editing sewing patterns. Our model fine-tunes state-of-the-art large multimodal models (LMMs) on a custom-curated large-scale dataset of over 120,000 unique garments, each with multimodal annotations including text, images, and sewing patterns. Additionally, we propose a novel tokenization scheme that concisely encodes these complex sewing patterns so that LLMs can learn to predict them efficiently. AIpparel achieves state-of-the-art performance in single-modal tasks, including text-to-garment and image-to-garment prediction, and enables novel multimodal garment generation applications such as interactive garment editing. The project website is at https: //georgenakayama.github.io/AIpparel/.

DOISemantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. Any repos shown below are weak matches — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{nakayama2026aipparel,
  title  = {AIpparel: A Multimodal Foundation Model for Digital Garments},
  author = {Kiyohiro Nakayama and Jan Ackermann and Timur Levent Kesdogan and Yang Zheng and Maria Korosteleva and Olga Sorkine-Hornung and Leonidas J. Guibas and Guandao Yang and Gordon Wetzstein},
  year   = {2026},
  doi    = {10.1109/CVPR52734.2025.00762},
  url    = {https://doi.org/10.1109/CVPR52734.2025.00762},
  journal = {CVPR 2025 2025}
}

Discussion