arXiv:2607. 23189v1 Announce Type: cross Abstract: AI-generated content (AIGC) has made significant progress, with 2D generative models becoming ready-to-use tools for the digital fashion industry.
By Shenghao Yang, Hongtao Zhang, Yuhan Yi, Zhihao Tang, Zihao Cui, Lian Wen, Han Yan, Yuan Gao, Mingbo Zhao
arXiv:2608.23302v1 Announce Type: new
Abstract: Fashion complementary image generation (CIG) aims to create garments that stylistically match a seed item based on user intent, making it a natural mul...
By Matteo Attimonelli, Claudio Pomo, Alessandro De Bellis, Danilo Danese, Dietmar Jannach, Tommaso Di Noia
MMTryon is a multi‑modal, multi‑reference virtual try‑on framework that generates high‑quality compositional try‑on results using text instructions and multiple garment images. It addresses three overlooked problems: supporting multiple try‑on items, allowing dressing style specification via text, and eliminating reliance on segmentation models by using a parsing‑free garment encoder and a scalable data generation pipeline. Experiments on high‑resolution benchmarks and in‑the‑wild test sets show MMTryon outperforms state‑of‑the‑art methods qualitatively and quantitatively.
By Xujie Zhang, Ente Lin, Michael Kampffmeyer, Zhenyu Xie, Jiang Li, Ting Liu, Xiaochao Qu, Luoqi Liu, Xiaodan Liang
FitControler introduces a fit-aware virtual try‑on system that adds garment fit control to existing VTON models. It uses a fit‑aware layout generator and a multi‑scale fit injector to redraw body‑garment layouts and render garments that match those layouts. The authors also release a new Fit4Men dataset of 13,000 body‑garment pairs and two fit consistency metrics to evaluate fit quality.
By Lu Yang, Yicheng Liu, Letian Zhou, Yanan Li, Xiang Bai, Hao Lu
Fashion complementary image generation (CIG) aims to create garments that stylistically match a seed item based on user intent, making it a natural multimodal grounding problem where models must inter...
arXiv:2511. 09483v3 Announce Type: replace Abstract: While multimodal large language models can describe visual content, their ability to generate executable procedures remains underexplored.
By Peiyu Li, Xiaobao Huang, Ting Hua, Nitesh V. Chawla
arXiv:2602.24043v2 Announce Type: replace
Abstract: Reconstructing 3D clothed humans from monocular images and videos is a fundamental problem with applications in virtual try-on, avatar creation, an...
By Yingxuan You, Ren Li, Corentin Dumery, Cong Cao, Hao Li, Pascal Fua
arXiv:2511. 18765v3 Announce Type: replace-cross Abstract: Existing industrial 3D garment meshes already cover most real-world clothing geometries, yet their texture diversity remains limited.
By Hui Shan, Ming Li, Haitao Yang, Kai Zheng, Sizhe Zheng, Yanwei Fu, Xiangru Huang
The paper introduces HyperBones, a real‑time garment simulation framework that combines a reduced‑space neural dynamics simulator with a lightweight neural network correcting Linear Blend Skinning (LBS) at a coarse level, and a convolutional MLP for fine‑scale wrinkle recovery in UV space. By decoupling identity‑specific computation from shape conditioning through a hypernetwork, the method achieves high performance without an offline simulator, delivering physically plausible dynamics across diverse motions and unseen body shapes. Experiments demonstrate a speedup of over 30× compared to state‑of‑the‑art autoregressive neural simulators, reaching interactive inference at roughly 1 ms per frame on a consumer GPU.
By Astitva Srivastava, Hsiao-Yu Chen, Ryan Goldade, Philipp Herholz, Zhongshi Jiang, Gene Wei-Chin Lin, Lingchen Yang, Nikolaos Sarafianos, Tuur Stuyck, Avinash Sharma, Egor Larionov
arXiv:2607. 09362v1 Announce Type: cross Abstract: Virtual try-on (VTO) has made significant progress in realistically transferring garments onto a target person.
By Seungyong Lee, Hyun Jun Jang, Sangoh Kim, Sungjoon Park
arXiv:2501. 13692v2 Announce Type: replace-cross Abstract: Diffusion models have recently unlocked new possibilities in editing images of real-world objects.
By Potito Aghilar, Vito Walter Anelli, Michelantonio Trizio, Eugenio Di Sciascio, Tommaso Di Noia
arXiv:2608. 05745v1 Announce Type: cross Abstract: Video Virtual Try-On (VVT) synthesizes a video of a person wearing a target garment while preserving identity, motion, and scene dynamics.
By Yushe Cao, Shikun Feng, Fei Shen, Haikuo Peng, Jianqiang Xia, Yiheng Zhu, Dianxi Shi, Chun Yu