Abstract:Recent advances in garment-centric image generation from text and image prompts based on diffusion models are impressive. However, existing methods lack support for various combinations of attire, and struggle to preserve the garment details while maintaining faithfulness to the text prompts, limiting their performance across diverse scenarios. In this paper, we focus on a new task, i.e., Multi-Garment Virtual Dressing, and we propose a novel AnyDressing method for customizing characters conditioned on any combination of garments and any personalized text prompts. AnyDressing comprises two primary networks named GarmentsNet and DressingNet, which are respectively dedicated to extracting detailed clothing features and generating customized images. Specifically, we propose an efficient and scalable module called Garment-Specific Feature Extractor in GarmentsNet to individually encode garment textures in parallel. This design prevents garment confusion while ensuring network efficiency. Meanwhile, we design an adaptive Dressing-Attention mechanism and a novel Instance-Level Garment Localization Learning strategy in DressingNet to accurately inject multi-garment features into their corresponding regions. This approach efficiently integrates multi-garment texture cues into generated images and further enhances text-image consistency. Additionally, we introduce a Garment-Enhanced Texture Learning strategy to improve the fine-grained texture details of garments. Thanks to our well-craft design, AnyDressing can serve as a plug-in module to easily integrate with any community control extensions for diffusion models, improving the diversity and controllability of synthesized images. Extensive experiments show that AnyDressing achieves state-of-the-art results.

Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation

Design2GarmentCode: Turning Design Concepts to Tangible Garments Through Program Synthesis

An Integrated Method of 3D Garment Design

FashionSD-X: Multimodal Fashion Garment Synthesis using Latent Diffusion

Design 3D Garments for Scanned Human Bodies

Multi-Garment Customized Model Generation

DressCode: Autoregressively Sewing and Generating Garments from Text Guidance

New Fashion: Personalized 3D Design with a Single Sketch Input

DiffCloth: Diffusion Based Garment Synthesis and Manipulation via Structural Cross-modal Semantic Alignment

Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing

Learning a Shared Shape Space for Multimodal Garment Design

Magic Clothing: Controllable Garment-Driven Image Synthesis

AIpparel: A Large Multimodal Generative Model for Digital Garments

GAN-Based Garment Generation Using Sewing Pattern Images

AnyDressing: Customizable Multi-Garment Virtual Dressing via Latent Diffusion Models

ChatGarment: Garment Estimation, Generation and Editing via Large Language Models

Automatic Digital Garment Initialization from Sewing Patterns

GarmentDreamer: 3DGS Guided Garment Synthesis with Diverse Geometry and Texture Details

Generating Datasets of 3D Garments with Sewing Patterns

A knowledge-supported approach for garment pattern design using fuzzy logic and artificial neural networks