ComposeMe
收藏资源简介:
ComposeMe 数据集是用于可控人体图像生成的新框架,特别针对身份属性的组合。该数据集包含各种姿势和表情的主题,用于促进自然组合和鲁棒的解耦。数据集创建过程包括使用专用分词器处理每个视觉组件的参考图像,并将这些属性特定的标记注入到预训练的文本到图像扩散模型中。ComposeMe 的目标是实现细粒度的可控图像合成,允许用户通过指定不同的身份属性来合成图像,如面部身份、发型和服装。此外,该方法还扩展到多人生成,即使在单个图像中也能组合多个不同的身份。数据集的应用领域包括自由形式的虚拟试穿系统,以及设计人员探索不同角色之间视觉特征的创造性工具。
The ComposeMe dataset is a novel framework for controllable human image generation, specifically targeting the composition of identity attributes. This dataset includes subjects with diverse poses and expressions, aimed at facilitating natural composition and robust disentanglement. The dataset creation process involves using a specialized tokenizer to process reference images of each visual component, and injecting these attribute-specific tokens into pre-trained text-to-image diffusion models. The goal of ComposeMe is to achieve fine-grained controllable image synthesis, allowing users to generate images by specifying different identity attributes such as facial identity, hairstyle, and clothing. Furthermore, this method can be extended to multi-person generation, enabling the combination of multiple distinct identities even within a single image. The application areas of this dataset include free-form virtual try-on systems, as well as creative tools for designers to explore visual features across different characters.
ComposeMe 数据集概述
基本信息
- 数据集名称:ComposeMe
- 发布机构:Snap Inc., USA
- 相关会议:SIGGRAPH Asia 2025
核心功能
- 支持对人类图像进行可控生成。
- 实现对多个视觉属性的解耦控制,如身份、发型和服装。
- 支持基于文本的控制。
技术方法
- 采用属性特定标记化技术,分别对身份、发型和服装进行表示。
- 使用多属性交叉参考训练策略。
- 基于预训练扩散模型进行嵌入合并和注入。
训练策略
第一阶段:单参考复制粘贴训练
- 学习每个属性的外观特征。
第二阶段:多属性交叉参考训练
- 将每个身份分解为不同的视觉属性。
- 从不同的输入图像中获取每个属性。
- 预测单独的目标图像。
- 能够从不对齐的属性输入生成自然对齐、连贯的输出。
实验内容
- 全身单人个性化
- 多属性单人个性化
- 仅面部双人个性化
- 多属性双人个性化
消融研究
- 面部和头发的交叉参考训练可有效控制表情和头部姿势。
- 服装的交叉参考训练可减轻来自服装区域的姿势泄漏。
- 多属性交叉参考训练使ComposeMe能够从不对齐的属性特定视觉提示实现高保真生成。
引用信息
bibtex @inproceedings{qian2025composeme, author = {Guocheng Gordon Qian and Daniil Ostashev and Egor Nemchinov and Avihay Assouline and Sergey Tulyakov and Kuan-Chieh Jackson Wang and Kfir Aberman}, title = {ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation}, booktitle = {ACM SIGGRAPH Asia 2025 Conference Proceedings}, year = {2025}, }

- 1ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image GenerationSnap Inc., USA · 2025年



