Mitsua/vroid-image-dataset-lite
收藏资源简介:
VRoid Image Dataset Lite是一个用于训练文本到图像模型的数据集,所有材料均为CC0或适当许可,确保无版权问题。数据集包含随机设置的相机角度、姿势、肤色和面部表情等参数的图像输出。数据集还用于训练Mitsua Diffusion One模型,该模型是一个潜在的文本到图像扩散模型,其VAE和U-Net仅使用公共领域/CC0或获得使用许可的版权图像进行训练。数据集使用了多种VRoid模型、姿势和动作、着色器以及其他纹理,所有材料均符合CC0或适当许可。元数据描述包括颜色偏移、VRoid模型名称、姿势剪辑编号、相机配置文件、面部表情、光照、光照颜色、轮廓、卡通阴影、皮肤配置文件、注视标签、相机位置、相机旋转、相机视野、头发颜色偏移、眼睛颜色偏移、服装和配饰颜色偏移、地平面材料、左手手势、右手手势和天空盒等信息。完整数据集包含约60万张图像,仅限非商业研究目的提供。
VRoid Image Dataset Lite is a dataset designed for training text-to-image models. All materials in the dataset are licensed under CC0 or appropriate licenses, ensuring no copyright issues. The dataset includes image outputs with randomly configured parameters such as camera angles, poses, skin tones, and facial expressions. It is also used to train the Mitsua Diffusion One model, a latent text-to-image diffusion model whose VAE and U-Net are trained solely using public domain, CC0-licensed, or properly authorized copyrighted images. The dataset employs various VRoid models, poses and motions, shaders, and other textures, all of which are licensed under CC0 or appropriate licenses. The metadata descriptions cover color offset, VRoid model name, pose clip number, camera profile, facial expression, lighting, lighting color, contour, cel shading, skin profile, gaze label, camera position, camera rotation, camera field of view, hair color offset, eye color offset, clothing and accessory color offset, ground plane material, left-hand gesture, right-hand gesture, and skybox, among other information. The full dataset contains approximately 600,000 images and is only available for non-commercial research purposes.
VRoid Image Dataset Lite 概述
数据集基本信息
- 名称: VRoid Image Dataset Lite
- 语言: 英语(en)、日语(ja)
- 大小: 1K<n<10K
- 任务类别: 文本到图像(text-to-image)
- 许可证: Creative Open-Rail++-M License
数据集内容
- 模型: 使用VRoid模型,所有模型均为CC0许可。
- 包括VRoid Project、pastelskies、yomox9、くつした、ろーてク等作者的模型。
- 姿势和动作: 使用自定义姿势和Unity Humanoid AnimationClip - PoseCollection的免费版子集,已获得作者直接授权。
- 着色器: MToon(MIT),由开发团队进行了一些修改。
- 其他纹理: 使用Poly Haven和ambientCG提供的CC0纹理。
数据集特点
- 图像生成: 通过随机设置相机角度、姿势、肤色和面部表情等参数生成图像。
- 颜色变换: 应用于皮肤、头发、眼睛、衣物和配饰的独立颜色变换,以增加图像多样性。
元数据描述
- 元数据项: 包括模型名称、姿势编号、相机配置、面部表情、光照、光照颜色、轮廓、阴影、皮肤配置、视线标签等。
- 颜色变换参数: 使用HSV颜色模型进行颜色变换。
数据集可用性
- 完整数据集: 包含约600k图像,仅限于非商业研究目的,需提供1TB在线存储或发送1TB物理硬盘至东京办公室。
开发团队
- 开发: Abstract Engine dev team
- 特别感谢: Mitsua Contributors




