官方服务:
资源简介:
A simple synthetic dataset for use in training classification algorithms
应用场景:
创建时间:
2019-11-29
相关数据集
LOKI
LOKI数据集由中山大学和上海人工智能实验室等机构联合创建,旨在评估大型多模态模型在检测合成数据方面的能力。该数据集包含视频、图像、3D、文本和音频五种模态,共计18,000条问题,覆盖26个详细子类别。数据集的创建过程包括使用多种合成模型生成高质量数据,并通过精细的异常标注进行分级。LOKI数据集主要应用于合成数据检测领域,旨在解决未来互联网中合成数据泛滥带来的真实性鉴别难题。
arXiv2024-10-13 更新3960
Synthetic data generated in Unreal Engine 4
Dataset generated with Unreal Engine 4 and Nvidia NDDS. Contains 1500 images of each object: Forklift, pallet, shipping container, barrel, human, paper box, crate and fence.
DataCite Commons2022-08-12 更新110
Synthetic Meets Authentic: Leveraging Text-to-Image Generated Datasets for Apple Detection in Orchard Environments
Training machine learning (ML) models for computer vision-based object detection process typically requires large, labeled datasets, a process often burdened by significant human effort and high costs
Mendeley Data2024-03-28 更新40
Nemotron-Personas-Singapore
Nemotron-Personas-Singapore 是一个开源(CC BY 4.0)的合成生成人物角色数据集,基于新加坡真实世界的人口统计、地理和人格特质分布,旨在捕捉新加坡人口的多样性和丰富性。该数据集包含 148,000 条记录,涵盖 20 个字段,包括 6 个角色字段(如专业角色、体育角色、艺术角色等)和 14 个上下文字段(如年龄、婚姻状况、教育水平、职业等)。数据规模为 888,00
Hugging Face2026-01-27 更新470
Smart Insurance datasets subset of the AEGIS project
The example dataset was produced within the Smart Insurance demonstrator of the AEGIS project. It contains a sample of the SYNTHETIC data that have been created for the demonstrator purposes.
NIAID Data Ecosystem80



