SCO dataset : A dataset for semantic communication
收藏资源简介:
The integration of communication and artificial intelligence has become a development trend, one of the applications is semantic communication, but the current research lacks the support of comprehensive datasets. To solve this problem, we built a new image and video dataset, named SCO dataset, for the researches on semantic communication and computing. First, we introduce the peculiarities of the dataset, which contains 5100 images and 138 video clips. Secondly, we we give the data generation and processing methods of the dataset, including images and videos. Then we describes the labeling method of the dataset for intelligent tasks, including classfication and object detection tasks, and the evaluation method of the dataset for subjective experience of human users. Next, we give and analysize the experimental results of the dataset. Finally, we give some practical applications of this dataset, including end-to-end semantic communication, joint source and channels semantic coding, and multi-user semantic communication resource management.
通信与人工智能的融合已成为当前的发展趋势,语义通信即为其典型应用之一,但目前相关研究缺乏全面数据集的支撑。为解决这一痛点,我们构建了一款全新的图像与视频数据集——SCO数据集(SCO dataset),用于语义通信与计算领域的相关研究。首先,本文介绍了该数据集的基本特性:其包含5100张图像与138段视频片段。其次,我们阐述了该数据集的图像及视频数据生成与处理流程。随后,我们详细描述了该数据集面向智能任务的标注方案,涵盖分类与目标检测两类任务,同时给出了面向人类用户主观体验的数据集评估方法。接下来,我们展示并分析了基于该数据集的实验结果。最后,我们介绍了该数据集的若干实际应用场景,包括端到端语义通信、信源与信道联合语义编码,以及多用户语义通信资源管理。




