遇见数据集

AdoCleanCode/clean_dirty_dac_dac_v11_higher_natural_noise_train

收藏
Hugging Face2025-10-14 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含三个字段:噪声类型(noise_type)、序列(sequence)和长度(length)。数据集分为7个批次,每个批次包含约2.5万个例子,总大小约为10GB。数据集适用于处理序列和噪声类型相关的任务。

The dataset consists of three fields: noise type (noise_type), sequence (sequence), and length (length). It is divided into 7 batches, each containing approximately 25,000 examples, with a total size of about 10GB. The dataset is suitable for tasks related to sequence and noise type processing.

提供机构:
AdoCleanCode
二维码
社区交流群
二维码
科研交流群
商业服务