遇见数据集

timonziegenbein/human-like-edit-sequences-eval

收藏
Hugging Face2025-10-22 更新2025-10-25 收录
官方服务:

资源简介:

这个数据集包含了文本序列(sequence)、标签(label)、原始句子(original_sentence)、不适当部分(inappropriate_part)、重写部分(rewritten_part)、文档ID(doc_id)、句子索引(sent_idx)、编辑索引(edit_idx)和类型(type)等字段。数据集分为测试集(test)和开发集(dev),分别包含76788和80838个示例。数据集的总大小为87714391字节,下载大小为31803782字节。

The dataset includes fields such as text sequence (sequence), label (label), original sentence (original_sentence), inappropriate part (inappropriate_part), rewritten part (rewritten_part), document ID (doc_id), sentence index (sent_idx), edit index (edit_idx), and type (type). It is divided into a test set (test) and a development set (dev), containing 76788 and 80838 examples respectively. The total size of the dataset is 87714391 bytes, and the download size is 31803782 bytes.

提供机构:
timonziegenbein
二维码
社区交流群
二维码
科研交流群
商业服务