遇见数据集

mmmaurer/MAGE-enriched

收藏
Hugging Face2025-09-09 更新2025-10-25 收录
官方服务:

资源简介:

这是一个增加了语言学特征的MAGE benchmark版本,用于文本分类任务,语言为英语,数据规模在10万到100万之间。这个数据集使用elfen工具从原始MAGE语料库中提取了额外的语言学特征。

This is an enriched version of the MAGE benchmark with linguistic features for text classification tasks, in English language, with a size range of 100K to 1M. The dataset has been augmented with additional linguistic features extracted from the original MAGE corpus using the elfen tool.

提供机构:
mmmaurer
二维码
社区交流群
二维码
科研交流群
商业服务