遇见数据集

canbingol/vngrs-web-corpus-200k

收藏
Hugging Face2025-10-15 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了文本内容、文本所属语料库和原始ID三个字段,具有200000个训练样本,适用于文本分析和处理任务。

The dataset includes three fields: text content, text corpus, and original ID, with 200,000 training samples, suitable for text analysis and processing tasks.

提供机构:
canbingol
二维码
社区交流群
二维码
科研交流群
商业服务