遇见数据集

fjxdaisy/word_split_example_finemath_4plus

收藏
Hugging Face2025-02-12 更新2025-02-15 收录
官方服务:

资源简介:

这是一个包含网页信息的多元语言数据集,包含网页的URL、抓取时间、MIME类型、文本内容、字符数、语言类型和评分等信息。数据集被分为训练集,共有1000个样本。

This is a multilingual dataset containing web page information, including URL, fetch time, MIME type, text content, character count, language type and score, etc. The dataset is split into a training set with a total of 1000 samples.

提供机构:
fjxdaisy
二维码
社区交流群
二维码
科研交流群
商业服务