登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
Top-ranked 1000 web sites for each ccTLD, their linguistic characteristics and the related text difficulty.
Top-ranked 1000 web sites for each ccTLD, their linguistic characteristics and the related text difficulty.
收藏
NIAID Data Ecosystem
2026-03-14 收录
网站语言特征分析
文本评估
数据链接:
https://figshare.com/articles/dataset/Top-ranked_1000_web_sites_for_each_ccTLD_their_linguistic_characteristics_and_the_related_text_difficulty_/22078887
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
(XLSX)
应用场景:
创建时间:
2023-02-10
相关数据集
aimlresearch2023/distilabel_12
文本生成
文本评估
该数据集是通过distilabel工具创建的,包含一个`pipeline.yaml`文件,用于重现生成数据集的流程。数据集的结构包括指令、生成内容、生成模型、评分和理由等字段。数据集的配置为默认配置,可以通过Hugging Face的`load_dataset`方法加载。
Hugging Face
2024-04-29 更新
24
0
Baidicoot/harmful-rlhf
强化学习
文本评估
--- dataset_info: features: - name: rejected dtype: string - name: chosen dtype: string - name: prompt dtype: string splits: - name: train num_bytes: 36118145.0 num_exa
Hugging Face
2024-05-06 更新
18
0
RyanYr/reflect_llama8b-t0_mistlarge-t12_om2-300k_correction_150k
自然语言处理
文本评估
该数据集包含四个主要特征:提示(prompt)、响应(response)、响应0的正确性(response@0_correctness)和响应2的正确性(response@2_correctness)。数据集分为一个训练集,包含152,866个样本,总大小为837,079,308.2287472字节。下载大小为306,448,410字节。数据集的文件路径为data/train-*。
Hugging Face
2024-12-15 更新
8
0
ZHLiu627/ultrafeedback_binarized_with_response_full_part2
自然语言处理
文本评估
--- dataset_info: features: - name: prompt dtype: string - name: prompt_id dtype: string - name: chosen list: - name: content dtype: string - name: role dtype:
Hugging Face
2024-03-08 更新
21
0
YYYYYYibo/ultrafeedback_binarized_dataset_offline_part1
文本评估
模型训练
--- dataset_info: features: - name: prompt dtype: string - name: prompt_id dtype: string - name: chosen list: - name: content dtype: string - name: role dtype:
Hugging Face
2024-05-06 更新
12
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广