遇见数据集

Supporting materials for "Widespread use of National Academies consensus reports by the American public"

收藏
Figshare2021-05-18 更新2026-04-28 收录
官方服务:

资源简介:

This repository contains supporting materials for "Widespread use of National Academies consensus reports by the American public." This includes (1) a table providing detailed information on each category into which comments left on NASEM reports were classified; (2) Python scripts for BERT implementation and associated *yml file; (3) BERT output files and model weights for 6-category and 64-category classification of user comments. This includes a compressed file containing the necessary: *.json, *.bin and vocab.txt files to use model weights for scientific replication. The specific file names are: config.json, special_tokens_map.json, tokenizer_config.json, vocab.txt, pytorch_model.bin

本仓库收录了论文《美国公众对美国国家科学院共识报告的广泛使用》的配套研究资料,具体包括: (1) 一张详细列明针对美国国家科学院(NASEM)报告所留评论的分类类别的完整信息表格; (2) 用于实现BERT模型的Python脚本及配套.yml配置文件; (3) 用于对用户评论进行6分类与64分类的BERT模型输出文件及模型权重,其中包含一个压缩文件,内含用于科学复现模型权重所需的*.json、*.bin与vocab.txt格式文件,具体文件名为:config.json、special_tokens_map.json、tokenizer_config.json、vocab.txt、pytorch_model.bin

创建时间:
2021-05-18
二维码
社区交流群
二维码
科研交流群
商业服务