遇见数据集

OCC dataset of all the bibliographic resources, made on 2017-06-25

收藏
Figshare2017-07-27 更新2026-04-08 收录
官方服务:

资源简介:

This archive contains the dump of the OpenCitation Corpus (OCC, http://opencitations.net) dataset about bibliographic resources, created regularly every month.<br><br>After unzipping the archive, Disk ARchive (DAR, http://dar.linux.free.fr/, a multi-platform archive tool for managing huge amount of data) is needed for recreating the whole structure. For extracting the DAR archive, please run the command<br><br>dar -x [archive-name]<br><br>Where "[archive-name"] is the name of the DAR file without final package number and extension. E.g.:<br><br>dar -x 2016-09-23-corpus_re<br><br>For further questions, comments, and suggestions please don't hesitate to contact Silvio Peroni at essepuntato@opencitations.net.

本归档文件包含开放引用语料库(OpenCitation Corpus, OCC,http://opencitations.net)的数据集转储内容,该语料库聚焦书目资源,每月定期生成更新。 解压本归档文件后,需使用磁盘归档工具(Disk ARchive, DAR,http://dar.linux.free.fr/,一款用于处理海量数据的多平台归档工具)重建完整的数据集结构。若要解压DAR归档文件,请执行以下命令: dar -x [archive-name] 其中"[archive-name]"为DAR文件的名称,无需携带末尾的包编号与扩展名。示例如下: dar -x 2016-09-23-corpus_re 如有进一步疑问、评论或建议,请随时联系Silvio Peroni,邮箱地址为essepuntato@opencitations.net。

提供机构:
OpenCitations ​
创建时间:
2017-07-24
二维码
社区交流群
二维码
科研交流群
商业服务