OpenCitations Corpus dataset of all the identifiers, archived on 2017-06-25
收藏资源简介:
This archive contains the dump of the OpenCitation Corpus (OCC, http://opencitations.net) dataset about identifiers, created regularly every month.<br><br>After unzipping the archive, Disk ARchive (DAR, http://dar.linux.free.fr/, a multi-platform archive tool for managing huge amount of data) is needed for recreating the whole structure. For extracting the DAR archive, please run the command<br><br>dar -x [archive-name]<br><br>Where "[archive-name"] is the name of the DAR file without final package number and extension. E.g.:<br><br>dar -x 2016-09-23-corpus_re<br><br>For further questions, comments, and suggestions please don't hesitate to contact Silvio Peroni at essepuntato@opencitations.net.
本归档文件包含每月定期生成的开放引用语料库(OpenCitation Corpus, OCC,http://opencitations.net)有关标识符的数据集转储。 解压该归档文件后,需使用磁盘归档工具(Disk ARchive, DAR,http://dar.linux.free.fr/,一款用于管理海量数据的多平台归档工具)来重建完整的文件结构。若要解压DAR归档,请执行以下命令: dar -x [archive-name] 其中"[archive-name]"为不含末尾包编号与扩展名的DAR文件名。例如: dar -x 2016-09-23-corpus_re 如有进一步疑问、评论或建议,请随时联系Silvio Peroni,邮箱地址为essepuntato@opencitations.net。



